Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mbowcut2
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
mbowcut2
3mo ago
I made myself reason it out, and came up with the exact same intuition. You need a sequence of 2n moves (n down moves, n over moves), but the sequence is completely determined by which moves are down (the others must be over). So it's
2.
▲
by
mbowcut2
3mo ago
Totally agree. Granting exemptions feels like trying to have their cake and eat it too. If regulations mean anything they need to be enforced so we can see the real downstream effects.
3.
▲
by
mbowcut2
7mo ago
Gotta hit that docker system prune -a
4.
▲
by
mbowcut2
7mo ago
Loved him in Secondhand Lions.
5.
▲
by
mbowcut2
8mo ago
If you thought we were getting bad bugs before, just wait until the 90% agent-coded PRs start landing. We're gonna have multiple crowdstrike-level blowups.
6.
▲
by
mbowcut2
8mo ago
It's an interesting concept, but I'm skeptical about how feasible this is. How much design/legwork/intervention will Seth actually contribute during the entire process? I'm thinking "growing corn" might be
7.
▲
by
mbowcut2
8mo ago
Wow, I didn't know about the "skills" feature, but with that as context isn't this attack strategy obvious? Running an unverified skill in Cowork is akin to running unverified code on your machine. The next super-genius
8.
▲
by
mbowcut2
10mo ago
It makes me wonder about the gaps in evaluating LLMs by benchmarks. There almost certainly is overfitting happening which could degrade other use cases. "In practice" evaluation is what inspired the Chatbot Arena right? But then p
9.
▲
by
mbowcut2
11mo ago
Seems like the less sexy headline is just something about the sample size needed for LLM fact encoding That's honestly a more interesting angle to me: How many instances of data X needs to be in the training data for the LLM to properl
10.
▲
by
mbowcut2
1y ago
I'm not surprised. People really thought the models just kept getting better and better?
11.
▲
Bezier Curve
(javascript.info)
2 points
by
mbowcut2
1y ago
|
0 comments
12.
▲
Magnus Carlsen Commentates Grok vs. OpenAI Finale [video]
(youtube.com)
3 points
by
mbowcut2
1y ago
|
1 comments
13.
▲
by
mbowcut2
1y ago
it looks like the 2nd and 3rd bar never got updated from the dummy data placeholders lol.
14.
▲
by
mbowcut2
1y ago
It's not a new problem (for individuals), though perhaps at an unprecedented scale (so, maybe a new problem for civilization). I'm sure there were black smiths that felt they had lost their meaning when they were replaced by indus
15.
▲
by
mbowcut2
1y ago
I've had similar experiences with vanilla ChatGPT as a DM but I bet with clever prompt engineering and context window management you could solve or at least dramatically improve the experience. For example, you could have the model exe
16.
▲
by
mbowcut2
1y ago
You can, and there has been some interesting work done with it. The technique is called LogitLens, and basically you pass intermediate embeddings through the LMHead to get logits corresponding to tokens. In this paper they use it to investi
17.
▲
by
mbowcut2
1y ago
The problem with embeddings is that they're basically inscrutable to anything but the model itself. It's true that they must encode the semantic meaning of the input sequence, but the learning process compresses it to the point th
18.
▲
by
mbowcut2
1y ago
LLMs are better at LaTeX than humans. ChatGPT often writes LaTeX responses.
19.
▲
by
mbowcut2
1y ago
I think I agree with you. My only rebuttal would be it's this kind of thinking that's kept any leading players form trying other architectures in the first place. As far as I know, SOTA for SSM's just doesn't suggest sig
20.
▲
by
mbowcut2
1y ago
I read this as "pirate space industry" and got real excited.
21.
▲
by
mbowcut2
1y ago
It's interesting how I couldn't tell whether the rocket was 1m tall or 10m tall in this video. Turns out it's actually 6m tall per the link.
22.
▲
by
mbowcut2
1y ago
Nah, I think they made it model agnostic, which is kinda smart.
23.
▲
by
mbowcut2
1y ago
To a topologist, everything is topology.
24.
▲
by
mbowcut2
1y ago
Pack it up boys, they finally made the killer app.
25.
▲
by
mbowcut2
2y ago
I had a smart TV that gradually got slower and slower until it became basically useless. I figured it was just running out of RAM as apps got larger with updates over the years.
26.
▲
by
mbowcut2
2y ago
So, is this just an example of the first-mover disadvantage (or maybe the problem of producing public goods?). The first AI models were orders of magnitude more expensive to create, but now that they're here we can, with techniques lik
27.
▲
by
mbowcut2
2y ago
Me thinks he doth protest too much.
28.
▲
by
mbowcut2
2y ago
I agree that the moats are weak for applications, but I think there are possible strategies to capture users. One way is to make it difficult to switch, similar to Apple Music vs Spotify or iPhone vs Android. Although these platforms offer
29.
▲
by
mbowcut2
2y ago
DeepSeek has demonstrated that there is no technical moat. Model training costs are plummeting, and the margins for APIs will just get slimmer. Plus model capabilities are plateauing. Once model improvement slows down enough, seems to me li
30.
▲
by
mbowcut2
2y ago
> The vast majority of our work is already automated to the point where most non-manual workers are paid for the formulation of problems (with people), social alignment in their solutions, ownership of decision-making / risk, action
More ›