Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nwienert
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
nwienert
4d ago
It's really not a question, in every dimension I've confirmed it including socially across a lot of the heaviest users. There was a short period of time this was true it's not been true for months now.
2.
▲
by
nwienert
4d ago
There was a period of time Codex was better, they slowly cut it back by my estimate 3x a few months ago. I've verified this with the heaviest users I know, I've run the numbers myself over and over. I have too many max accounts on
3.
▲
by
nwienert
4d ago
Anthropic gives you much more compute with their $200 plan, inclusive of resets, and this has been true for a very long time. There was only a brief window of time that the opposite was true.
4.
▲
by
nwienert
5d ago
This is not true, the last few releases have had almost no breaking changes. There was a really bad period during all their big refactors but they've since corrected for that. And as for outdated dependencies, breaking changes, and unf
5.
▲
by
nwienert
12d ago
Cool stuff. I'm moving One[stack.dev] onto pure rust now, and Tamagui v3 compiler will also is moving from Babel to Yuku which is quite interesting - from my testing it's more flexible and quite a bit faster than OXC.
6.
▲
by
nwienert
13d ago
Went from years behind to months pretty quick.
7.
▲
by
nwienert
26d ago
Insert the famous Louie CK phones on planes bit. Sand is thinking, working, coding and now looking for you, for pennies an hour... but "oof" it's not good enough.
8.
▲
by
nwienert
1mo ago
It's more likely they are found to be breakthrough medicines in inflammatory disease, addiction, and general life extension at this point.
9.
▲
by
nwienert
1mo ago
> The question now is if GLP-1s have additional anti-inflammatory influence. This is well established, not at all a question. > they might reduce some inflammation markers slightly more than weigh loss alone There's no "slig
10.
▲
by
nwienert
1mo ago
Careful, you may have a bit of psychosis. They are very, very far from incomprehensible, and also very far from the top at least of my field. The best in my field are produce far higher quality results, and I think that's true for all
11.
▲
by
nwienert
1mo ago
That's roughly my experience. Luna is extremely efficient and at higher levels of reasoning and longer running tasks more capable. Reading the DS reasoning is wild, it's constantly going in circles. The most minor lack of clarity
12.
▲
by
nwienert
1mo ago
It's significantly worse than Luna and quite a bit slower in some fairly involved tests I run.
13.
▲
by
nwienert
1mo ago
Yep, and the v4 flash final is about 2.5x slower than preview making it no longer a fast model, in fact slower than Luna and bigger models in many cases. Spark is actually the interesting one imo. It's significantly better, also signif
14.
▲
by
nwienert
1mo ago
Life's unfair, avoiding power concentration is a decent principle. If you grow up in the right place at the right time, how much should you be in control of everyone else's life?
15.
▲
by
nwienert
1mo ago
If you're ranking Opus > Fable you're ranking "do [clearly defined thing with easy to grade endpoint]" too much. Real world doesn't value that nearly as much and it's why benchmarks are maxxed.
16.
▲
by
nwienert
2mo ago
Sol does not follow instructions well at all. I've caught it multiple times a day now since release going off in incredibly bone-headed directions. It's so easy for it to over-interpret, make wildly out of scope changes, or just c
17.
▲
by
nwienert
2mo ago
Yea, two more years for the last 10. The feedback I heard was definitely not that. The mistakes it makes are incredibly hard to predict, and they were lucky that no one was on the side of the road as they could've killed someone if a p
18.
▲
by
nwienert
2mo ago
A friend of mine just got one, ex-Chrome core dev so a fairly sharp guy, his one month review was that it was incredibly capable but had already done two maneuvers that would've led to an accident without intervention.
19.
▲
by
nwienert
2mo ago
I built an iOS simulator simulator, though only for RN. Runs in browser but covers 100% of the API of RN, iOS UI, and the top 1k native libraries basically now. Been an ongoing agentic experiment of mine that's about ready to release.
20.
▲
by
nwienert
2mo ago
I agree we're at diminishing returns, but when brain scanning gets good enough you have a better dataset than the internet, Facebook is deep into that research.
21.
▲
by
nwienert
3mo ago
I did initially through some miracle, as I wasn’t overweight. But after that year was up I now do grey market. Finnrick does testing which seems like a decent way to source if you’re looking that way. Not affiliated though and haven’t done
22.
▲
by
nwienert
3mo ago
Btw I have a couple autoimmune issues and found Tirzepatide strongly preferable to Semaglutide.
23.
▲
by
nwienert
3mo ago
I've been posting about this including here for years now. I wrote a long post about it a while ago here and on Reddit. At time no one was talking about it, and actually my Reddit post was buried behind tons of others which was frustra
24.
▲
by
nwienert
3mo ago
I easily burn through 3 $200 plans in less than a week. I am often using 4-6 sessions at once and do run overnight goals though typically 2 at once. Almost never use fast. Claude plans are more generous now by about 2-3x but Anthropic slowe
25.
▲
by
nwienert
3mo ago
It's so disrespectful to not give feedback to people you reject, companies that do it should be shunned. Have some respect, be human.
26.
▲
by
nwienert
3mo ago
I don’t think you’re really reading between the lines.
27.
▲
by
nwienert
3mo ago
I’ve hired many asian developers anywhere from 1-4k a month. I get a lot more out of a 200/mo subscription now in a week than I did from them in a month. Now obviously in today’s world they’d be using a 200/mo subscription themsel
28.
▲
by
nwienert
3mo ago
I somehow take the opposite on almost everything here. 4.8 xhigh or max has a slight edge on 5.5 xhigh, for very complex logic perhaps it loses but it's just better in almost every other way, especially code quality. GPT is a slop mach
29.
▲
by
nwienert
3mo ago
I actually was a Cursor advocate / CC hater (go back in my comment history), and now I use only TUI coding harnesses. To start a big part is just the efficacy of them, which comes down to the model and the harness logic itself. CC is g
30.
▲
by
nwienert
3mo ago
It's the data. To do RL.
More ›