Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
EugeneOZ
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
EugeneOZ
12d ago
I love Astra pelicans! Awesome results!
2.
▲
by
EugeneOZ
14d ago
Absolutely BRUTAL! :) Thank you for doing this, I love your benchmark the most!
3.
▲
by
EugeneOZ
14d ago
Impressive pelicans!
4.
▲
by
EugeneOZ
15d ago
So the issue was in large context size of these old sessions.
5.
▲
by
EugeneOZ
15d ago
These pelicans are awful.
6.
▲
by
EugeneOZ
22d ago
Also not a fan of country/folk, but today I learned about her special albums trilogy: The Grass Is Blue, Little Sparrow, and Halos & Horns. Quite different music! Some songs I added to my playlist :)
7.
▲
by
EugeneOZ
2mo ago
> You can still see what you’re billed on the Spending page. No, you can not: https://www.pasteboard.co/dNXUdT-h8Giy.png If you want to say that "admin can" - it doesn't matter, I'm not going to ping
8.
▲
Cursor removed cost information from the usage page and CSV export
(forum.cursor.com)
337 points
by
EugeneOZ
2mo ago
|
153 comments
9.
▲
by
EugeneOZ
2mo ago
“Physically accurate” indeed raises the bar significantly. You’ve already done great work here. That said, the feedback seems to come from someone who spent considerable time analyzing your work. Even if only a few of the suggestions are ul
10.
▲
by
EugeneOZ
2mo ago
And now your CLAUDE.md is only good for Claude 5 models, not previous ones.
11.
▲
by
EugeneOZ
2mo ago
> Then: Give Claude rules > Now: Let Claude use judgement No, it should follow my rules exactly. I don't care what code examples it was trained on - it will either write code the way I want, or I'll use another model.
12.
▲
by
EugeneOZ
2mo ago
Why we should waste 46 minutes instead of briefly looking at charts for 20 seconds? To pay their ads? No, thanks.
13.
▲
by
EugeneOZ
2mo ago
Any benchmark where Sol is better than Fable at coding is ridiculous.
14.
▲
by
EugeneOZ
2mo ago
The article is titled "Why Vanilla JS," but the takeaway seems to be "here are the custom abstractions I wrote."
15.
▲
by
EugeneOZ
2mo ago
GPT 5.6 Sol is a token hog. After implementing the task, it started some "reviews" I didn't ask for - they consumed 19.5M and 11.9M tokens, while the task itself was below 5M tokens.
16.
▲
by
EugeneOZ
2mo ago
If for 2% of users a webpage will not look as awesome as intended (it's not guaranteed that it will be broken), that's ok. It's not poisoning - it's a 98% chance of getting a top mark.
17.
▲
by
EugeneOZ
3mo ago
Start fixing the unfixable and doing the undoable things ;)
18.
▲
by
EugeneOZ
4mo ago
2 comments in total there
19.
▲
by
EugeneOZ
4mo ago
Doesn't look like a sport car. From above it actually looks like a phone. The main thing is that the charging port isn’t on the bottom.
20.
▲
by
EugeneOZ
4mo ago
> Ask it if a microservices architecture makes sense for your three-person team and it’ll explain why microservices are an excellent choice If you ask it to be fair and non-biased and provide pros and cons and give possible alternatives
21.
▲
by
EugeneOZ
5mo ago
If you think that you can just silently modify the model without any announcements and only react when it doesn't go through unnoticed, then be 100% sure that your clients will check every possible alternative and will leave you as soo
22.
▲
by
EugeneOZ
5mo ago
> people didn't understand to use /effort to increase intelligence, and often stuck with the default -- we should have anticipated this UI is UI. It is naive to expect that you build some UI but users will "just magically&
23.
▲
by
EugeneOZ
5mo ago
Absolutely awesome, thank you, Lewis!
24.
▲
by
EugeneOZ
5mo ago
> Skills are great for pure knowledge and teaching an LLM how to use an existing tool. But for giving an LLM actual access to services, the Model Context Protocol (MCP) is the far superior That's it. For some things you need MCP, fo
25.
▲
by
EugeneOZ
6mo ago
There are open-source alternatives: https://mochi1ai.com/ https://wan.video/ and others. There are free to use tools also.
26.
▲
by
EugeneOZ
6mo ago
This market will not be abandoned, and other tools already exist: https://klingai.com/global/ https://aistudio.google.com/models/veo-3 https://runwayml.com
27.
▲
by
EugeneOZ
6mo ago
I don't know - it works okay (yet to be tested whether it is actually smarter than Opus 4.6), but it is not bad at all. So far, it works quite fine (I'm not testing the "fast" version).
28.
▲
by
EugeneOZ
6mo ago
Not in my experience. Quoting my tweet: Gave the same prompt to GPT 5.4 (high) and Opus 4.6 (high). GPT 5.4 implemented the feature, refactored the code (was not asked to), removed comments that were not added in that session, made the code
29.
▲
by
EugeneOZ
7mo ago
I do, 100%, every line.
30.
▲
by
EugeneOZ
7mo ago
It depends on how much value their talents can bring to humankind, I guess.
More ›