Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
docheinestages
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
docheinestages
6d ago
Nice try, Dario.
2.
▲
by
docheinestages
6d ago
Same here.
3.
▲
by
docheinestages
9d ago
A dead giveaway that I think it's AI-made (not necessarily a con) is the complexity of the UI and redundant information. Things like the green dot paired with "interface is up". It looks polished though.
4.
▲
by
docheinestages
13d ago
Now I'm starting to doubt the credibility of Artificial Analysis.
5.
▲
by
docheinestages
13d ago
My gut feeling tells me it has something to do with Cloudflare. Along with AWS, they're two of the main suspects in such incidents.
6.
▲
by
docheinestages
13d ago
All I see is slop.
7.
▲
by
docheinestages
14d ago
Is this a compressed and retrained version of GLM-5.2?
8.
▲
by
docheinestages
15d ago
I mean, I'm interested to know if the frontier models also get to see all questions at once. Then it's more fair game than if they just see one question at a time.
9.
▲
by
docheinestages
15d ago
It's what happens when you don't do market research.
10.
▲
by
docheinestages
15d ago
Isn't this cheating? Or rather, are frontier agents only looking at one question at a time? If I understand correctly, you're looking at all the examples of the exam questions. If the exam was adjusted so that you can only look at
11.
▲
by
docheinestages
16d ago
I don't think writing is an "AI-complete" problem. It's just that the models are inefficient at it right now, due to training and/or architecture. Good writing requires thinking, reflection, and doing multiple passe
12.
▲
by
docheinestages
16d ago
Memory is not necessarily a good thing. What we need is a combination of sufficiently large context windows (10-100 million tokens) along with curate, bloat-free data.
13.
▲
by
docheinestages
16d ago
> How can I judge what is a good memory to store? How can I avoid filling my memory with crap? > This is a common fear with memory systems but doesn't really apply to memoryfields. Irrelevant material is simply never surfaced by
14.
▲
by
docheinestages
19d ago
Looks like the repo teleported.
15.
▲
by
docheinestages
21d ago
This is only a problem because people keep using X. Just stop.
16.
▲
by
docheinestages
22d ago
But did it produce anything meaningful? If so, show me.
17.
▲
by
docheinestages
25d ago
Oh fantastic! It was already subpar and they want to make it even worse. One day we'll look back at history and see how Anthropic went down.
18.
▲
by
docheinestages
26d ago
Kudos to Kagi!
19.
▲
by
docheinestages
1mo ago
Would it work with "giggity" too?
20.
▲
by
docheinestages
1mo ago
Can we get a tool that does the opposite for the readers?
21.
▲
by
docheinestages
1mo ago
I think one of the best use cases of AI is as a natural language interface to programming. The syntax of a programming language is an opinionated part of its design that often adds to the complexity of learning the language.
22.
▲
by
docheinestages
1mo ago
I watched the intro video and all I saw were numbers in an Excel sheet. That doesn't make this fun. I'd rather have AI ELI5 it to me.
23.
▲
by
docheinestages
1mo ago
Anthropic should build a harness (and model) that smartly takes care of all these points. Not requiring the user to do the manual work. All I see are excuses because they cannot handle the load and enforce strict quotas on users, all while
24.
▲
by
docheinestages
1mo ago
Are the benchmarks comparing just the models while keeping the harness the same (the open-source Toast harness)?
25.
▲
by
docheinestages
1mo ago
My uses cases are mainly research and prototyping, often starting from scratch in greenfield projects. The 5-hour quota is shared between Claude web and Claude Code, so it doesn't really matter what interface I use. When it comes to r
26.
▲
by
docheinestages
1mo ago
Claude has essentially become useless for agentic development or research. Doesn't matter what model you use. A few rounds and bam, you've burned through your quota. Doesn't matter how "intelligent" their models are
27.
▲
by
docheinestages
1mo ago
YCombinator's investments are unfortunately very questionable nowadays.
28.
▲
by
docheinestages
1mo ago
Your landing page is very hard to read. The font size is literally 10px for some content, while animations distract the reader. Dear YC, please force the startups to dedicate some of the 500k funding for standard web design.
29.
▲
by
docheinestages
1mo ago
They conveniently decided not to include the Qwen range of models in the Artificial Analysis graph, except the out-of-league Max variant. At least be brave and honest.
30.
▲
by
docheinestages
1mo ago
I'm interested to learn from the prompts you used to generate the images and texts. Perhaps a skill you could share to generate such content?
More ›