Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
achrono
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
achrono
7d ago
Ignoring the checkbox is an utterly offensive move, but repeatedly manipulating it contrary to stated consumer intent is a whole other level. I didn't know we were supposed to take 'frontier' literally in every sense of t
2.
▲
by
achrono
7d ago
I've been suspecting over the last couple of years of the frontier companies using data for training anyway, regardless of training-use consent. "Using" the data doesn't have to mean they literally upload chat transcript
3.
▲
by
achrono
12d ago
My point is simply this: Chinese labs going to great lengths to distill Claude is evidence that distillation is useful but not that that is what makes them so good. That one doesn't even need them to admit it. Nathan Lambert has made t
4.
▲
by
achrono
12d ago
There was also this follow-up [1], pointing out that most of these 7 studies only had something like a 50–65% estimated chance of producing a significant result, yet all 7 did. In other words, if a whole series of fairly noisy experiments a
5.
▲
by
achrono
12d ago
To whoever downvoted this, it would be helpful if you actually reply with something substantive. The post I'm responding to makes sweeping characterizations, and I challenge it with a relatively good heuristic and indirect evidence, an
6.
▲
by
achrono
12d ago
Yes, evidence is needed but especially for the claim that distillation is what makes these open models good. Serious citation needed. Think about it: even if they distill the shit out of frontier models, the model still gotta learn, right?
7.
▲
by
achrono
1mo ago
Certainly, real-life is the ultimate benchmark. But for various reasons that isn't always immediately possible to go evaluate a model on. Maybe my Greek idea sounded too high falutin' or simply seemingly clever (I give an exampl
8.
▲
by
achrono
1mo ago
Here's an example: ** The stone remembers not what the cities gave, but what the goddess kept. Begin with the first reckoning, under Ariston. Find those who carried the Greeks in their name and the silver in their care. Take their name
9.
▲
by
achrono
1mo ago
Sounds obvious but just try using the models for anything outside the evals. Take something arcane from Greek history, use it to create a masked linguistic puzzle, which you then ask the model to solve mathematically, all wrapped as an ask
10.
▲
LFM2.5-2.6B: On-Device Agents
(docs.liquid.ai)
3 points
by
achrono
1mo ago
|
0 comments
11.
▲
by
achrono
2mo ago
At the risk of sounding like a conspiracy theorist, this sounds like a great opportunity to make a statement . US or China, but likelier to be the former. Maybe Clem's on a call with the US government right now?
12.
▲
by
achrono
2mo ago
Well, sorry but look at the methodology on this. The study the post leans on did not study learning . 36 adults went around copying familiar words, and their "typing" is key presses using one index finger (!). I'm afraid thi
13.
▲
by
achrono
2mo ago
The Greeks themselves, and us in the modern age, really have not given Egypt credit enough for it being the fount from where Greece drew to formulate much of its technology. There is very little direct textual attestation of course, and I&#
14.
▲
by
achrono
3mo ago
Nope, GLM 5.2 is only the latest and greatest in a long line of open-weights models. There are even fully open source models that are comparable to o1-mini (OLMo), or almost-fully-open ones that are comparable to o3 (Nemotron). I'm s
15.
▲
by
achrono
3mo ago
Beats Opus 4.5 on reasoning you say? Prompt: If A goes to B who then goes to C, can A send something to C? Response: We need to interpret best. The phrase "If A goes to B who then goes to C, can A send something to C?" could be a
16.
▲
by
achrono
3mo ago
Take a step back and look at this article's diction and the rest of this entire website. Completely AI generated. All those tokens have to go somewhere
17.
▲
by
achrono
3mo ago
After my own very exhaustive survey, I can just say '+1' and also good to note that OLMo has actually had one independent reproduction (albeit not open) done: https://www.amd.com/en/developer/resources&#x
18.
▲
by
achrono
3mo ago
How do we know that today's frontier models are merely scaled up versions of that? Genuine question, since the labs have narrowed what they share over the years to now almost nothing, in terms of how the model was trained and how it wo
19.
▲
by
achrono
5mo ago
Across 52 professional domains, current frontier models degrade 25% of document content after just 20 interactions!
20.
▲
LLMs Corrupt Your Documents When You Delegate
(arxiv.org)
4 points
by
achrono
5mo ago
|
2 comments
21.
▲
by
achrono
11mo ago
Towards maximizing the sum of individual happiness, power, beauty and knowledge. Maybe a few other attributes in there, but these are the bare minimum that no civilization would deny for itself. The question of course is 'how'. Fo
22.
▲
by
achrono
1y ago
From the article: > According to the 115-page complaint, Baig discovered through > internal security testing that WhatsApp engineers could “move > or steal user data” including contact information, IP addresses > and profile pho
23.
▲
by
achrono
1y ago
I love even more how it's a .md file from well before Markdown even existed.
24.
▲
Square is now sharing its roadmap publicly
(squareup.com)
2 points
by
achrono
1y ago
|
0 comments
25.
▲
by
achrono
1y ago
Really want to know what these "more interesting bits" are that GPT-5-thinking and other models of this calibre cannot do. Unless of course you choose to do them even though these models can in fact do them, in which case, please
26.
▲
by
achrono
1y ago
Other than banks & ticketing, there is a whole host of things that do in fact need an app. * Mobile payments * Navigation * All manner of IoT devices * Wearables! * Digital versions of ID (Mobile Passport Control) etc. So no, you can&#x
27.
▲
by
achrono
1y ago
I think this just further demonstrates the truth behind the truly small & scrappy teams culture at OpenAI that an ex-employee recently shared [1]. Even with the way the presenters talk, you can sort of see that OAI prioritizes speed abo
28.
▲
by
achrono
1y ago
Key highlights in addition to the model quality itself: * real-time router that quickly decides which model to use based on conversation type, complexity, tool needs, and explicit intent (for example, if you say “think hard about this” in
29.
▲
GPT-5 Announcement
(openai.com)
2 points
by
achrono
1y ago
|
1 comments
30.
▲
by
achrono
1y ago
> Toronto, where immigrants from the subcontinent grow up in enclaves surrounded by other immigrants. Citation please, because this is sweeping. Two questions to consider: 1. Are these enclaves representative of the subcontinent, or of a
More ›