Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sigmar
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
sigmar
3d ago
>Just remove hacking (bio-weapon, etc.) data from the training dataset and you're done. Reasoning about how to write secure software uses the same knowledge as reasoning about how to break/hack it.
2.
▲
by
sigmar
4d ago
Lots of private benchmarks already exist, where you have to trust the tester (ex Artificial Analysis, Arc-agi).
3.
▲
by
sigmar
4d ago
Sure, if you remove almost all of my comment my argument disappears. To spell things out for people that don't know about the topic: Zitron is neither an expert on the topic, nor a credible source of information: https://tec
4.
▲
by
sigmar
4d ago
>To understand the truth about the Hugging Face hack, you could do a lot worse than to listen to Ed Zitron and Cal Newport's recent podcast conversation Lol, okay... The crux of this piece is Doctorow saying that the hack was just a
5.
▲
by
sigmar
5d ago
>or there is a more deliberative approach to assigning credit than who was "first" to solve some problem it feels like an unintended consequence of the millennium prize is that people view the [last contributor to the solution]
6.
▲
by
sigmar
6d ago
Were the agents ever tasked with algorithm improvements? Post just says he didn't find any ("report essentially no algorithm advancements"). These LLMs are useful for optimization tasks where they can attempt a change and the
7.
▲
by
sigmar
7d ago
>We have a lot of physics based simulation tools, but they tend to focus on small subsets of the full design problem and they make limiting approximations Do you think it is possible that better math will lead to better physics models?
8.
▲
by
sigmar
12d ago
>The speed with which we were able to produce this proof demonstrates that it is now possible to formalize large swaths of mathematics, which may both catch errors in the common body of mathematical proofs and reduce the burden of refere
9.
▲
by
sigmar
19d ago
https://en.wikipedia.org/wiki/Attempted_assassination_of_Don... what evidence is there that this registered republican was "far left"?
10.
▲
by
sigmar
2mo ago
>We are currently working closely with our inference partners and open-source maintainers to align the technical details and ensure the model can be reliably deployed across the ecosystem. The full model weights will be released by July
11.
▲
by
sigmar
2mo ago
Most of what an LLM does "could have" been done by a human if you throw enough human hours at it. But the reality in this circumstance is that a new tool helped find this leak. Saying this could have happened in a "non LLM wo
12.
▲
by
sigmar
2mo ago
>Evaluators validate each tag individually — for example, protein, preparation, or health, individually rather than judging the item as a whole. Am I reading this right that the jury is multiple LLMs each iterating through each tag and v
13.
▲
by
sigmar
2mo ago
>I think it's reasonable for people to say, hey, if you're going to trash the reputation of Zig (in a pretend-objective way) what specifically is this referring to? Not aware of any comments from Anthropic on this topic.
14.
▲
by
sigmar
2mo ago
To me, this addendum makes it worse. Making small edits to a post like this makes it seem like you're doubling down on the original resentful points, especially with all the new justifications like "a trillion dollar company fired
15.
▲
by
sigmar
2mo ago
>One, this constant bullshit about some window closing, or the perpetual underclass, or falling hopelessly behind. This is negative valence hype, not only is it not true, it’s mostly designed to make you feel bad about yourself and move
16.
▲
by
sigmar
2mo ago
Proven to be safe? Do you want a randomized trial? Do you demand that of every math tutorial video that goes up on YouTube?
17.
▲
by
sigmar
2mo ago
How many years do you use them?
18.
▲
by
sigmar
2mo ago
They don't make any specific claims about what conditions it will diagnose. At 16:30, he says they are only initially doing "body composition" because anything more would add 9+ months to the deployment timeline. I assume the
19.
▲
by
sigmar
3mo ago
that's an obsequious Altman, not their model being banned for being too good
20.
▲
by
sigmar
3mo ago
I visited CERN last July. Was lucky enough to get into a group tour. The tour guide was a postdoc researcher who said the only times that public tours are allowed to take an elevator down is during long shutdowns. So while they do this work
21.
▲
by
sigmar
3mo ago
>publish these incredible papers explaining how they achieved their gains - something the American labs no longer do unfortunately. Google is still releasing a lot of llm architecture research. They introduced speculative decoding of LLM
22.
▲
by
sigmar
3mo ago
The ATF was created by an act of congress. https://en.wikipedia.org/wiki/Gun_Control_Act_of_1968
23.
▲
by
sigmar
3mo ago
>the language in the docs is awfully indirect. writes this^ and then proceeds to highlight a bold title from the docs that says "summarized thinking" that explains things clearly in the first sentence. lol
24.
▲
by
sigmar
3mo ago
Qualified it with "100%" because claude4 models show the first few lines of the chain of thought: >On Claude 4 models, the first few lines of thinking output are more verbose, providing detailed reasoning that's particular
25.
▲
by
sigmar
3mo ago
Fable/mythos are the first models from anthropic that hide 100% of reasoning tokens. So it seems to me like we're about to get a lot more data about to what extent Chinese model progress has been a consequence of distillation tech
26.
▲
by
sigmar
3mo ago
>Some administration officials have said that a resolution should include an acknowledgment on Anthropic’s part that its rollout of Fable and communication with the White House could have been improved, people familiar with the talks sai
27.
▲
by
sigmar
3mo ago
I should have contextualized the quote- "chat is dead" is from an openai employee which was describing how they're shifting focus to more agentic consumer products, and putting less focus on the back-and-forth chatbot interfa
28.
▲
by
sigmar
3mo ago
I like that "chat is dead" framing I heard recently because too many people are having interpersonal relations with these LLMs and want to tune their "emotions"/tone. Humanity would be in a better place if we though
29.
▲
by
sigmar
3mo ago
All science-based conclusions come with uncertainty. Only ideologues (and siths) write in absolute terms.
30.
▲
by
sigmar
3mo ago
Wacky decision. I know they want ads but it's an endurance sport! Imagine if they chopped a sport like 10km run into quarters.
More ›