Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
tshadley
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
1.
▲
by
tshadley
8mo ago
https://en.wikipedia.org/wiki/Liquid_droplet_radiator
2.
▲
by
tshadley
8mo ago
> LLMs cannot offer that promise by design, so it remains your job to find and fix any deviations from the abstraction you intended. LLMs are clumsy interns now, very leaky. But we know human experts can be leak-proof. Why can't L
3.
▲
by
tshadley
1y ago
As IMO medalists they would be expected to I'm sure. But this can be verified because the results are public: https://github.com/aw31/openai-imo-2025-proofs/
4.
▲
by
tshadley
1y ago
Yes, OpenAI: https://x.com/alexwei_/status/1946477754372985146 > 6/N In our evaluation, the model solved 5 of the 6 problems on the 2025 IMO. For each problem, three former IMO medalists independently grad
5.
▲
by
tshadley
1y ago
Thank you, amazing, fresh.
6.
▲
by
tshadley
1y ago
The goal here is not to replace transformers but combine them with RNN so you get both good short-term memory (self-attention) and much improved long-term memory (ATLAS recurrent memory). "Empirically, our models—OmegaNet, Atlas, DeepT
7.
▲
by
tshadley
1y ago
100% agreed with your experience, AI provides little value to one's area of expertise (10+ years or more). It's the context length -- AI needs comparable training or inference-time cycles. But just wait for the next doubling of l
8.
▲
by
tshadley
2y ago
> To understand the capabilities of LLMs, we evaluate GPT3 (text-davinci-003) [11], ChatGPT (GPT-3.5-turbo) [57] and GPT4 (gpt-4) Oh dear, this is embarrassing. Anil Anathaswamy, are you aware a year in AI research now is like 10 years
9.
▲
by
tshadley
2y ago
Well the Franks study probably destroyed any chance for natural sleep conditions. Nedergaard is scathing: https://www.thetransmitter.org/glymphatic-system/new-method-... > The new paper used many of the techniques
10.
▲
by
tshadley
2y ago
Seems to me o3 prices would be what the consumer pays, not what OpenAI pays. That would mean o3 could be more efficient in-house than paying subject-matter experts.
11.
▲
by
tshadley
2y ago
I always get the feeling he's subconsciously inserting a "magical" step here with reference to "synthesis"-- invoking a kind of subtle dualism where human intelligence is just different and mysteriously better than
12.
▲
by
tshadley
2y ago
> One random example to illustrate the distinction: training gaps can easily decrease uncertainty. You have lots of mammals in your training data, and none of them lay eggs. You ask "The duck-billed platypus is my favorite mammal! D
13.
▲
by
tshadley
2y ago
> ...proving that this one particular piece of the hallucination problem may be conceptually simple. Everything mentioned in the article boils down to that one particular piece-- non-detected uncertainty. The architecture constraints re
14.
▲
by
tshadley
2y ago
The article referenced the Oxford semantic entropy study but failed to clarify that the issue greatly simplifies LLM hallucination (making most of the article outdated). When we are not sure of an answer we have two choices: say the first t
15.
▲
by
tshadley
2y ago
"Why PCIe Risers suck and the importance of using SAS Device Adapters, Redrivers, and Retimers for error-free PCIe connections." I'm a believer! Can't wait to hear more about this.
16.
▲
by
tshadley
2y ago
Sure looks like a typo. Contact author? https://x.com/fchollet https://x.com/arcprize https://x.com/mikeknoop
17.
▲
by
tshadley
3y ago
https://mathshistory.st-andrews.ac.uk/HistTopics/Bakhshali_m... has some examples. |One person possesses seven asava horses, another nine haya horses, and another ten camels. Each gives two animals, one to each of the
18.
▲
by
tshadley
3y ago
So this is old news?
19.
▲
by
tshadley
3y ago
All cynicism aside, there's vastly more in the collective writings of humans on empathy than medicine.
20.
▲
by
tshadley
3y ago
From the article: "April 3, 2023 - Real Humans Can’t Tell the Difference Between a 13B Open Model and ChatGPT Berkeley launches Koala, a dialogue model trained entirely using freely available data. They take the crucial step of measuri
21.
▲
by
tshadley
3y ago
Ah, that's it; polite fictions are scored higher than uncomfortable facts.
22.
▲
by
tshadley
3y ago
That's weird. Having the community study this would certainly help them. They're afraid this is giving too much insight into their proprietary training/modeling methods?
23.
▲
by
tshadley
3y ago
That should be okay though, 10 good answers will still report the score of the best one chosen. I think the GPTs are using beam search which is projecting out a "beam" (looks more like a tree to me) of probable answers each of w
24.
▲
by
tshadley
3y ago
> A probable guess will lower loss much better than "I don't know" or whatever equivalent. Guessing only reduces loss as much as the dataset allows -- a bad guess will give a higher loss. The model learns to assign probab
25.
▲
by
tshadley
3y ago
Earth's crust: not quite the same as Cu/Zn but way more than I expected: https://periodictable.com/Properties/A/CrustAbundance.an.htm... Lithium: 0.0017% Copper: 0.0068% Zinc: 0.0078% Se
26.
▲
by
tshadley
4y ago
I'm assuming that you disagree specifically that few artists can essentially create derivative art (i.e avoiding plagiarization but being clearly influenced by artist X or paying homage to artist Y, etc.) as well as top SD models. Wel
27.
▲
by
tshadley
4y ago
My initial definition was "like a super-humanly talented artist": this is very different from a human being who also happens to be an artist. Stable Diffusion does only art with text-prompting well, nothing else, and will take a
28.
▲
by
tshadley
4y ago
> An NN is simply an approximation of a multi-valued function, whose parameters are adjusted by minimizing the difference between the output of the NN and the output of the real function for a certain input. Right, but that equally fits
29.
▲
by
tshadley
4y ago
"[The complaint] argues that the Stable Diffusion model is basically just a giant archive of compressed images (similar to MP3 compression, for example) and that when Stable Diffusion is given a text prompt, it “interpolates” or combin
30.
▲
by
tshadley
4y ago
"I cannot emphasize this enough: ChatGPT is not generating meaning. It is arranging word patterns." To reconcile this statement with his admission that ChatGPT reliably turns out passable results, we must assume the average studen
More ›