Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
a_wild_dandan
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
a_wild_dandan
15d ago
Guessing SH meant Steven Hawking, who kicked the bucket. Metaphorically.
2.
▲
by
a_wild_dandan
21d ago
Maybe everyone will migrate to Hugging Bay for downloading models via torrent?
3.
▲
by
a_wild_dandan
2mo ago
Wow, you weren't kidding. I looked at their chart, and the cost-per-task for Fable is more than double Sol's. And DeepSeek absolutely stomps. Four cents per-task vs Sol's $1 and Fable's $3. I might need to check out
4.
▲
by
a_wild_dandan
2mo ago
solve p=np make no mistakes
5.
▲
by
a_wild_dandan
3mo ago
Backward compatibility with current meatspace tooling.
6.
▲
by
a_wild_dandan
7mo ago
You’re not stupid. That’s terrible UX. The button is completely disconnected from its modal, and is placed in a bizarre/nonstandard location.
7.
▲
by
a_wild_dandan
8mo ago
Speaking of tricks, does anyone here know how many angels can dance on the head of a pin?
8.
▲
by
a_wild_dandan
9mo ago
Taiwan’s geopolitical position is vastly more complex than the fantasy where invasion would follow merely from fab parity.
9.
▲
by
a_wild_dandan
9mo ago
> Unlike the previous GPT-5.1 model, GPT-5.2 has new features for managing what the model "knows" and "remembers to improve accuracy. Dumb nit, but why not put your own press release through your model to prevent basic thi
10.
▲
by
a_wild_dandan
9mo ago
Businesses do whatever’s cheap. AI labs will continue making their models smarter, more persuasive. Maybe the SWE profession will thrive/transform/get massacred. We don’t know.
11.
▲
by
a_wild_dandan
9mo ago
No. I like being able to ignore them. I can’t do that if people chop off their disclaimers to avoid comment removal.
12.
▲
by
a_wild_dandan
9mo ago
Someone will make a killing on a rechargeable version of this. The ergonomics are a good idea.
13.
▲
by
a_wild_dandan
9mo ago
If the claims in the abstract are true, then this is legitimately revolutionary. I don’t believe it. There are probably some major constraints/caveats that keep these results from generalizing. I’ll read through the paper carefully thi
14.
▲
by
a_wild_dandan
11mo ago
His specific thesis is that pods fundamentally clean worse than powder because they're inherently single-stage releases of detergent in machines designed for two-stage releases. Despite this, he still explicitly says that pods have t
15.
▲
by
a_wild_dandan
11mo ago
How does having management strategies over an alleged addiction imply that it isn’t an addiction?
16.
▲
by
a_wild_dandan
11mo ago
Intelligence is whatever an LLM can’t do yet. Fluid intelligence is the capacity to quickly move goal posts.
17.
▲
by
a_wild_dandan
1y ago
I would bet that it's far lower now. Inference is expensive we've made extraordinary efficiency gains through techniques like distillation. That said, GPT-5 is a reasoning model, and those are notorious for high token burn. So who
18.
▲
by
a_wild_dandan
1y ago
This might be a dumb question but like...why does it matter? Are other companies reporting training run costs including amortized equipment/labor/research/etc expenditures? If so, then I get it. DeepSeek is inviting an appl
19.
▲
by
a_wild_dandan
1y ago
Thank you for explaining. I was so confused at how AMD was improving Quake performance with duck-like monikers.
20.
▲
by
a_wild_dandan
1y ago
You're right! China is presently terrified of involution. They're dealing with wage deflation and immense debt right now. Beijing is telling its companies to scale back subsidies and stop price wars. The flood of cheap batteries
21.
▲
by
a_wild_dandan
1y ago
Is it possible to run Cursor entirely with local models? My Mac can comfortably run relatively massive models. I would experiment so much more with AI in my codebases knowing that I won't slam into a brick wall due to quotas, connectio
22.
▲
by
a_wild_dandan
1y ago
That's absolutely wild. I've been loving using the 96GB of (V)RAM in my MacBook + Apple's mlx framework to run quantized AI reasoning models like glm-4.5-air. Running models with hundreds of billions of parameters (at ~14 tok
23.
▲
by
a_wild_dandan
1y ago
> What am I doing wrong? Providing a woefully inadequate descriptions to others (Claude & us) and still expecting useful responses?
24.
▲
by
a_wild_dandan
1y ago
GLM-4.5-air produces tokens far faster than I can read on my MacBook. That's plenty fast enough for me, but YMMV.
25.
▲
by
a_wild_dandan
1y ago
Oh absolutely, AI labs certainly talk their books, including any safety angles. The controversy/outrage extended far beyond those incentivized companies too. Many people had good faith worries about Llama. Open-weight models are now v
26.
▲
by
a_wild_dandan
1y ago
I'll accept Meta's frontier AI demise if they're in their current position a year from now. People killed Google prematurely too (remember Bard?), because we severely underestimate the catch-up power bought with ungodly piles
27.
▲
by
a_wild_dandan
1y ago
Right? I still remember the safety outrage of releasing Llama. Now? My 96 GB of (V)RAM MacBook will be running a 120B parameter frontier lab model. So excited to get my hands on the MLX quants and see how it feels compared to GLM-4.5-air.
28.
▲
by
a_wild_dandan
1y ago
"Perfection is achieved, not when there is nothing more to add, but when there is nothing left to take away." - Claude, probably
29.
▲
by
a_wild_dandan
1y ago
"You haven't contorted your comically simple query enough to make the brittle tool work. Throw the chicken bones better next time."
30.
▲
by
a_wild_dandan
1y ago
Huh? Grammar-based sampling has been commonplace for years. It's a basic feature with guaranteed adherence. There is no "carefully crafting" anything, including safeguards.
More ›