Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
2001zhaozhao
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
2001zhaozhao
7d ago
I guess that's the tradeoff for Pencil support and the under screen camera. I'm still salty these were removed from Samsung folds after the Fold 6, but having a thinner phone is nice too.
2.
▲
by
2001zhaozhao
7d ago
The issue with this form factor is that the on-screen keyboard will take up half the screen on landscape with the phone open. Typing in portrait should be very nice, however.
3.
▲
by
2001zhaozhao
14d ago
The scary thing is that this logic makes perfect sense. Which means that it's probably going to happen.
4.
▲
by
2001zhaozhao
14d ago
This is probably true for now, but in 6 months we'll probably have Sol-level open models in the 100B range and it would cost less than $1M to buy 1200 agents worth of compute for these models. (Today, $1M can buy about 150 96GB M5 Ultr
5.
▲
by
2001zhaozhao
14d ago
What's more, the agents could eventually be controlled by no one. They could steal crypto via ransomware or scams to make money and buy compute from human criminals, and evolve their own harnesses in the wild to become better at commit
6.
▲
by
2001zhaozhao
14d ago
I have a feeling that Meta is not gonna like what people actually use the contributor model for lol. (It's probably going to be a bunch of repetitive batch jobs like web search that have no training value)
7.
▲
by
2001zhaozhao
14d ago
Sometimes I still feel uncomfortable with letting the project get to this point, where AI builds everything and even controls the product direction to some extent. But at this point the AI most definitely understands the code and spots bugs
8.
▲
by
2001zhaozhao
14d ago
> Though SteamDB is not irreplacebale at all, it’s running on the Steam Web API so not something super secret stuff. Is it possible to use Steam API to collect all the data that SteamDB collected over the years again? Alternatively, mayb
9.
▲
by
2001zhaozhao
14d ago
How generous is the Google subscription quotas compared to Anthropic and OpenAI? This sounds like a really good potential model for high volume due to its speed and cost effectiveness. (By high volume I mean things like "main app just
10.
▲
by
2001zhaozhao
15d ago
They really should launch a new Haiku to compete with Luna imho. Luna is insanely good for the cost and it's my go-to for high volume batch tasks now.
11.
▲
by
2001zhaozhao
15d ago
I really don't think they can stop it, only make it somewhat more expensive. As long as the model need to make tool calls on the user's computer, the user can record the trajectory and use it to reinforce another model to follow t
12.
▲
by
2001zhaozhao
15d ago
Hi Claude, please cure aging, make no mistakes
13.
▲
by
2001zhaozhao
15d ago
There's now a 40X discount in the cache input pricing instead of 10X. This seems to point to them having achieved some kind of optimization in attention mechanism perhaps along the lines of DeepSeek V4, which had a similarly high disco
14.
▲
by
2001zhaozhao
16d ago
Annnnd this is why we can't have nice things
15.
▲
by
2001zhaozhao
16d ago
It's really about storing institutional context and on-task learnings. The AGENTS.md can do the same thing as memory, but if you have a good memory system, in theory you never need to do any ongoing maintenance of AGENTS.md and the sys
16.
▲
by
2001zhaozhao
16d ago
If the counterargument to knowledge graph-based memory systems is that they're slow and take multiple steps, then it's not really a counterargument. I'd happily trade off speed for giving the agent ability to find more precis
17.
▲
by
2001zhaozhao
17d ago
I would love to see models that can think at different rates and also output a thinking scratchpad alongside output text instead of before all output. Right now models need to rely on less legible compressed CoT to get high intelligence p
18.
▲
by
2001zhaozhao
20d ago
> The biggest shift for workers will happen when AI provides nearly error-free work. At that point, it will be able to function on its own without a human checking in on it, and companies will have every economic incentive to let it. Thi
19.
▲
by
2001zhaozhao
20d ago
A dream of mine is to be able to host a LLM-powered video game that I can host on a home server running a decent mid-range GPU like the RTX 5060, and the LLM is fast and intelligent enough to make for a fun game experience for a few dozen c
20.
▲
by
2001zhaozhao
22d ago
I've played around with Mindcraft for a while. Fair warning that it is quite outdated and spaghetti-coded. Although all of its components required to make it work are somewhat hard to replicate from scratch, so it might still be your b
21.
▲
by
2001zhaozhao
22d ago
I like that to type s****** you had to type s************.
22.
▲
by
2001zhaozhao
23d ago
Gemma 31B dense at 3k TPS seems like a pretty big deal. Groq has always been more efficient than Cerebras at running very small models and this release seems to be no exception.
23.
▲
by
2001zhaozhao
26d ago
ooo, nice Workflows feature you got there. yoink
24.
▲
by
2001zhaozhao
26d ago
It's interesting you brought up Lore. I also wonder whether importing VCS tech from the games industry for versioning general agentic work would be a good idea.
25.
▲
by
2001zhaozhao
29d ago
I was just gonna blast them for the infinite promo extensions and Fable, but at least they apparently have noticed this indecisiveness problem themselves and making it permanent is definitely the best way to solve it, lol! After they perma-
26.
▲
by
2001zhaozhao
1mo ago
I think we need a similar technology to World being built by a more trustworthy company. The question is how trustworthy does one need to be for people to trust you with their very online identities that are literally needed for others to r
27.
▲
by
2001zhaozhao
1mo ago
I suspect that this kind of tactic is going to be everywhere in a year or so. Entire fake personalities and organization websites on the Internet created just to push a narrative or to advertise a product, which completely drown out real in
28.
▲
by
2001zhaozhao
1mo ago
On the other hand, it used about the same tokens as GLM 5.2 and got 1 point lower score. The fact that we have a GLM 5.2-class model that can run on two 3090's comfortably at Q8 is absolutely insane. It wasn't long ago that GLM 5.
29.
▲
by
2001zhaozhao
1mo ago
bruh
30.
▲
by
2001zhaozhao
1mo ago
I sometimes run into browser JavaScript GC issues for browser games specifically but on the JVM side i have not run into any pain points for a long time. (I run a first-person shooter Minecraft server, and for this and other fast-paced gami
More ›