Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
zhyder
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
zhyder
6d ago
AI dramatically reduces the cost of writing code (especially when you have a reference), so the scale between native and cross-platform is going to tilt more towards native now compared to before. But I'm curious what their update will
2.
▲
by
zhyder
2mo ago
A big part of this announcement does seem to be _delegation_ in the background; they give the example of web search but that could be any tool. I haven't tried it yet either but sounds like they've found a reasonable UX that mixes
3.
▲
by
zhyder
4mo ago
Aside from the issue of platform owners (Apple, Google, Microsoft) offering storage sync as an integrated feature, which others have pointed out, the other reason growth is limited is that filesystem storage and sync thereof has become less
4.
▲
Google's AI Studio now integrates with Firebase for vibe coding production apps
(blog.google)
2 points
by
zhyder
6mo ago
|
3 comments
5.
▲
by
zhyder
7mo ago
Looks like the best display you can get in laptops at this price: 2408x1506 resolution, 500 nits, antireflective coating (!). And bonus points for no silly notch.
6.
▲
by
zhyder
7mo ago
I guess it could warn about it but the VM sandbox is the best part of Cowork. The sandbox itself is necessary to balance the power you get with generating code (that's hidden-to-user) with the security you need for non-technical users.
7.
▲
Netflix drops bid for Warner Bros after Paramount offer
(theverge.com)
20 points
by
zhyder
7mo ago
|
1 comments
8.
▲
by
zhyder
7mo ago
Model card: https://deepmind.google/models/model-cards/gemini-3-1-flash-... Pretty close to Gemini 3 Pro Image (aka Nano Banana Pro) in most benchmarks, even without thinking+search, and even exceeding it in 2 mos
9.
▲
More plugin support in Claude Cowork
(claude.com)
1 points
by
zhyder
7mo ago
|
0 comments
10.
▲
by
zhyder
7mo ago
Agree, can't wait for updates to the diffusion model. Could be useful for planning too, given its tendency to think big picture first. Even if it's just an additional subagent to double-check with an "off the top off your hea
11.
▲
by
zhyder
7mo ago
Surprisingly big jump in ARC-AGI-2 from 31% to 77%, guess there's some RLHF focused on the benchmark given it was previously far behind the competition and is now ahead. Apart from that, the usual predictable gains in coding. Still is
12.
▲
by
zhyder
7mo ago
"the value of a human eyeball" / attention is and always will be the limited resource. But I wish the way the economy worked wasn't that attention is sold for money, which makes money the moat, and sets a floor on how lo
13.
▲
by
zhyder
7mo ago
Hmm the whole point of checkpoints seems to be to reduce token waste by saving repeat thinking work. But wouldn't trying to pull N checkpoints into context of the N+1 task be MUCH more expensive? It's at odds with the current prac
14.
▲
by
zhyder
7mo ago
So 2.5x the speed at 6x the price [1]. Quite a premium for speed. Especially when Gemini 3 Pro is 1.8x the tokens/sec speed (of regular-speed Opus 4.6) at 0.45x the price [2]. Though it's worse at coding, and Gemini CLI doesn'
15.
▲
by
zhyder
8mo ago
Love it. Wonder if it's viable for citizen journalism in warzones and areas of civil unrest, with the larger size of photos (and short videos), given the inherently slow transfer rates and battery life implications of going thru multip
16.
▲
by
zhyder
8mo ago
Sounds like antirez, simonw, et al are still advocating reviewing the code output of these agents for now. But presumably soon (within months?) the agents will be good enough such that line-by-line review will no longer be necessary, or hum
17.
▲
Ask HN: How do you review the code from agents?
2 points
by
zhyder
8mo ago
|
1 comments
18.
▲
by
zhyder
8mo ago
Most car manufacturers made this mistake because they started mimicking the then leader for innovation (and customer satisfaction), Tesla, too much. General cautionary tale: just coz a company is successful, doesn't mean it's doin
19.
▲
by
zhyder
9mo ago
"Almost anyone can prompt an LLM to generate a thousand-line patch and submit it for code review. That’s no longer valuable. What’s valuable is contributing code that is proven to work." I'd go further: what's valuable i
20.
▲
by
zhyder
9mo ago
Have you tried them with providing a grounding resource, e.g. attaching a file to ChatGPT or NotebookLM? Yes need some human expert to create (or curate) that grounding resource in the first place, but LLMs handle the rest well: presenting
21.
▲
by
zhyder
9mo ago
End of an era: video (with broadband Internet penetration) was the best tool we had for 15+ years. But LLMs are now good enough, including in image+infographic generation and factuality (especially when grounding resources are provided... w
22.
▲
by
zhyder
9mo ago
Glad to see big improvement in the SimpleQA Verified benchmark (28->69%), which is meant to measure factuality (built-in, i.e. without adding grounding resources). That's one benchmark where all models seemed to have low scores unti
23.
▲
CC: Google Labs AI agent for email+calendar
(blog.google)
2 points
by
zhyder
9mo ago
|
0 comments
24.
▲
Google GenTabs: Labs variant of Chrome with generated mini-apps
(blog.google)
7 points
by
zhyder
9mo ago
|
1 comments
25.
▲
by
zhyder
9mo ago
Big knowledge cutoff jump from Sep 2024 to Aug 2025. How'd they pull that off for a small point release, which presumably hasn't done a fresh pre-training over the web? Did they figure out how to do more incremental knowledge upda
26.
▲
by
zhyder
10mo ago
It's all about the chip economics. I don't know how the _manufacturing cost_ of Google's TPUs compares to Nvidia's GPUs, for inference of equivalent token throughput. But at the moment Nvidia's 75-80% gross margin i
27.
▲
by
zhyder
11mo ago
"complementing the Neural Accelerators in the CPU and GPU" seems to be a misprint; I don't believe they have the accelerators in the CPU too. Still super interesting architecture with accelerators in each GPU core _and_ a ded
28.
▲
by
zhyder
1y ago
Plug for our https://uphop.ai/app : it's for adult learning / corporate training. We break down a desired job skill into small chunks, and engage the user with practice & give nuanced feedback. And of course l
29.
▲
by
zhyder
1y ago
Neural band is huge, glad they're shipping it already rather than waiting (years?) for a production version of Orion (the full AR glasses they demo'd a year ago together with this neural band). TheVerge found the controls great, e
30.
▲
by
zhyder
1y ago
Neural band is huge, glad they're shipping it already rather than waiting (years?) for a production version of Orion (the full AR glasses they demo'd a year ago together with this neural band). TheVerge found the controls great, e
More ›