Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
timbilt
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
NanoClaw's architecture is a masterclass in doing less
(jonno.nz)
47 points
by
timbilt
5mo ago
|
13 comments
2.
▲
Google open-sources experimental agent orchestration testbed Scion
(infoq.com)
230 points
by
timbilt
5mo ago
|
62 comments
3.
▲
Kubernetes Resource Optimization Strategies That Work in Production
(scaleops.com)
7 points
by
timbilt
1y ago
|
1 comments
4.
▲
by
timbilt
1y ago
Yes, but in a case like this it's a neutral third-party running the benchmark. So there isn't a direct incentive for them to favor one lab over another. With public benchmarks we're trusting the labs not to cheat. And it'
5.
▲
by
timbilt
1y ago
> Unlike many public benchmarks, the PR Benchmark is private, and its data is not publicly released. This ensures models haven’t seen it during training, making results fairer and more indicative of real-world generalization. This is key
6.
▲
Google's Reverse Acquihire of Windsurf and the Future of AI Developer Tools
(qodo.ai)
14 points
by
timbilt
1y ago
|
5 comments
7.
▲
Making file encryption fast and secure for teams with advanced key management
(dropbox.tech)
2 points
by
timbilt
1y ago
|
0 comments
8.
▲
Obesity drugs show promise for treating a new ailment: migraine
(nature.com)
2 points
by
timbilt
1y ago
|
0 comments
9.
▲
Slashing CI Costs at Uber
(uber.com)
1 points
by
timbilt
1y ago
|
0 comments
10.
▲
Detecting and Countering Malicious Uses of Claude
(anthropic.com)
2 points
by
timbilt
1y ago
|
0 comments
11.
▲
Uber's Journey to Ray on Kubernetes: Ray Setup
(uber.com)
1 points
by
timbilt
1y ago
|
0 comments
12.
▲
Why we chose LangGraph to build our coding agent
(qodo.ai)
17 points
by
timbilt
1y ago
|
9 comments
13.
▲
by
timbilt
2y ago
anyone else concerned that training models on synthetic, LLM-generated data might push us into a linguistic feedback loop? relying on LLM text for training could bias the next model towards even more overuse of words like "delve"
14.
▲
The Hidden Costs of Men's Social Isolation
(scientificamerican.com)
35 points
by
timbilt
2y ago
|
5 comments
15.
▲
Looking back at our Bug Bounty program in 2024
(engineering.fb.com)
1 points
by
timbilt
2y ago
|
0 comments
16.
▲
by
timbilt
2y ago
Twitter thread about this by the author: https://x.com/jonasgeiping/status/1888985929727037514
17.
▲
Scaling up test-time compute with latent reasoning: A recurrent depth approach
(arxiv.org)
149 points
by
timbilt
2y ago
|
44 comments
18.
▲
AI is accelerating scientific production, not progress
(twitter.com)
5 points
by
timbilt
2y ago
|
0 comments
19.
▲
The Sport of Tuning Quantum Dot Arrays: QDarts Simulator
(quantum-machines.co)
1 points
by
timbilt
2y ago
|
0 comments
20.
▲
by
timbilt
2y ago
Until we get real-time learning to work in production, every AI tool feels like it's getting dumber over time. It goes very quick from "wow this is magic" to starting to notice all the little gaps. I think we have a fundament
21.
▲
by
timbilt
2y ago
The weirdness of LLMs is that they're so damn good at so many things but then you see these glaring gaps that instantly make them seem dumb. We desperately need benchmarks and evals that test these kinds of hard to pin down cognitive a
22.
▲
Improving Search Ranking for Maps
(medium.com)
1 points
by
timbilt
2y ago
|
0 comments
23.
▲
Do We Live in a Special Part of the Universe?
(scientificamerican.com)
3 points
by
timbilt
2y ago
|
0 comments
24.
▲
by
timbilt
2y ago
Unit tests are more commonly written to future proof code from issues down the road, rather than to discover existing bugs. A code base with good test coverage is considered more maintainable — you can make changes without worrying that it
25.
▲
by
timbilt
2y ago
> validates each test to ensure it runs successfully, passes, and increases code coverage This seems to be based on the cover agent open source which implements Meta's TestGen-LLM paper. https://www.qodo.ai/blog
26.
▲
Introducing Qodo Cover: Automate Test Coverage
(qodo.ai)
16 points
by
timbilt
2y ago
|
8 comments
27.
▲
Does Sleep Training Work?
(scientificamerican.com)
2 points
by
timbilt
2y ago
|
0 comments
28.
▲
Bird-inspired leg enables robots to jump into flight
(nature.com)
3 points
by
timbilt
2y ago
|
0 comments
29.
▲
Dwarf planet might have its own ice volcano
(nature.com)
2 points
by
timbilt
2y ago
|
0 comments
30.
▲
by
timbilt
2y ago
https://archive.md/Bn3Dz
More ›