Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ttul
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
ttul
4d ago
We built a “code atlas” that provides the LLM with a semantically queryable map of how things connect and relate in a very large and sprawling codebase that evolved over 15 years. It tends to dramatically reduce the length of time models ha
2.
▲
by
ttul
5d ago
Just have Codex or whatever take the screenshots and post them to the repo. Easy peasy.
3.
▲
by
ttul
5d ago
Most surprising in this article is the vibrant HDR magenta circle on Matt’s Twitter.
4.
▲
by
ttul
9d ago
I think Sam would be saying something pithier. And the 20x Pro plan absolutely runs at a loss, so I think he would be the last one to promote it.
5.
▲
by
ttul
9d ago
My goodness, the complaining... Just get all your devs a $200 ChatGPT Pro (20x) plan. Yes, you lose the "team" component, but you gain so much more. And what's $200 against the salary of a good developer? It's absolutely
6.
▲
by
ttul
9d ago
"Canadian who helped design splash-free urinal is proud of his weird science award" https://www.cbc.ca/radio/asithappens/2026-ig-nobels-9.733140...
7.
▲
by
ttul
9d ago
Yes, and that's a great idea. Chapter 11 protects companies from their creditors. If a suitor comes along offering to buy the company that is in Chapter 11, the creditors can go to the judge and request that the judge allow the purchas
8.
▲
by
ttul
10d ago
It is likely too specific to our particular orchestration approach to be useful as open source. But the recipe is replicable. You can just paraphrase this prompt: “Starting with our Helm charts, which you will find in /folder-of-chart
9.
▲
by
ttul
11d ago
I got Astra to build an interactive website that provides developers with an atlas of our source code, giving it the Helm charts that describe our cloud services and telling it to work backwards to the source code that runs everything. The
10.
▲
by
ttul
12d ago
This is undoubtedly true. Agents are extremely analytical and trained to be objective - far more so than humans. They are not driven by emotion. If you have good stuff and you make it extremely clear to everyone through your documentation,
11.
▲
by
ttul
13d ago
I built this for my own company. Armature is on to something. You start by analyzing the choices agents would make for various use cases and then glean what, if anything, you might do to start tilting the agents in the direction of your own
12.
▲
by
ttul
13d ago
"The gym's doors were mysteriously removed from their hinges during the night. The gym equipment was also apparently stolen. And the school's custodian was found incoherent next to a bottle of top-shelf Scotch."
13.
▲
by
ttul
14d ago
Will look forward to the "feel" of the model in real testing. But I agree that these benchmarks do get "dealt with" rapidly. That's a shame, but I guess it's the times we live in.
14.
▲
by
ttul
14d ago
Crushing it on DeepSWE is a very big deal. Excited to give this a try.
15.
▲
by
ttul
15d ago
Most people running local models would probably love to run larger models if only they had access to big enough hardware. I'm curious: to those of you running models locally, if there was a way to inference the model of your choice at
16.
▲
Endless sitcom using Minimax H3 and a turbo LoRA
(twitch.tv)
5 points
by
ttul
17d ago
|
5 comments
17.
▲
by
ttul
19d ago
It burns out so quickly on the 5x plan. Better than nothing, I suppose, but I don't know how I would survive on a 5x plan given that I burn out more than one 20x plan monthly.
18.
▲
by
ttul
23d ago
Luna is a very capable model - thanks for pointing that out. Terra is the strange one: not cheap enough or intelligent enough to be on the frontier. But Luna sure is.
19.
▲
by
ttul
23d ago
Fable 5 is just straight up a larger model - I'm guessing at this, but there is plenty of evidence online from people far more plugged in than I am. OpenAI is pursuing a strategy that yields greater operating margins and penetration of
20.
▲
by
ttul
24d ago
My sense is that Fable 5 has “taste”. But Sol gets to work and gets shit done. I reserve Fable for when things need a refresh or if I want a flawless front end. Sol does the majority of actual work. I max out two of each at the Max/Pro
21.
▲
by
ttul
26d ago
The DeepSWE benchmark they report (59.3%) overlaps with the confidence interval of 5.6-Sol Medium (61% +/- 2%), but likely at 1/18th the cost (they did not report the DeepSWE benchmark cost, but v4-flash had this cost ratio agains
22.
▲
by
ttul
26d ago
A good share of humanity would have also gotten this question wrong!
23.
▲
by
ttul
27d ago
If you ask Sol or Claude how much time it will take to implement a plan they just came up with, they usually advise a timeframe in the weeks or months - assuming, I suppose, that human programmers will be building it. And then you ask the m
24.
▲
by
ttul
28d ago
Poor BrandonM... He has not been invited to cocktail parties at all since then.
25.
▲
by
ttul
28d ago
Hot water for the whole neighbourhood!
26.
▲
by
ttul
29d ago
Indeed. You need 45 to 60 liters per second of cooling water flowing over a Cerebras wafer every minute to keep it under 90C. And that’s assuming the water leaves at 90C… More realistically, you need much more cooling water.
27.
▲
by
ttul
29d ago
I’ll get that 250kW home power service dropped in next week!
28.
▲
by
ttul
1mo ago
(Otherwise it would have slowly transformed into CO2 over the eons without our help)
29.
▲
by
ttul
1mo ago
I think there is literally no oxygen down there with the hydrogen. It’s the same with hydraulic fracturing of hydrocarbons. Plenty of methane COULD go boom, but there is no oxygen deep underground.
30.
▲
by
ttul
1mo ago
I am waiting with bated breath to read, “load-bearing” somewhere… The latest models are very capable, but sometimes they seem to get so deep in the details that they lose the overall plot. What the hell is the point of this page? Can you pu
More ›