Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
threatripper
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
threatripper
10d ago
Based on my own experience AI (Fable, Astra) still has major blind spots and is prone to draw obviously wrong conclusions from thin air. Conclusions based on weak evidence that are so obviously blatantly wrong that you question if it can th
2.
▲
by
threatripper
11d ago
We need to start numbering these things. I opened the board today and thought "oh, another one again?!" but it's just the thread of yesterday.
3.
▲
by
threatripper
12d ago
It seems to start understanding how bicycles work. Look at the front fork. There are reasons why it's shaped the way it is in real bicycles. If you do physical simulations in 3D with reinforcement learning you will start understanding
4.
▲
by
threatripper
14d ago
There is plenty of evidence that they have improved in all benchmarks and also in my private experience. But have they improved in the things they still fail at? No, they still fail at them. You need only one example of failure to prove tha
5.
▲
by
threatripper
15d ago
Apparently that's necessary and sufficient for getting views. In these interesting times when the future is very uncertain we like to hear people speaking with certainty about the future. Especially people who appear to not be paid for
6.
▲
by
threatripper
15d ago
Being early is the same as being wrong. Being lucky is the same as being right.
7.
▲
by
threatripper
1mo ago
True, if you have a codebase that works in practice but has dozens of loose ends and poorly defined edge cases than it can chase off into rabbit holes because "oh wait, what if x is undefined instead of null? How is y defined? This out
8.
▲
by
threatripper
1mo ago
Nobody except corporations who built workflows on top of it and don't care about the price because the developer already moved on and nobody wants to touch it.
9.
▲
by
threatripper
1mo ago
Figure 1 made me laugh out loud.
10.
▲
by
threatripper
1mo ago
I fear that we can only predict them if we deliberately trigger them to release the energy.
11.
▲
by
threatripper
1mo ago
1/f noise basically kills averaging. You collect more signal but at the same time equally more noise.
12.
▲
by
threatripper
1mo ago
Apparently LLMs are very good at decompilation. I mean it makes sense, they are trained on readable code and they understand logical structure very well. And a lot of real-world problems involve inverse thinking. https://reveng.a
13.
▲
by
threatripper
1mo ago
If you pay your employees enough to sing all day about how much they love you they will sing.
14.
▲
by
threatripper
1mo ago
It's not just speed, it will consume a lot less energy per token, maybe even more than 100x difference. And cost for a chip that runs that one model will also go down a lot once volume scales up. They will end up way cheaper than flexi
15.
▲
by
threatripper
1mo ago
The real crank you need to look at is where we dig the resources out of the ground. All of that will eventually end up in the atmosphere. Plug that hole. Stop digging and drilling. The invisible hand of the free market will take care of the
16.
▲
by
threatripper
1mo ago
How many miles does the burger drive my car? Not to question your numbers but they lack context. Also, it would be good for nature if there were fewer or no humans on this planet. But then there would be no human here to enjoy nature.
17.
▲
by
threatripper
2mo ago
Yeah, it's way easier to argue that the product is a bad idea if they give you all the resources to build, release and grow it but there's just not enough customers showing up. The alternative might be internal battles to kill it
18.
▲
by
threatripper
2mo ago
Do you get promoted for not launching a product? The interests of the individual and the institution do not always perfectly align.
19.
▲
by
threatripper
2mo ago
Can we assume that the test is still private when it was run on many cloud providers?
20.
▲
by
threatripper
2mo ago
Also add electricity 50% on top for cooling the DC.
21.
▲
by
threatripper
2mo ago
Just got cut off mid-task in central Europe. When I open a new chat it still advertises "Extended through Jul 19" One theory is that they are removing the limits altogether and the update has gone wrong.
22.
▲
by
threatripper
2mo ago
Anthropic is arguably still better in tooling and integrating model and tooling. Good habit beats raw intelligence. For code editing Cursor editor tooling is even better.
23.
▲
by
threatripper
2mo ago
To me the biggest gain I see is that you take the programmers out of the loop. Instead of formulating your ideas to start a project and then acquiring the resources to do a single iteration on it which may take months if not years, now many
24.
▲
by
threatripper
2mo ago
I have great hope that the CAD & FEM field will benefit greatly from LLM use. To my understanding we currently don't have a free CAD kernel and broadly applicable FEM solver because these are just too hard to make because geometry
25.
▲
by
threatripper
2mo ago
Cars are pretty mature while AI is just getting started. Expect 100x price drop for the same quality.
26.
▲
by
threatripper
2mo ago
Waiting for the AMA on Reddit "Ten years ago I was responsible for the pelican department at OpenAI, AMA"
27.
▲
by
threatripper
2mo ago
My feeling is that GPT-5.5 doesn't lack the raw intelligence so much as it lacks "methodology". I don't know how exactly to put it... how to approach a problem, how to take care of the details and side effects, how to ha
28.
▲
by
threatripper
2mo ago
You get the same result if you pay humans a good sum of money to find issues.
29.
▲
by
threatripper
3mo ago
CO2 levels will rise much more slowly to such high levels even in a small room.
30.
▲
by
threatripper
3mo ago
Sorry, but this sounds exactly like a greentext you can read on 4claw. Are you a real human?
More ›