Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
brap
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
brap
4d ago
Right, but this leads to my main question: is this still necessary? By analogy with code, do we still need code to be maintainable/readable if machines write it all? (Obviously for now the answer is yes, but I’m not sure this will be t
2.
▲
by
brap
4d ago
China doesn't give a fuck, next
3.
▲
by
brap
4d ago
I know it's incredibly presumptuous for me, a nobody, to say this to 25 Fields Medalists, but: Perhaps you are misaligned. Who decided the goal of math must be human insight? First off, some mathematical truths might simply be far be
4.
▲
by
brap
5d ago
There’s one thing I constantly see agents tripping over, I’m not sure what the right word for it would be, but it basically boils down to “making changes in the right places”. They seem to have very poor grasp of where things are supposed t
5.
▲
by
brap
6d ago
I’m not a fan of OAI to say the least, but having worked at similar companies, my guess is that it’s just too difficult to prove/disprove beyond a doubt, and they have other priorities
6.
▲
by
brap
6d ago
I think the line between regular LLM "endpoints" and agents/harnesses is going to become more and more blurry until it's a meaningless distinction. When you're using ChatGPT/Claude/Gemini etc. you're
7.
▲
by
brap
7d ago
Folded phones are the essence of “just because you can doesn’t mean you should”
8.
▲
by
brap
7d ago
How do you manage your frustration in these interactions? I often find myself getting pissed off
9.
▲
by
brap
14d ago
I believe the older models are being gradually phased out, newer ones have no availability issues
10.
▲
by
brap
14d ago
Just like Claude Code and others it has the same —-dangerously-skip-permissions flag, auto approves everything
11.
▲
by
brap
14d ago
Antigravity has been also rapidly improving lately, and your can also use any of the open coding harnesses. But I mostly meant “harness” as in your workflow/loop setup.
12.
▲
by
brap
14d ago
People have been sleeping on Gemini lately but these last few Flash releases (which were very rapid) are damn good. These sort of fast and cheap models are great for tasks that are verifiable and can be retried infinitely (like coding), you
13.
▲
by
brap
16d ago
Incredible. This is the kind of weaponized OCD I want in my team
14.
▲
by
brap
21d ago
Best case scenario, this ends up being abused in order to feed LLMs crap responses (or worse).
15.
▲
by
brap
23d ago
Surely more social workers will stop people from pissing on the floor
16.
▲
by
brap
23d ago
This is incredible, I wish I could have this in my city. Can’t help but wondering, how will this look like if we had AI try to “augment” these maps in real time (maybe using street view images?). I wonder if it would be playable in reasonab
17.
▲
by
brap
23d ago
>A good idea, a terrible implementation Idea is terrible, implementation is WAI.
18.
▲
by
brap
25d ago
I mean… so just HTTP + OpenAPI spec?
19.
▲
by
brap
28d ago
Nowadays most “LLM” endpoints include some sort of server side harness as well, and I’d bet more than one model involved, so it’s really just agents all the way down
20.
▲
by
brap
29d ago
Yes because governments are known for keeping things running smoothly
21.
▲
by
brap
29d ago
> only one can win Why? I mean, if you assume that the moment some threshold of intelligence is reached it will suddenly explode and self-improve at a pace no one would be able to ever catch up with, then yes probably only one can win. B
22.
▲
by
brap
29d ago
In case you were wondering, responsiblestatecraft.org is owned by "Quincy Institute for Responsible Statecraft", founded and led by no other than reknowned IRI-propagandist Trita Parsi (currently being probed by US authorities for
23.
▲
by
brap
1mo ago
I’m entirely confident that this technically pointless, especially when you consider open models exist. I believe they know damn well that this will lead nowhere, and are only doing this to mitigate criticism.
24.
▲
by
brap
1mo ago
It baffles me that someone can just send a PR without knowing the gist of how it works and why it’s done this way. If you can’t answer this basic question then why tf do we even need you around?
25.
▲
by
brap
1mo ago
I also had Squeak in my curriculum! From what I remember it had a unique object hierarchy or something like
26.
▲
by
brap
1mo ago
While the demo is incredible, I think that in most practical use-cases, models aren't very useful without tools (search, code execution, etc.). Even if we assume reasoning latency drops to ~0ms (AFAIK this demo doesn't include rea
27.
▲
by
brap
1mo ago
Everybody thinks they have good taste don’t they. How convenient it is that we’re all so awesome according to the one metric that can’t be measured
28.
▲
by
brap
1mo ago
Many years ago, back when I was a junior, I wrote something like this to my manager, complaining about another employee who has wronged me. I eventually decided not to send it because it just seemed so unprofessional. Wild that a company th
29.
▲
by
brap
2mo ago
>Let’s examine another crazy version. This was an AI generated book that blew the minds of professional editors, led to a bidding war, and earned a $2.4M advance. Let’s examine a third version: the people investing $2.4M into this book u
30.
▲
by
brap
2mo ago
I tried letting Fable have a few passes at my docs (ultracode) and the output was basically unreadable. Nothing a human would ever write. I wonder if limiting them to a certain style like STE upfront would make them perform better/wors
More ›