Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
somebodythere
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
somebodythere
21d ago
Did you read the article?
2.
▲
by
somebodythere
21d ago
Not with the help of the models!
3.
▲
by
somebodythere
2mo ago
This prediction can't be scored until the 2028 election cycle. You may think it's very unlikely the prediction will have turned out to be correct by the 2028 election cycle, but that is not the same thing as the prediction being s
4.
▲
by
somebodythere
2mo ago
The thing about exponentials is if you admit 60%, it's pretty easy to admit 95%.
5.
▲
by
somebodythere
3mo ago
That is not true. You can tell you are on the latter part of the S-Curve you are on, if the rate of change of capabilities has decreased compared to before. That is not what we are seeing right now. The rate of change is increasing, or is a
6.
▲
by
somebodythere
5mo ago
There is some evidence suggesting that "blue zones" are largely about pension fraud. https://fortune.com/europe/2024/12/14/are-blue-zones-myth-ex...
7.
▲
by
somebodythere
6mo ago
Using your API key in third-party harnesses has always been allowed. They just don't like using the subsidized subscription plan outside of first-party harnesses. So this seems to be out of spite
8.
▲
by
somebodythere
7mo ago
even a squirrel that needs guidance from a human grandmaster, is heavily inspired by existing games, and who can use Piece Mover library is incredible. 5 years ago the squirrel was just a squirrel. then it was able to make legal moves. now
9.
▲
by
somebodythere
8mo ago
No. You can point e.g. Opencode/Cline/Roo Code/Kilo Code at your inference endpoint. But CC has high install base and users are used to it, so it makes sense to target it.
10.
▲
by
somebodythere
9mo ago
Why would I ask the model to reverse the string 'glorbix,' especially in the context of software engineering?
11.
▲
by
somebodythere
9mo ago
You personally wouldn't use live captions and dubbing, so there's no point building it for the millions of people who need it as an accessibility feature?
12.
▲
by
somebodythere
10mo ago
Rufus is a Claude Haiku, yes.
13.
▲
by
somebodythere
10mo ago
I've seen a few of this type of thing pop up in search results ("DeepWiki" by Cognition.) I'm not a fan. It is just LLM contentslop, basically. Actual wikis written by humans are made of actual insight from developers an
14.
▲
by
somebodythere
10mo ago
I think this wound up being close enough to true, it's just that it actually says less than what people assumed at the time. It's basically the Jevons paradox for code. The price of lines of code (in human engineer-hours) has decr
15.
▲
by
somebodythere
1y ago
Roughly, this is the Electronic Frontier Foundation (and comparable lobbying orgs in other countries.) However, an org like this doesn't have much power to compel individuals to give them $1.
16.
▲
by
somebodythere
1y ago
LLM argumentative essays tend to have this "gish-gallop" energy; say a bunch of tenuously related and vaguely supported things, leave the reader wondering if it was the author who failed to connect the dots, or them
17.
▲
by
somebodythere
1y ago
Maybe it's my engineer-brain talking, but "lab-grown" actually biases me towards the diamonds. Feels precise and futuristic.
18.
▲
by
somebodythere
1y ago
I don't know if it matters. Even if the best we can do is get really good at interpolating between solutions to cognitive tasks on the data manifold, the only economically useful human labor left asymptotes toward frontier work; work t
19.
▲
by
somebodythere
1y ago
My guess is that they did RLVR post-training for SWE tasks, and a smaller model can undergo more RL steps for the same amount of computation.
20.
▲
by
somebodythere
1y ago
I see what you are getting at. My point is that if you train and agent and verifier/governor together based on rewards from e.g. RLVR, the system (agent + governor) is what will reward hack. OpenAI demonstrated this in their "Lear
21.
▲
by
somebodythere
1y ago
Because if the agent and governor are trained together, the shared reward function will corrupt the governor.
22.
▲
by
somebodythere
1y ago
I took your original post to mean that AI researchers' and AI safety researchers' expectation of AGI arrival has been slipping towards the future as AI advances fail to materialize! It's just, AI advances have been material
23.
▲
by
somebodythere
1y ago
AGI timelines have been steadily decreasing over time: https://www.metaculus.com/questions/5121/date-of-artificial-... (switch to all-time chart)
24.
▲
by
somebodythere
1y ago
Did you see the supplemental material that explains how they arrived at their timelines/capabilities forecasts? https://ai-2027.com/research
25.
▲
by
somebodythere
2y ago
Federal interests can easily tell the local prosecutor "hey, don't prosecute this, it risks setting bad precedent".
26.
▲
by
somebodythere
2y ago
Seems not tinfoil and rather plausibly a pragmatic decision by the prosecution.
27.
▲
by
somebodythere
2y ago
The market is liquid enough to absorb a sale for $400K.
28.
▲
by
somebodythere
2y ago
Sure. Presumably also the developers at FooLabs would like to continue having a job developing Foo, and the broader software community would like to continue benefiting from additional features and improvements to Foo, which probably wouldn
29.
▲
by
somebodythere
2y ago
To start an X session after logging in
30.
▲
by
somebodythere
2y ago
Instruction is not the only way to interact with an LLM. In tuning LLMs to the assistant persona, they become much less useful for a lot of tasks, like naming things or generating prose.
More ›