Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
einrealist
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
einrealist
4d ago
Can't imagine that RSI will result in something useful in practice. There are serious obstacles like alignment drift and model collapse. All while we still cannot solve the accuracy problem with current frontier models.
2.
▲
by
einrealist
5d ago
And there is another problem: LLMs generating too much code, code that is doing more than was asked. And that cannot be fixed by tests. Usually, we create tests for wanted behavior and expected exceptions. But we don't create tests for
3.
▲
by
einrealist
15d ago
Client-side anti-cheat is a band-aid that will always punish legit players and merely hinder cheaters. 'The client cannot be trusted' applies to simple HTTP exchanges as well as to game state. Of course, if the server has to track
4.
▲
by
einrealist
20d ago
A dedicated status code + Location header allows to decouple the way an advisory is presented completely. At that location, there can be any representation, HTML or structured data like JSON-LD. The RFC extension would be small and clean.
5.
▲
by
einrealist
20d ago
There is something I'd call Subscription Fatigue. People and businesses are asked to pay subscription fees left and right. If I convert my bank statements only each quarter in a spike or once or twice a year, why should I pay for each
6.
▲
by
einrealist
26d ago
But I want to specifically signal a security concern to the requesting party. For example, a artifact proxy (like Artifactory or CodeArtifact) can use this response to warn developers.
7.
▲
by
einrealist
27d ago
If anyone wants to create a revision to RFC 9110 :) HTTP/1.1 309 Security Advisory Location: https://acme/aaargh-another-advisory
8.
▲
by
einrealist
29d ago
Or 2x less. 800B valued is not 800B invested. And it wouldn't be the first time investors sell at a loss if sentiment declines.
9.
▲
by
einrealist
29d ago
This is so true! I rarely order anything from Amazon these days, maybe once or twice a year. Everything feels like a rip-off or fake. I don't trust the search function at all. Nor do I trust the customer ratings. The only exception is
10.
▲
by
einrealist
2mo ago
Yes, but isn't football already overly commercialised? Why is advertising for sports betting or alcohol any better than the FIFA issue? UEFA should take a look at itself, too.
11.
▲
by
einrealist
3mo ago
Yet another band aid.
12.
▲
by
einrealist
3mo ago
Nice summary. Reading this reminds me about the strong encryption discussion. > We optimize what we can measure, not what we actually want to achieve. We hope and pray that these are the same thing, but they often aren’t. He points out t
13.
▲
by
einrealist
4mo ago
Maintaining support for Windows is free?
14.
▲
by
einrealist
4mo ago
"We do not anticipate declaring or paying any cash dividends to holders of our common stock in the foreseeable future." Sounds like 'never' to me.
15.
▲
by
einrealist
4mo ago
Must be retail investors believing: big number == good.
16.
▲
by
einrealist
4mo ago
So good SEO will require prompt injection now?
17.
▲
by
einrealist
4mo ago
I strongly doubt it will even provide us with a roof over our heads. In an unconstrained market, the pressure to extract as much as possible from the UBI will be enormous. The amount of UBI will probably always lag years behind the actual a
18.
▲
by
einrealist
4mo ago
Those price increases will increase the pressure to use cheaper / free models (commoditization), thus cutting into the revenue projections of the frontier model vendors. Its going to be exciting to see what happens to these huge invest
19.
▲
by
einrealist
4mo ago
Isn't alignment a dilemma? Because what is aligned, how and for whom? And who decides how that alignment should look like? There are probably many domains in which required alignment is in conflict with each other (e.g. using LLMs for
20.
▲
by
einrealist
4mo ago
Compute in science was already subsidized by public funding or by donations. Most supercomputers are financed this way. And that's a good thing. If you have a good science problem that can be computed, apply for compute time. There is
21.
▲
by
einrealist
4mo ago
Did I praise our animal agriculture anywhere?
22.
▲
by
einrealist
4mo ago
"After 16 minutes and 41 seconds, it came back" ... "further 47 minutes and 39 seconds" ... "After 13 minutes and 33 seconds" ... "After 9 minutes and 12 seconds" ... "After 31 minutes and 40 sec
23.
▲
by
einrealist
5mo ago
I wonder how this figure was settled. Is it based on consumer pricing? Can't Microsoft and OpenAI just make a number up, aside from a minimum to cover operating costs? When is the number just a marketing ploy to make it seem huge, impo
24.
▲
by
einrealist
5mo ago
Also funny how people (including LLM vendors, like Cursor) think that rules in a system prompt (or custom rules) are real safety measures.
25.
▲
by
einrealist
5mo ago
Is 'refactoring Markdown files' already a thing?
26.
▲
by
einrealist
5mo ago
True. I didn't expect it to provide novel designs. Maybe Anthropic should find a better replacement for 'Design'. In my example, I expected it to create UI elements for a business application / expert system. And it did
27.
▲
by
einrealist
5mo ago
Good for crunching out some prototypes, ideas and getting inspirations I guess. Two prompts - the initial one and one refinement - took about ten minutes and used up 90% of the token budget. I wonder what the real costs are. After the IPO,
28.
▲
by
einrealist
5mo ago
They are trying to optimize the circus trick that 'reasoning' is. The economics still do not favor a viable business at these valuations or levels of cost subsidization. The amount of compute required to make 'reasoning'
29.
▲
by
einrealist
5mo ago
I can slow down the compute by a factor of a thousand. It would not change the result. But it changes the economics. We only call it intelligent, because we can do the backpropagation, the inference (and training) fast enough and with enoug
30.
▲
by
einrealist
5mo ago
I don't trust anyone who claims that LLMs today are superhumanly intelligent. All they do is perform compute-intensive brute-force attacks on the problem/solution space and call it 'reasoning', all while subsidising the
More ›