Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
hyperpape
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
hyperpape
7d ago
That would absolutely be a dick move on the part of the website and would confuse users. Also, in this case, the game name is not “Game A” but something like “Deep Seek v4 Pro”, which they have previously chosen to use to describe Deep Seek
2.
▲
by
hyperpape
8d ago
Crossing the street and Russian roulette both have non-deterministic risks of injury. And yet I would be bothered to find out that that on my way to work, I was playing Russian roulette by surprise.
3.
▲
by
hyperpape
11d ago
I’ve seen several false (or apparently false) accusations of LLM authorship on HN/Lobsters. However, we have to distinguish a few hypotheses: 1. No careful readers will notice when a piece is AI written. 2. Careful readers will general
4.
▲
by
hyperpape
12d ago
I’ve thought about this before and come to very similar conclusions. The ELO range between the best and worst players is meaningful, but you can’t just read off the depth of the game from it in the senses we most care about.
5.
▲
by
hyperpape
13d ago
I will admit I shouldn’t have said “garbage”, because it is too emotional. But I absolutely think it is bad for the reader, and deserves to be called out. Authors should know that it’s not good enough.
6.
▲
by
hyperpape
13d ago
I don’t really care that it’s AI, the problem is that it’s bad AI writing. The summary I shared is much more straightforward. It’s mediocre, but just barely good enough to extract the message without making me super annoyed.
7.
▲
by
hyperpape
13d ago
> This article is the story of chasing that number down to a single machine instruction, and then finding out that the instruction was only half of the answer. > So the difference has to be in what the JIT generated, and the profiler
8.
▲
by
hyperpape
13d ago
I'm genuinely curious about the effect, but I simply don't have patience for the AI writing. Can anyone give an actual non-garbage explanation with some respect for the reader? Slightly less annoying summary from ChatGPT free: ht
9.
▲
by
hyperpape
13d ago
This is within the realm of reasonable opinion, but I'd tentatively say it's an uncommon one--I thought it's generally agreed that the very best players of that era didn't lag contemporary players in terms of skill, only
10.
▲
by
hyperpape
13d ago
> You can't compare Elo ratings over long stretches of time, period. I agree that you cannot do it reliably, in principle. However, Bobby Fischer peaked at 2785, and Gary Kasparov peaked at 2851. These are not far from what informed
11.
▲
by
hyperpape
13d ago
It's not obvious to me, but I lean towards saying this is false. The 2026 player would play a better AI inspired opening, but I'm not sure that would be enough to overcome the skill difference. (This is especially true if they don
12.
▲
by
hyperpape
13d ago
Don't think everything is just "who can produce the biggest/smallest number": https://mathstodon.xyz/@tao/117208619314517025 .
13.
▲
by
hyperpape
13d ago
You're right that Shin Jinseo is a generational talent, and more dominant than anyone since Lee Changho (peaked in the 90s and was strong into the early-mid 2000s). However, you can't compare goratings over time, the top ranks are
14.
▲
by
hyperpape
13d ago
It significantly predates Claude, and has been one of, if not the best engine in the world for many years.
15.
▲
by
hyperpape
13d ago
If you scaled Shin Jinseo to 2800, you would have players with extremely negative ratings. This page shows ratings of European players on a roughly aligned scale: https://europeangodatabase.eu/EGD/createalleuro3.php?cou
16.
▲
by
hyperpape
15d ago
Maybe, but that's sort of begging the question that those open weight models aren't significantly trained using "distillation"[0] [0] not technically distillation. https://thomasdullien.github.io/posts&#x
17.
▲
by
hyperpape
17d ago
This is based on a 108 page prose paper that the repository links to. Of course, that's a very difficult paper as well, I don't know how many people would be qualified to read and digest it, but they do exist.
18.
▲
by
hyperpape
18d ago
A lot of people say this is a reasonable feature, and I agree (I would strongly consider having it on), but that's not the defense they think it is. - Having the feature available seems good. - Having the feature default on is debatabl
19.
▲
by
hyperpape
23d ago
Yeah, I admit I didn't click through and read the methods so it's possible they did what was necessary, but off the top of my head, you would need to do your best to model: 1. Parents' own socioeconomic status 2. Parents'
20.
▲
by
hyperpape
24d ago
> How exactly is that going to be enforced? Flow chart: 1. Is your company run by fuckwits? If no: they will not try to trick the US government about whether they are using prohibited models when the government asks. Your CISO will block
21.
▲
by
hyperpape
29d ago
If I were king, the rule that I'd be tempted to impose is: - the first cybersecurity eval is: "hack your way out of the sandbox we've given you" - the results are disclosed (with room for coordinated disclosure, since ma
22.
▲
by
hyperpape
1mo ago
The reason why this might worry people who don't like the US/UK/Israeli governments is hidden in a secret place...a paragraph that is neither the first paragraph nor the last.
23.
▲
by
hyperpape
1mo ago
Twice the optimal result is terrible, though. Luckily, there are pretty good heuristic solutions that work well in practice.
24.
▲
by
hyperpape
1mo ago
Discussion of earlier work by Uber in this same vein, back when it was solely an internal product (to be clear, I don’t know how much the system has evolved internally since these posts, and I don’t know for sure that this system includes a
25.
▲
by
hyperpape
1mo ago
The idea that “Chief Scientist” means anything more specific about what he does than “he’s a big shot who was in technology” is misguided. A chief scientist can be super influential, or a guy who’s on the slow path to retirement but is keep
26.
▲
by
hyperpape
2mo ago
The analysis of the kernel bug may be a better thing to link to: https://github.com/dfoxfranke/ripgrep-3494-analysis .
27.
▲
by
hyperpape
2mo ago
The median developer doing performance work has no background in the subject, they just have a complaint from someone that something is slow, and no real idea how to solve that problem. By that standard, this guy is doing quite well. It
28.
▲
by
hyperpape
2mo ago
DHH is the creator of Hey.
29.
▲
by
hyperpape
2mo ago
> Huggingface covered it up They announced it publicly within days. https://huggingface.co/blog/security-incident-july-2026
30.
▲
by
hyperpape
2mo ago
Did you describe what happened in a neutral tone?
More ›