Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
niklassheth
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
niklassheth
7d ago
I'm happy to answer any questions :)
2.
▲
by
niklassheth
5mo ago
And on secondary markets shares are trading for at least twice that...
3.
▲
by
niklassheth
6mo ago
So many problems with this: The benchmark is totally useless. It measures single prompts, and only compares output tokens with no regard for accuracy. I could obliterate this benchmark with the prompt "Always answer with one word"
4.
▲
by
niklassheth
9mo ago
Nice! Your comparison site is probably the best one out there for image models
5.
▲
by
niklassheth
10mo ago
I put the output from this tool into GPT-5-thinking. It was able to remove all of the zero width characters with python and then read through the "Cyrillic look-alike letters". Nice try!
6.
▲
by
niklassheth
10mo ago
This is more evidence that Cognition's SWE-1.5 is a GLM-4.6 finetune
7.
▲
by
niklassheth
1y ago
It seems like the repo is mostly if not entirely LLM generated; not a great sign.
8.
▲
by
niklassheth
1y ago
I know some consumer cards have artificially limited FP64, but the AI focused datacenter cards have physically fewer FP64 units. Recently, the GB300 removed almost all of them, to the point that a GB300 actually has less FP64 TFLOPS than a
9.
▲
by
niklassheth
1y ago
I've found the same, but I also haven't gained much value out of "deep research" products as a whole. When I last tested them with topics I'm familiar with, I found the quality of research to be poor. These tools se
10.
▲
by
niklassheth
1y ago
The majority of phones in the US are iPhones, especially in big cities where phone theft is most common.
11.
▲
by
niklassheth
1y ago
I've also found it to be good at digging deep on things I'm curious about, but don't care enough to spend a lot of time on. As an example, I wanted to know how much sugar by weight is in a coffee syrup so I could make my own