Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Azantys
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
Azantys
12d ago
The whole point was for the formalization to be clean enough so it could be reused in other parts of mathematics as I understand it. 13M lines of AI slop which have never been checked do not sound like what the original goal for such a form
2.
▲
by
Azantys
25d ago
Yes, but you can use local models with opencode/pi
3.
▲
by
Azantys
1mo ago
Its 27B not 37B and having just 27B in total and 3T and like 30B active of those is still totally different. A 120B with 5B active is still much slower than a proper 5B. Just like the new Ling 3.0 Tiny with 8B and 1B active only gets around
4.
▲
by
Azantys
1mo ago
At one point there were specific -Coder release of Qwen e.g. 2.5 but they dropped that, still wondering how much better 3.6-Coder or 3.8-Coder would be when they ignore everything else
5.
▲
by
Azantys
1mo ago
Benchmarks are basically tied https://artificialanalysis.ai/models/comparisons/deepseek-v4...
6.
▲
by
Azantys
1mo ago
Dysfunctional vibeslop
7.
▲
by
Azantys
1mo ago
Here you have shown yourself that progress slows down and doesnt speed up. 8.9/0.35 = ~25x more performance in 10 years from 2006 to 2016. 104.8/8.9 = ~12x more performance in 10 years from 2016 to 2026. Growth has dropped 50%.
8.
▲
by
Azantys
1mo ago
You already can with AMD/Intel gpus. Kraid is a compiler for Arm Mali which is mostly used on mobile stuff, nothing AI related.
9.
▲
by
Azantys
2mo ago
I questioned the speed not the output quality, that is another discussion
10.
▲
by
Azantys
2mo ago
? Readme says 60-70s per token
11.
▲
by
Azantys
2mo ago
0.01 tk/s is unusable for anything, you would wait a whole day for just 1000 token of output, what is the point of projects like this?
12.
▲
by
Azantys
2mo ago
Free month is insane for a bit of downtime
13.
▲
by
Azantys
2mo ago
Just looking with ncdu/du or another tool which sorts folders by size would have been much faster and easier, was there any need to use Claude here?
14.
▲
by
Azantys
2mo ago
If you could render fully in OpenCL and output it via another device it would probably work. But they dont have ROPS so they cant do traditional video rendering. Even something like Blender where you would use it for calculations and not fo
15.
▲
by
Azantys
2mo ago
It really depends on the language, popular languages work pretty good
16.
▲
by
Azantys
2mo ago
You are comparing a 35B models to a 635B+ frontier model, of course thats not even close
17.
▲
by
Azantys
2mo ago
Old datacenter GPUs could game, but new ones can only do OpenCL/CUDA/ROCm etc. and have no display out. Im using an MI50 right now and I would wish newer datacenter cards could also be used for everything like them.
18.
▲
by
Azantys
2mo ago
Useless comment
19.
▲
by
Azantys
2mo ago
The wording wasnt very good I ment compared to programming or math the amount of logic and reasoning is small (Research level math hardly compares to writing a book in raw reasoning and logic). And I thing the smaller models have enough &qu
20.
▲
by
Azantys
2mo ago
Of course there is logic but its nowhere near the complexity of math or programming
21.
▲
by
Azantys
2mo ago
Do people really use 100B+ models for writing? I am no writer but to me it seems like writing is one of the easiest tasks with barely any logic or reasoning and as long as its not longer than a handful of pages I expect even 8B models to pe
22.
▲
by
Azantys
3mo ago
I think model training is pretty hard to do efficiently on a vastly distributed network. If the model cant fit into the VRAM of the node your performance becomes so bad its useless, so a distributed model could only be properly trained if t
23.
▲
by
Azantys
3mo ago
+1
24.
▲
by
Azantys
3mo ago
Game studios should choose CryEngine/Decima again over UE5
25.
▲
by
Azantys
3mo ago
Just dont upgrade the Mainboard firmware then
26.
▲
by
Azantys
3mo ago
Yeah but it wasnt close to Opus etc. Still a good local model when it released
27.
▲
by
Azantys
3mo ago
Galaxy AI 3.8-Flash-Plus Max (xhigh)
28.
▲
by
Azantys
3mo ago
Isnt that more Perplexitys thing anyways?
29.
▲
by
Azantys
3mo ago
Career and personal advice from LLMs, not sure if thats your best bet
30.
▲
by
Azantys
4mo ago
Is LPDDR5X not too slow for inference, atleast compared against HBM?
More ›