Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kristianp
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
Chess.com Leak Exposes 7.3M Users, Evidence Points to Scraping
(securityaffairs.com)
85 points
by
kristianp
3d ago
|
24 comments
2.
▲
by
kristianp
5d ago
That first diagram is striking: that deepseek's own inference is at least 5% higher on tool calling (TAU Bench) than most other providers. I wonder if they make sure their responses are valid json at the token generation level using a
3.
▲
by
kristianp
7d ago
That's not far off the 4:3 of the Samsung Z fold 8. Around 4.24:3.
4.
▲
by
kristianp
9d ago
Sector C might have some insights into how to make this even smaller or more featurefull. It has some interesting hacks. https://xorvoid.com/sectorc.html
5.
▲
by
kristianp
9d ago
Npm always finds its way in, usually for building web artifacts.
6.
▲
by
kristianp
10d ago
> Any "w" is a "while" Meaning that something as simple as "w = 4" would fail? A little too nasty for my liking. Not a choice I would have made, but admire the amount of work done here and the readability o
7.
▲
by
kristianp
10d ago
Are you talking about using a thunderbolt dock and GPU? I doubt that's a use case high in the developers minds. I'm curious though, I wonder if the open source NVIDIA driver can be built for it.
8.
▲
by
kristianp
12d ago
They should use their portal to de-claude the writing.
9.
▲
by
kristianp
15d ago
LLMs token generation is memory bandwidth constrained. If the M7 has double the memory bandwidth as some speculate [1], then it will help with LLM performance. There is also expectation (possibly unfounded) that the GPU will have improved
10.
▲
by
kristianp
15d ago
The data center growth questions he raised have been where I found him interesting. I could never find any other articles to corroborate his predictions though. My question is when will the AI bubble burst? I didn't know he'd had
11.
▲
by
kristianp
16d ago
I agree about getting the Ultra if you're interested in AI (LLM) inference speed. I'm a little perplexed as to why there isn't a RAM option in between 96GB and 256GB, though. For instance, I believe Deepseek v4 flash runs a
12.
▲
by
kristianp
17d ago
One problem with Anubis is that once you've solved the POW once, you just need to hold the cookie to avoid solving it again. Scrapers have probably learnt to do that by now. So Anubis isn't as effective as it used to be before it
13.
▲
by
kristianp
20d ago
> I'd skip M5 and M6 chips for LLM work and wait for a year for M7. Another note on this, it's likely the M7 ultra will be released around 6 months after the M7 Max, if the M1 and M5 are taken as reference. So the M7 ultra may
14.
▲
The choices we make about AI now are critical
(gatesnotes.com)
3 points
by
kristianp
21d ago
|
1 comments
15.
▲
by
kristianp
21d ago
Both of you should mention what quant you're using. And as another comment said, what tasks you're doing, i.e. coding, classification, summarizing etc.
16.
▲
by
kristianp
22d ago
My Intel laptop is pretty close to silent on Linux. ThinkPad p14s. As soon as it boots into Windows 11 the fan is very audible.
17.
▲
by
kristianp
22d ago
> I'd skip M5 and M6 chips for LLM work and wait for a year for M7. Or you could lease an M5 max/ultra until the M7 equivalent comes out. At least in the US a leasing option is available.
18.
▲
by
kristianp
22d ago
I've had a 1GiB VPS for a while. I'm starting to think that I could have done with 512MB for the fairly simple jobs I have on it. As long as I get 1 whole CPU slice (no slowdowns) I'd be ok. The 10GB SSD would be tight, but
19.
▲
What the Vancouver Stock Exchange Can Teach Us About Rounding Numbers in VBA
(nolongerset.com)
4 points
by
kristianp
22d ago
|
0 comments
20.
▲
Hot Chips 2026: Samsung makes LPDDR5X smart with logic unit in memory
(tomshardware.com)
4 points
by
kristianp
22d ago
|
1 comments
21.
▲
by
kristianp
22d ago
Their statement about sharing an Ultra for an office LLM server is a little misleading. The tokens/second would be too slow for any business use case. You'd want to buy multiple GPUs for that purpose, probably Nvidia, despite th
22.
▲
by
kristianp
23d ago
Thanks. I see but 2 RAM chips on one photo.
23.
▲
by
kristianp
23d ago
https://xcancel.com/IntCyberDigest/status/209195400829849193... I wonder if these are M5 Ultra machines?
24.
▲
by
kristianp
26d ago
This is by the author of "Designing Data-Intensive Applications", one of HNs favourite books. https://dataintensive.net/
25.
▲
by
kristianp
26d ago
Agreed. The AI news cycle is constant and promises the eventual end to human usefulness. I feel like one day we'll be but pets to the machines (like those in Iain M. Banks' culture novels), but real life won't be a communist
26.
▲
by
kristianp
1mo ago
There are some posts on https://www.cpu-world.com/forum/viewtopic.php?p=330329 , however the images are not visible. I tried registering, but it says "bot registration detected".
27.
▲
by
kristianp
1mo ago
The "incredibly expensive Intel Pentium" image link 404s.
28.
▲
Nvidia backing $105B in financing for OpenAI data center in Ohio
(cnbc.com)
3 points
by
kristianp
1mo ago
|
0 comments
29.
▲
by
kristianp
1mo ago
Enjoying the "DOS" font of the code examples. The wikipedia page on the Pentium has multiple references to an Intel publication called "Solutions", May/June 1993. It would be interesting to see that, but can't
30.
▲
by
kristianp
1mo ago
Does it get loud and hot with a smaller MOE model like this?
More ›