Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
zaptrem
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
zaptrem
13d ago
Per hour per 8x node (so $1-1.25 per hour per gpu)
2.
▲
by
zaptrem
13d ago
Per hour per 8x node (so $1-1.25 per hour per gpu)
3.
▲
by
zaptrem
14d ago
Not the cheapest. Many providers have spot H100s for $8-12, and H200s for slightly less than the above prices
4.
▲
by
zaptrem
2mo ago
Can you explain how the above event doesn't count as evidence alignment is an actual risk?
5.
▲
by
zaptrem
2mo ago
Given these models could not have been trained in the first place if they had to license every line of random fan fiction on the internet, I think distillation also being fair game is a tradeoff everyone should be willing to take (unless th
6.
▲
by
zaptrem
2mo ago
This is my number one complaint about the M-series MBP line. Especially true of the cutout in the middle that has points so sharp they can cut you if you accidentally scrape it with your hand.
7.
▲
by
zaptrem
2mo ago
Why haven’t we seen any queues or the like over the past week then? If it’s truly a capacity limitation why not just boot subscription users to a lower priority queue or limit usage to outside peak hours?
8.
▲
by
zaptrem
3mo ago
Should we require the destruction of the brains of those that watch pirated movies?
9.
▲
by
zaptrem
3mo ago
Needs more WebGL spinning rubik's cube
10.
▲
by
zaptrem
3mo ago
Can you include GPT 5.5 non-pro (extra high thinking I guess) in your comparison? GPT Pro is the "I am willing to torch cash for a sooometimes slighty better result" option, not the one people are actually expected to use daily. T
11.
▲
by
zaptrem
4mo ago
OOM on CUDA GPUs is relatively graceful (the process crashes). However, on macOS if torch MPS tries to allocate too much memory, the whole kernel will simply lock up and the only option is to reboot the computer. I have no idea why Apple do
12.
▲
by
zaptrem
4mo ago
Love me some JSD. Here is a problem most people don't consider with generative modeling (e.g., AI text, image, music, video models): basically all standard pre-training algorithms for generative models (i.e., cross entropy, basically a
13.
▲
by
zaptrem
4mo ago
V4-Pro is about 2.4× total params and 1.3× active params of V3.2.
14.
▲
by
zaptrem
4mo ago
Seems pretty clear, Claude and Codex were getting a lot of free publicity by instructing their models to do the same and MS wanted similar results. However, a bug caused this to be applied to all commits instead of all Copilot-influenced
15.
▲
by
zaptrem
5mo ago
I bumped from $20 -> $100 today but the Codex CLI lacking code rewind and "you can change files but ask me every time" mode from Claude Code is quite annoying. Sometimes I want to code, not vibe code lol.
16.
▲
by
zaptrem
5mo ago
Agreed, that’s why I specified end to end (I.e., text to waveform)
17.
▲
by
zaptrem
5mo ago
My point is you should consider creating truly undetectable audio end to end with AI to be effectively impossible for the foreseeable future (i.e., I would bet money it is still trivially detectable five years from now). It won't be de
18.
▲
by
zaptrem
5mo ago
I train music generation models. They are very trivial to detect. In fact, detecting them then training them to evade detection by the detection model is a big part of training them! But the detectors win instantly without some hardcore reg
19.
▲
by
zaptrem
5mo ago
What's your reasoning effort set to? Max now uses way more tokens and isn't suggested for most usecases. Even the new default (xhigh) uses more than the old default (medium).
20.
▲
by
zaptrem
6mo ago
YouTube et al's automated copyright systems put way too much trust in the hands of those making the claims.
21.
▲
by
zaptrem
6mo ago
Many of the games that actual kids spend time on are the purest expression of gaming slop (half-broken microtransaction gambling hell with schizophrenic flashing colors). Roblox and Fortnite's Islands system are both guilty of this. Th
22.
▲
by
zaptrem
6mo ago
In my experience, the Epic Games Store downloads faster, installs more efficiently, and launches games faster than Steam. The social features I actually use (i.e., add a friend, join them in a game) work fine. I'm not aware of any feat
23.
▲
by
zaptrem
6mo ago
I have Max 20x and they're still separate on 2.1.75.
24.
▲
by
zaptrem
8mo ago
Data centers don't do anything other than sit there and turn electricity into heat. They only emit nothing but heat (which could be useful to others in the building).
25.
▲
by
zaptrem
8mo ago
What did Epic do?
26.
▲
Wikipedia: Sandbox
(en.wikipedia.org)
93 points
by
zaptrem
8mo ago
|
37 comments
27.
▲
by
zaptrem
8mo ago
"Previous data from the trial reported that 107 participants received the mRNA vaccine and Keytruda treatment, while the remaining 50 only received Keytruda. At the two-year follow-up, 24 of the 107 (22 percent) who got the experimenta
28.
▲
by
zaptrem
8mo ago
See here for a truly random sample of human music: https://0xbeef.co.uk/random/soundcloud Thankfully, most of it doesn't reach your Spotify feed. I think most of it is garbage, but I'd fight for the right of
29.
▲
by
zaptrem
8mo ago
I’m a founder of one of these AI music companies and that noise you’re describing (it differs between co’s for us it’s loud vocals, for Suno it’s vocal aliasing/sandiness and mushy instrumentals, etc) is exactly why I think these songs
30.
▲
by
zaptrem
9mo ago
Not sure it’s a cultural thing since most of the copy coming out of DeepSeek has been pretty straightforward.
More ›