Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
oxcidized
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
oxcidized
10mo ago
While true, you do still technically get pop up ads occasionally when set that way.
2.
▲
by
oxcidized
10mo ago
Thanks for the info! Definitely much better than I expected.
3.
▲
by
oxcidized
10mo ago
Admittedly I've not tried running on system RAM often, but every time I've tried it's been abysmally slow (< 1 T/s) when I've tried on something like KoboldCPP or ollama. Is there any particular method required t
4.
▲
by
oxcidized
10mo ago
> That's small enough to run well on ~$5,000 of hardware... Honestly curious where you got this number. Unless you're talking about extremely small quants. Even just a Q4 quant gguf is ~130GB. Am I missing out on a relatively c
5.
▲
by
oxcidized
1y ago
Considering there were two generations (around 4.5 years) of top-tier consumer GPUs (3090/4090) stuck at 24GB VRAM max, and the current one (5090) "only" bumped it up to 32GB, I think you'll be waiting more than 5 years
6.
▲
by
oxcidized
2y ago
I believe double-jeopardy laws wouldn't allow this, but I could be wrong.