Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
notnullorvoid
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
notnullorvoid
11d ago
They are more than a accessory. You wouldn't say someone who funds a contract killing is merely an accessory to murder. VCs should likewise be considered as principal offenders, unless there is proof that the company was not transparen
2.
▲
by
notnullorvoid
16d ago
Yes they are quite good, but are not able to run on a 16GB RX 9070. Quantized Qwen 3.8 Flash Next could maybe run eventually on that card with a highly optimized inference engine that dynamically caches the hottest layer experts. Even then
3.
▲
by
notnullorvoid
22d ago
Traffic jams can be a feature, if it's the game Prototype where you could barrel through cars like a freight train.
4.
▲
by
notnullorvoid
22d ago
Qwen 3.8 27B is so far the closest I have run, though it suffers on speed compared to Gemma4 26B MoE model which I still use. Neither are going to match Opus or Sol though, but they can be as fast or faster depending on what you are using t
5.
▲
by
notnullorvoid
22d ago
I haven't yet, though plan to when this model is released. The models that I've been daily driving (Gemma 4 26B, Qwen 3.8 27B) have fit nicely on my 3090. I think FreeToken only offers a perf increase for MoE models that you can&#
6.
▲
by
notnullorvoid
22d ago
> It is still slow, a lot slower than what you are used to with claude and co. That really depends on the model, I run a few models locally. All at speeds comparable to or faster than Opus. In general we haven't reached the ceiling
7.
▲
by
notnullorvoid
22d ago
True, but you can link them up over thunderbolt or Ethernet. If your goal is to run local LLMs, not all weights need to live on the same computer. You can segment the workload by layers and pass the activations along the lower bandwidth int
8.
▲
by
notnullorvoid
22d ago
It will be interesting to see the intersection of this with inference engines like FreeToken which improve distribution of work for MoE models across CPU/RAM and GPU/VRAM. If all it takes for a competitive model to run locally at
9.
▲
by
notnullorvoid
22d ago
It needs to be a little less than double the overall price for 256GB option, otherwise it's better to get 2 256GB and link them.
10.
▲
by
notnullorvoid
23d ago
And you'd be okay with Anthropic being the ones that determine what "significantly large" means?
11.
▲
by
notnullorvoid
23d ago
Yes I want so badly to work at Antropic. - 50% or more of my compensation locked up with what I believe has a less than 5% chance of keeping significant value - Work with engineers who believe there job now is to vibe out a code editor, and
12.
▲
by
notnullorvoid
23d ago
No that is not inconceivable. What is inconceivable is believing that group of individuals is not a small minority, and the question will filter everyone else out.
13.
▲
by
notnullorvoid
23d ago
In my experience hype driven companies are very resistant to that even when equity/cash ratio offered is more sane.
14.
▲
by
notnullorvoid
23d ago
It would be a decent interview question if such a large portion of compensation did not come from equity. Otherwise it's basically the same as asking how someone would feel if all the sudden they lost 50% or more of their wealth. No on
15.
▲
by
notnullorvoid
24d ago
It seems like at every turn the word private is avoided and replaced with "non-public". I understand that the system probably is sound, the data just isn't encrypted, which I wouldn't necessarily expect. Maybe it's
16.
▲
by
notnullorvoid
1mo ago
Try it out with your own prompt, I suspect you'll be surprised.
17.
▲
by
notnullorvoid
1mo ago
I've done a few variations, I've been impressed with all of them. My favourite so far has been "Generate an SVG of a turtle flying a kite", result: https://imgur.com/a/bdKJPV4 . Some will say conflat
18.
▲
by
notnullorvoid
1mo ago
Congrats on the release! Nice to see transitions API go away, though I am not entirely convinced of the async model. It's capability is impressive, yet I do not like hiding what values are async, and throwing to await like React has ne
19.
▲
by
notnullorvoid
1mo ago
Yes latency, and the usual preference of ownership over rentership. Similarly their are benefits to running local AI too, like data privacy and control.
20.
▲
by
notnullorvoid
1mo ago
The gains wouldn't be "free lunch", it's the result of time and effort researching optimal design and architecture. Even if the idea of "no free lunch" was taken liberally discounting the cost of research, it w
21.
▲
by
notnullorvoid
1mo ago
If the brain does rely on quantum effects, it's still possible the quantum effects in use are able to be simulated efficiently on a classical computer. For example if it's a matter of signal transfer rather than quantum computatio
22.
▲
by
notnullorvoid
1mo ago
You can find it today for gaming. Even despite the outlandish rise in hardware costs, there is very little demand for cloud gaming.
23.
▲
by
notnullorvoid
1mo ago
It may not make financial sense for someone retired, not into tech, and/or data privacy to host their own LLMs. However if usage of AI in day to day lives continues to increase, I think it will eventually make sense for the majority. M
24.
▲
by
notnullorvoid
1mo ago
We've barely even started on optimizations like advanced language aware grammars, and specialization routing (dynamically loading fine tunes or seperate weights for specific tasks or languages).
25.
▲
by
notnullorvoid
1mo ago
They should've done $0.01 off. Effectively the same discount, while being more outrageous.
26.
▲
by
notnullorvoid
1mo ago
No. I think it'll mostly cause a dip in the average intelligence of knowledge workers. We will recover when we realize the valuable workers in a field are the ones who want to be there and actively increase their knowledge rather than
27.
▲
by
notnullorvoid
1mo ago
> has me questioning how much longer I want to remain in the industry. It's not a bad time to take a break from it. Keep yourself from cognitive decline, then in 2 or 3 years when the industry has come to its senses you'll be a
28.
▲
by
notnullorvoid
1mo ago
It'll only get worse unless companies start having strict policies around it, making repeat offenses fireable. Juniors should have a little more wiggle room, but there's no excuse for senior engineers to participate in such carele
29.
▲
by
notnullorvoid
1mo ago
> The problem is, your employer doesn’t care whether your brain is creating new neurons and connections. They care about productivity and profit. It was/is a struggle to get them to recognize and balance tech dept, we must now push
30.
▲
by
notnullorvoid
1mo ago
> that makes me wonder if the trillion dollar valuations for OpenAI and Claude are even justified. They aren't, not even if we forget about the capable Chinese models. I suspect Anthropic will implode soon when employees are unable
More ›