Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ch_sm
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
ch_sm
5d ago
I can imagine that the new vapour chamber they’re introducing helps or is necessary to run this chip at high speeds.
2.
▲
by
ch_sm
7d ago
I don’t know if i want to expose my child to auto-censorship (or online gaming anyway).
3.
▲
by
ch_sm
10d ago
unpopular opinion apparently, but i don‘t understand the finder hate. as a real-life, non-software-developer user, i couldn‘t care less about library folders, root folders, or a unix-tradition file system at all. i want my folders with my d
4.
▲
by
ch_sm
15d ago
I’m not an expert, but my understanding is that MTPs are smaller LLMs fine-tuned to "mimic" / predict a specific model’s response. It’s possible that the MTP you’re using isn’t trained well enough on Qwen 3.8. What accept rat
5.
▲
by
ch_sm
15d ago
It depends on your use case, but the smaller Gemma 4 models or qwen3.6:9b would probably run OK on that. I recommend trying it, even just for fun. It‘s easy with omlx.
6.
▲
by
ch_sm
15d ago
I‘ve been considering applying at WikiMedia for a while now, but [my country] is not in the listed countries. Is there a chance you‘ll expand this list in the future?
7.
▲
by
ch_sm
18d ago
beautiful!
8.
▲
by
ch_sm
18d ago
great tip, thanks!
9.
▲
by
ch_sm
19d ago
To be fair, that is what it says on their website.
10.
▲
by
ch_sm
21d ago
that is a very fair point i haven‘t seen being made thus far. it took a long time for gemma4 and qwen 3.6 to reach the speed and reliability that they have now, and many (including myself) made assumptions about their viability on day one.
11.
▲
by
ch_sm
27d ago
prooof-read by Qwen
12.
▲
by
ch_sm
27d ago
> I'm unable to find a local model that comes close to the effectiveness of GPT models in Codex I agree with you there, the local models are not as capable as frontier remote LLMs. If you‘re used to letting fable run for an hour to
13.
▲
by
ch_sm
28d ago
> I tried running a smaller model locally, and it's not usable for me. If you have the hardware, a MacBook Pro for Qwen 3.6 35B A3B and Gemma 4 26B A4B for example, they are absolutely usable, both in terms of speed and quality. Ane
14.
▲
by
ch_sm
1mo ago
Right, my bad; I know they claim to not store _all_ prompts and responses, that was a bit cynical/hyperbolical of me. However, regulation could force them to do it — with all the downsides that come with that. Your privacy argument, on
15.
▲
by
ch_sm
1mo ago
here‘s what i don‘t get about this whole discussion. AI companies already store all prompts and responses for future training. just make an API that returns the string distance between a previously generated paragraph and the query? that wo
16.
▲
by
ch_sm
1mo ago
yes. that‘s the first half of the cause (the other being that he‘s a shameless scammer).
17.
▲
by
ch_sm
1mo ago
While these arguments are valid, you’re missing the forrest for the trees. Renewable energy has better long-term ROI because it doesn’t require constant purchase of fuel. Renewable energy plants also produce far, far less CO2 that goes into
18.
▲
by
ch_sm
2mo ago
> bought it off CraigsList for $10, handed off in Times Square haha that‘s a nice nostalgic image. I remember installing it on a friend’s laptop with my DVD in a café, to „fix her macbook“.
19.
▲
by
ch_sm
2mo ago
you can try ollama, omlx or llama.cpp for instance to download a model and get an inference server running locally. They expose „open ai compatible“ endpoints, so you can configure almost any harness to use them.
20.
▲
ChatGPT, Roblox to fall under strictest EU rules for platforms
(bloomberg.com)
74 points
by
ch_sm
2mo ago
|
54 comments
21.
▲
by
ch_sm
2mo ago
oh, it‘s the CEO of a well known AI company educating us all – and all he wants is us to hear his hunch on open weight models?! Amazing, please help us understand the situation a bit better, thanks
22.
▲
by
ch_sm
2mo ago
ah finally. Compiz for react.
23.
▲
by
ch_sm
2mo ago
In my experience, yes. A bit more reliable than gemma for me. I mostly use A3B (35B, mix of experts) though, because it‘s faster, and in the same ballpark intelligence wise as the dense 27B, so it’s the sweetspot for me. I want to try cohe
24.
▲
by
ch_sm
2mo ago
i see! would you mind sharing your model string?
25.
▲
by
ch_sm
2mo ago
WHAT. that‘s amazing. thanks so much for sharing, i‘ve been looking for ways to speed a3b up for days. It‘s 11pm here but i‘ll try this right now
26.
▲
by
ch_sm
2mo ago
Nice, I tried it too with oMLX — agreed, it seems very capable for coding! Was slightly underwhelmed by performance though. I got about ~24 t/s on the ternary version on my M2 Max 64. That’s quite a bit slower than Qwen A3B 35B (4 bit
27.
▲
by
ch_sm
2mo ago
https://archive.is/sTtAN
28.
▲
by
ch_sm
2mo ago
that‘s … not really how it works. Agents get your code into their context via tool calls, not by uploading the entire file to a server „where the weights live and thus the code has to be too“. Small but crucial difference. Aside from that:
29.
▲
by
ch_sm
2mo ago
the point the article makes is good (albeit not new). the style sounds very LLM to me.
30.
▲
by
ch_sm
3mo ago
>The way royalties get assigned is based on a percentage of your listening versus your monthly payment. For example, spend an entire month listening to Taylor Swift’s new album, she gets the entire royalty share. But if you listen to the
More ›