Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
syntaxing
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Shapelearn Qwen 3.8 27B (13.1 GB VRAM)
(byteshape.com)
4 points
by
syntaxing
3h ago
|
0 comments
2.
▲
by
syntaxing
5d ago
Wow thanks for the link. I have zoom on my personal laptop which isnt ideal. I always wanted to run it sandboxed
3.
▲
by
syntaxing
5d ago
Has anyone have good success using AI generated CAD parts? I’ve been trying but it’s always 95% there, but with all hardware, you need 100% right. It’s often quicker and cheaper for me to do it by hand (but I was a mechanical design enginee
4.
▲
by
syntaxing
8d ago
Surprised no one is talking about it but the 0.1 version bumped the parameters from 284B to 552B but “more efficient”, particularly kv cache usage
5.
▲
by
syntaxing
8d ago
I wonder if that’s why 3.8 got so much better? Mixing the reasoning traces from both sides seems to be effective.
6.
▲
by
syntaxing
8d ago
I’m surprised they allow open lid drinks in the lab. One wrong bump and poof 300K easy.
7.
▲
by
syntaxing
9d ago
> 23 agents total. This hit a bit too close to home. Sol has the same issue, spawns a lot of agents for no good reasons (besides burning tokens).
8.
▲
by
syntaxing
10d ago
I’m more curious how each 4 bit quant compares. It seems like NVFP4 outperforms Q4_K_M in terms of speed and top 1 but is only good for expensive Nvidia cards
9.
▲
by
syntaxing
13d ago
Qwen 3.8 27B is the real deal BUT remember to use froggeric template and/or medium reasoning. https://huggingface.co/froggeric/Qwen-Fixed-Chat-Templates
10.
▲
by
syntaxing
14d ago
I swear, Qwen 3.8 27B @ Q8 is smarter than Sonnet 5 most of the time. Why wouldn’t corporate America self host at this point, especially with better options like Deepseek Flash and GLM 5.3 flash that’s a middle ground between Sonnet and Opu
11.
▲
by
syntaxing
18d ago
I’m honestly surprised this is better benchmark wise than the text only model. I figured the addition of vision would take away from some of the text capabilities.
12.
▲
by
syntaxing
18d ago
How do people bypass captcha or robot checks? All I wanted is a price aggregator but it always gets blocked by major retailers.
13.
▲
by
syntaxing
21d ago
If you get a children’s card, they print your kids name right on it. It’s a nice little souvenir to keep.
14.
▲
by
syntaxing
23d ago
Ironically, our administration pushing for ban of the AI chips to China is forcing them to make smaller and more efficient models which seems like a requirement for running on Chinese chips. I wouldn’t be surprised this model was tailored t
15.
▲
by
syntaxing
23d ago
I’m more curious on the size. If it’s smaller than or equal size to GLM 5.3, this would be a crazy good model. If it’s closer to deepseek pro, it would be a good model. If it’s near Kimi K3, I think it’s competitive but nothing particularly
16.
▲
by
syntaxing
24d ago
With MTP? I get 25-30 TPS on a strix halo. 50+ on a M5 max should very doable. Dflash (2) will push your TG even further
17.
▲
by
syntaxing
24d ago
Really looking forward to this, 27B is a struggle with a strix halo and Laguna 2.1 can do stupid things for tooling calls.
18.
▲
by
syntaxing
25d ago
Agreed, and you can write cool infra agents to do stuff for you that runs during off hours like nightly tests and triage.
19.
▲
by
syntaxing
25d ago
GLM series has made it very practical to self host. If the new update for Deepseek flash holds up, I think it would be silly for some companies to not self host.
20.
▲
by
syntaxing
27d ago
It’s wild how broken search has become, LLM made it worse but it was already downhill prior to that. Even Reddit forces you login to search within a subreddit and displays the annoying “Best place on internet” overlay after staying on the s
21.
▲
by
syntaxing
28d ago
Overkill is the point. Same thing for iPads and iPhones. It gives a luxury feeling. That and Chinese manufacturing have made milling a very cost effective manufacturing method.
22.
▲
by
syntaxing
28d ago
iPhones are iPads have a milled CNC “unibody” as well… Being able to dictate where to have more thickness helps a ton on the luxury feel. Most of the chips are recycled anyway into new billets.
23.
▲
by
syntaxing
29d ago
I’m kinda surprised Apple doesn’t do something like this. I would imagine it’s a lot easier with phones with LiDAR
24.
▲
by
syntaxing
1mo ago
I think the fun takeaway from this is that GPT 5.4 is probably 45B active parameters and GPT 5.6 Sol is closer to 50B.
25.
▲
by
syntaxing
1mo ago
Thanks for the update! I wonder if a Blackwell GPU would be noticeably faster. Which vendor did you end up using? I want to get a gigabyte one but thats be OOS for months.
26.
▲
by
syntaxing
1mo ago
I think the hardest part is how to define “waste”. Are you “wasting” your time if you go to a school event for your kids during the weekday? Is it a waste of time walking your kids to the bus? Is it a waste of time reading to my kids before
27.
▲
by
syntaxing
1mo ago
What speed do you get on this setup? Im tempted to use the same GPU.
28.
▲
by
syntaxing
1mo ago
The issue with “AI” is how toxic it feels. Feels grimy like social media. Just like social media, I haven’t introduced any of it to my young kids nor do I plan to until they’re teens. At most I give them access to local AI through home assi
29.
▲
by
syntaxing
1mo ago
Would I be surprised there’s bench maxing happening? Yes. But some users also use Q4 quantized and complain how dumb local models are.
30.
▲
by
syntaxing
1mo ago
Is there a reason why so many of these agent harness are written in node.js?
More ›