Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
wizee
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
wizee
16d ago
For software development tasks, Qwen 3.8 27B is genuinely excellent, but you need 32+ GB of VRAM to run it well with decent context, and enough memory bandwidth and compute to run it at a decent pace. With an M5 Max Mac Studio, you can do t
2.
▲
by
wizee
23d ago
Wow, super cool! It works quite well, looks nice, and really covers the whole city. I'd love to build something similar for my own home city.
3.
▲
by
wizee
3mo ago
Qwen 3.6 27B is quite good for agentic coding, and practical to run on consumer hardware. You need a system with either 32+ GB VRAM, or a unified memory system with 48+ GB VRAM and a decent integrated GPU. While not cheap, such a setup is s
4.
▲
by
wizee
5mo ago
They're comparing to Opus 4.6, not 4.5. It was Anthropic's best public model up until last week.
5.
▲
by
wizee
5mo ago
I run Qwen 3.5 122B-A10B on my MacBook Pro, and in my experience its capability level for programming and code comprehension tasks is roughly that of Claude Sonnet 3.7. Honestly I find that pretty amazing, having something with capability r
6.
▲
by
wizee
6mo ago
It’s worth also comparing Qwen 3.5, it’s a very strong model. Different benchmarks give different results, but in general Qwen 3.5, GLM 5, and Kimi K2.5 are all excellent models, and not too far from current SOTA models in capability/i
7.
▲
by
wizee
10mo ago
I disagree about the image quality at typical sizes - I find JPEG-XL is generally similar or better than AVIF at any reasonable compression ratios for web images. See this for example: https://tonisagrista.com/blog/2023
8.
▲
by
wizee
10mo ago
JPEG-XL provides the best migration path for image conversion from JPEG, with lossless recompression. It also supports arbitrary HDR bit depths (up to 32 bits per channel) unlike AVIF, and generally its HDR support is much better than AVIF.
9.
▲
by
wizee
11mo ago
Those ratios seem way off if you're referring to the M1 Max and not the base M1. If we use Geekbench CPU performance, the Ryzen 9 7945HX (which is from 2023) is around 12% faster single core and 32% faster multicore than the M1 Max (wh
10.
▲
by
wizee
11mo ago
153 GB/s is not bad at all for a base model; the Nvidia DGX Spark has only 273 GB/s memory bandwidth despite being billed as a desktop "AI supercomputer". Models like Qwen 3 30B-A3B and GPT-OSS 20B, both quite decent, sh
11.
▲
by
wizee
11mo ago
It's colleges they they have been clamping down on, as they were bringing in absolutely massive numbers of mostly Indian students who were coming mainly to work in low-end jobs and get out of India rather than to legitimately study. Th
12.
▲
by
wizee
1y ago
Mistral models are largely along the likes of what you were asking for. However, Grok (any version) absolutely is not a “don’t say gay” model, it talks about sexuality of all forms quite openly and fairly and is happy to produce creative co
13.
▲
by
wizee
1y ago
That’s the thing - old German compact luxury sedans from the 80s had the control feel, balance, and light weight you get from a Porsche, while also being practical family cars. There’s nothing like that made today. They were also decently s
14.
▲
by
wizee
1y ago
Aside from the Mazda MX-5 (which isn’t the most practical car), almost all small, simple, and light cars made today are econoboxes. They’re not designed to have the rich control feel, balanced and satisfying handling near the limits, respon
15.
▲
by
wizee
1y ago
In general, I agree. However, many older cars were small, light, simple, and raw - characteristics that have largely disappeared from modern cars. Automatic transmissions from the mid-90s and earlier generally sucked, though good old manual
16.
▲
by
wizee
1y ago
Qwen 3 Coder 30B-A3B has been pretty good for me with tool calling.
17.
▲
by
wizee
1y ago
While cloud models are of course faster and smarter, I've been pretty happy running Qwen 3 Coder 30B-A3B on my M4 Max MacBook Pro. It has been a pretty good coding assistant for me with Aider, and it's also great for throwing code
18.
▲
by
wizee
1y ago
Privacy, both personal and for corporate data protection is a major reason. Unlimited usage, allowing offline use, supporting open source, not worrying about a good model being taken down/discontinued or changed, and the freedom to use
19.
▲
by
wizee
1y ago
On my M4 Max MacBook Pro, with MLX, I get around 70-100 tokens/sec for Qwen 3 30B-A3B (depending on context size), and around 40-50 tokens/sec for Qwen 3 14B. Of course they’re not as good as the latest big models (open or closed)
20.
▲
by
wizee
1y ago
You should use flash attention with KV cache quantization. I routinely use Qwen 3 14B with the full 128k context and it fits in under 24 GB VRAM. On my Pixel 8, I've successfully used Qwen 3 4B with 8K context (again with flash attenti
21.
▲
by
wizee
1y ago
They just recently released the r1-0528 model which was a massive upgrade over the original R1 and is roughly on par with the current best proprietary western models. Let them take their time on R2.
22.
▲
by
wizee
1y ago
People supported families with single incomes with less than high school education for centuries before the 1950s.
23.
▲
by
wizee
1y ago
KDE 6 is quite stable in my experience and faster/more efficient than Gnome too.
24.
▲
by
wizee
1y ago
The excessive translucency makes contrast much worse and complex backgrounds poke through to distract from the test. Readability suffers severely. This is a terrible design direction. Kill it with fire.
25.
▲
by
wizee
1y ago
Is reading and memorizing a copyrighted text a breach of copyright? I.e. is creating a copy of the text in your mind a breach of copyright or fair fair use? Is it a breach of copyright if a digital “mind” similarly memorizes copyrighted tex
26.
▲
GLM-4-32B-0414: New MIT-licensed SOTA LLM from Zhipu AI
(huggingface.co)
3 points
by
wizee
1y ago
|
1 comments
27.
▲
by
wizee
1y ago
Demo available at https://chat.z.ai/ In my experiments, its non-reasoning variant is the strongest non-reasoning LLM in its size class, and outperforms much larger LLMs like Llama 4 Maverick in programming and mechanical en
28.
▲
by
wizee
2y ago
I tried it out locally and it's pretty good. It has a good writing tone and style, and a good level of knowledge, in line with expectations for its size. It's slightly worse than Mistral Large 2411 at STEM tasks, but very close in
29.
▲
by
wizee
2y ago
It seems to have been very benchmark-tuned for LMArena. In my own experiments, it was roughly in line with other comparably sized models for factual knowledge (like Mistral Small 3), and worse than Mistral Small 3 and Phi-4 at STEM problems
30.
▲
by
wizee
2y ago
In my own experiments with Gemma 3 27b, I was underwhelmed and rather disappointed. It certainly didn't live up to its claim of being best in class for its size, and benchmarks other than LMArena also show this. On various simple (high
More ›