Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kmike84
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
kmike84
6d ago
Unfortunately no, it's 200B + 552B. It's not as bad as it sounds though, because most of 552B is in 4bit natively.
2.
▲
by
kmike84
8d ago
Yeah, makes sense; e2e is different and valuable, KLD is not a replacement. As for KLD, have you tried it on something which is even closer to e2e task, like agentic traces from https://huggingface.co/datasets/nvidia&#x
3.
▲
by
kmike84
8d ago
Measuring quality e2e definitely makes sense. But I think there is a bit more to this: > Measuring token prediction differences (KL-divergence, top-1 predictions) is easy, but it does not tell us whether the model gets worse at solving t
4.
▲
by
kmike84
13d ago
The big news is that it's going to be open weight - https://x.com/finkd/status/2095232032896946311 .
5.
▲
by
kmike84
14d ago
Pass. When articles keep mentioning models like DeepSeek R1, or Llama 3.1, or Qwen3 32B, it is a pretty robust indicator of AI slop. LLMs love to suggest DeepSeek R1, etc. - training data cut-off? No person with real practical experience an
6.
▲
by
kmike84
17d ago
So, open models will be better on this benchmark, which is deserved
7.
▲
by
kmike84
1mo ago
Heh, a good point.
8.
▲
by
kmike84
1mo ago
> Which harness for the benchmark ? pi, with a plugin to do web search / web fetch. > You have previously commented on using OC/GLM. R u going to stock with it? For personal use - probably yes, z.ai + kimi + opencode go subs
9.
▲
by
kmike84
1mo ago
DeepSeek needs more RAM for weights, Qwen requires more compute. Also, DeepSeek's KV cache requires less RAM than Qwen's. In concurrent situations (on servers) you load model weights once, but you have different context in each pa
10.
▲
by
kmike84
1mo ago
I have an internal automated benchmark, which roughly follows my workflow, and I've been testing various models on it, local and cloud. Qwen 3.8 27B did awesome. Its understanding is correct, research is better than e.g. glm's (an
11.
▲
by
kmike84
2mo ago
omlx is quite similar to LM Studio, so there are "options like LM Studio" which are open source
12.
▲
by
kmike84
3mo ago
You care if you run it on a laptop. It's getting hot, fans are spinning, and you may want to use laptop for other things while the agent is working.
13.
▲
by
kmike84
3mo ago
I do have this experience. I've used Claude Code (with Opus mostly), and then switched to opencode (mostly with Kimi 2.6) for my personal projects; it's based on a couple months of use. Claude Code is better. But Opencode + kimi 2
14.
▲
by
kmike84
3mo ago
Go for it! It's very satisfying :)
15.
▲
by
kmike84
3mo ago
* plugin for Logic Pro to A/B mix with reference tracks, with ai-based stem splitter (e.g. isolate vocals in ref track, and compare with your vocal track) * plugin for Logic Pro to simulate how a mix will sound on my macbook and phone
16.
▲
by
kmike84
8mo ago
I'm not sure I want AI to touch me emotionally. It feels insincere and manipulative, especially when I don't know upfront if the content (music, video, text) is from another human being or from AI. AI will become good enough to wr
17.
▲
by
kmike84
1y ago
> whereas the best engines average 99.something%? To compute accuracy, you compare the moves which are made during the game with the best moves suggested by the engine. So, the engine will evaluate itself 100%, given its settings are the
18.
▲
by
kmike84
2y ago
I think he/she is reacting mostly to this quote from the article, not to the main article topic: > I have a good answer: my job is to double our value-add capacity over the next three years. Essentially, to double our output without
19.
▲
by
kmike84
3y ago
The URL parsing in httpx is rfc3986, which is not the same as WHATWG URL living standard. rfc3986 may reject URLs which browsers accept, or it can handle them in a different way. WHATWG URL living standard tries to put on paper the real bro
20.
▲
by
kmike84
3y ago
A great initiative! We need a better URL parser in Scrapy, for similar reasons. Speed and WHATWG standard compliance (i.e. do the same as web browsers) are the main things. It's possible to get closer to WHATWG behavior by using urllib
21.
▲
by
kmike84
4y ago
Exit Through the Gift Shop - an amusing documentary about somebody trying to find Banksy (a street artist), and much more, supposingly directed by Banksy himself. There is some debate if it is documentary or not (the story is almost too goo
22.
▲
by
kmike84
4y ago
No.
23.
▲
by
kmike84
4y ago
The advice to use lru_cache is good. But there is an issue if lru_cache is used on methods, like in the example given in the article: 1. When lru_cache is used on a method, `self` is used as a part of cache key. That's good, because th
24.
▲
by
kmike84
4y ago
Not sure about the autofocus advice; I'm pretty happy with manual focus. It requires static camera placement, and fixed distance to the person, but isn't this happening anyways? Are people really walking around the room or moving
25.
▲
by
kmike84
4y ago
Hm, I haven't noticed any increased latency when using a DSLR as a webcam.
26.
▲
by
kmike84
4y ago
As I understand, the drivers (webcam utility? not sure) are built for x86. For some reason they don't work in apps which are built for M1, so the camera only works if an app which needs a video is running in emulation mode. So, if you
27.
▲
by
kmike84
4y ago
Is it such a big issue? My Canon DSLR turns off every 30min, but that's only for a couple of seconds, it then turns back on. On a positive side, it's now easy to notice when 30min or 1hr meeting is running over, it's a nice r
28.
▲
by
kmike84
5y ago
That's interesting.. We're working on web data extraction in Zyte (former Scrapinghub); we have an Automatic Extraction product ( https://docs.zyte.com/automatic-extraction-get-started.html ) which combines ML and m
29.
▲
by
kmike84
6y ago
I'm unsure about the advice of sticking to 1440p at 27". I have a non-retina imac 27 (1440p), external LG 27" 4K USB-C monitor and a macbook pro 13 with a real "retina", and use them all regularly. For my eyes, scal
30.
▲
by
kmike84
6y ago
My son (4.5yo) became a huge fan of Moomin tales recently. Books, audio books, cartoons; he likes them more than super heroes these days. These tales are not only nostalgia material fo adults, these are still great children stories.
More ›