Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Phemist
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
Phemist
4d ago
Also debunked in the Gamer Nexus video. The TV will scan for public wifis and connect. Watching the video, it also seemed like a "missed opportunity" for LG to not create a mesh network of LG devices that would help eachother conn
2.
▲
by
Phemist
8d ago
It definitely is solvable though. Data versioning is a thing and it can work quite transparently to the mutations done on the data. To not know who made and who approved a set of mutations on data can easily become equally as mind-blowingly
3.
▲
by
Phemist
10d ago
Yeah, vanilla MacOS is super focused on the set of "all windows of a given Application" as the useful "unit of work". Whereas in real work, you actually handle individual windows of a set of applications (e.g. firefox wi
4.
▲
by
Phemist
12d ago
My knowledge is a few years outdated by now, but I remember digging into this and realizing that most of the chinese open-source libs were license-washing software. E.g. PaddleOCR is licensed under Apache 2.0, a very permissive license, how
5.
▲
by
Phemist
12d ago
Yes, but expect those prison sentences will be more popular than you might expect. The current admin expends an insane amount of energy pumping the size of their supposed fanbase.
6.
▲
by
Phemist
13d ago
What is the intuition. Higher quality turns due to more reasoning results in significantly fewer turns taken?
7.
▲
by
Phemist
14d ago
Potentially because they run from the same colossus DCs?
8.
▲
by
Phemist
17d ago
This default-to-auto-mode and the misleading marketing is begging for a class action once damages accumulate. Especially considering the Auto Mode even can actively prevent the clean-up!
9.
▲
by
Phemist
19d ago
With every newly released open weight model, the clock on the issues you describe is reset. I can see a marketplace arising for paid updates to common lines of open weight models, which will incentivize those with the hardware to train to f
10.
▲
by
Phemist
20d ago
> They didn't leave my computer I guess, I just showed me the output of some `ls` commands Not sure exactly how Zed works, but wouldn't the results of the `ls` tool call be fed back into claude?
11.
▲
by
Phemist
21d ago
I am not arguing that there are perhaps other models that can run at the same quality, can coordinate between the different modalities, but are way less power hungry. My point is exactly about the comparison between the token output of the
12.
▲
by
Phemist
21d ago
They are already working on it. https://huggingface.co/unsloth/Qwen3.8-Flash-Next-GGUF https://unsloth.ai/docs/models/qwen3.8-next > You will need at least 75 GB of RAM or unified memory t
13.
▲
by
Phemist
22d ago
They are not. If the robot speech is a tool call, then for a fair comparison we need to take the tool call scaffolding (and probably the reasoning too) into account. So rather than a sentence of 10 tokens worth of speech being the output, t
14.
▲
by
Phemist
22d ago
It seems that e.g. OpenAI is mostly power constrained. To go to the middle of nowhere would mean also basically no power grid to speak off?
15.
▲
by
Phemist
22d ago
> I don’t see how tokens can’t produce speech or track metabolic needs. It probably could, but the point is this would require additional tokens, blowing up the comparison. The token output of LLMs and "token output" of speech
16.
▲
by
Phemist
22d ago
The 20W number includes EVERYTHING else the brain does. The chips/models are literally only producing tokens. Let's see an LLM drive a robot harness and have the robot produce speech, as well as move through 3D space, keep track o
17.
▲
by
Phemist
23d ago
RAM and SSD in apple gear has always been way over-priced. There was a short blessed period in March where the M5 Max macbook pro was out, but the general 30% price hike had not yet happened. In this period, given the insane inflated RAM pr
18.
▲
by
Phemist
23d ago
Single turn set-ups may work like this. You control the thing you input, the LLM outputs something and then nothing happens further for that specific context. (Simple question/answer style interactions..) (Multi-turn) tool calling set-
19.
▲
by
Phemist
25d ago
Maybe detecting watermarked text is a skill to be attained. Not allowing proper feedback and training will not allow people to notice the difference on time, thus ruining the data? Best practice is to allow a number (scaled based on complex
20.
▲
by
Phemist
26d ago
I had Kagi configured to rewrite links old.reddit.com. Looks like I need to update it. Edit: Actually I think this was part of their documentation to explain the intended use of that feature. Too bad this will now no longer work! ( https:&#
21.
▲
by
Phemist
29d ago
On the same day that OpenAI cuts per token cost by 50% on GPT5.6 Sol. :')
22.
▲
by
Phemist
1mo ago
Ok. I dont have a good feeling for the actual completion distributions. The noise sounds problematic. I can imagine it relates to the size of the context used for hashing. You want this as long as possible so that the entropy is higher, but
23.
▲
by
Phemist
1mo ago
I am going off the explanation in the declaude page (and related papers). But I see now anthropic mentions Aaronson's distortion-free watermarking. Random watermarking functions colour the tokens based on (small) contexts and a secret
24.
▲
by
Phemist
1mo ago
1. As watermarked text is added to the training data, watermark-related tokens will be associated more with AI outputs and thus lower quality outputs which will hasten model collapse. Especially because every provider has its own secret key
25.
▲
by
Phemist
1mo ago
I am interested however in why fine-tuning on reasoning traces of a frontier model is such an effective way of improving an (open-weight) base model. See e.g. https://huggingface.co/hesamation/Qwen3.6-35B-A3B-Claude-4.6
26.
▲
by
Phemist
1mo ago
Of course we are all guessing, but both things can be true: they don't want you to hit the cache, because cache writes are more profitable than cache reads, and they are supply limited on the compute. Pre-filling input tokens is very d
27.
▲
by
Phemist
1mo ago
Incentives to optimize cache-usage are only aligned between anthropic and you, dear user, when there is not enough compute available to serve the tokens fresh. Apparently enough compute is now available that frugal usage is no longer a requ
28.
▲
by
Phemist
1mo ago
Did you try the claude reasoning traces finetune for qwen3.6? I find that it works muuch better. I assume the same 3.8 finetune will be released at some pointas well. Edit: link - https://huggingface.co/rico03/Qwen3.6-2
29.
▲
by
Phemist
1mo ago
This kills the isolation.
30.
▲
by
Phemist
1mo ago
You let the smarter model explore the traces and figure out where the current harness' bottlenecks are for the current LLM. Then you can adjust prompts or tools to fix those.
More ›