Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
richardfey
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
10 ms
·
1.
▲
by
richardfey
25d ago
Yes; or something which has a similar effect.
2.
▲
by
richardfey
25d ago
It feels lke they replace older models with "optimised" versions which are cheaper to run, but keep the same name.
3.
▲
by
richardfey
1mo ago
For pixel owners: the Google photos app cannot be uninstalled, but you can disable it using adb.
4.
▲
by
richardfey
1mo ago
> "Instead of improving an agent against a fixed test, we let the evaluation evolve alongside the agent" This quote should have been highlighted earlier in the article.
5.
▲
by
richardfey
1mo ago
It defines and introduces a lot of concepts/acronyms in the thinking blocks which we normally don't read.
6.
▲
by
richardfey
1mo ago
> If the model was designed specifically to quantize down to 1.58b, then it's different. > AFAIK, there's no large models designed for this yet. Isn't BitNet b1.58 2B4T what you are looking for? (haven't tried it m
7.
▲
by
richardfey
1mo ago
> Quis custodiet ipsos custodes? is a question that is roughly 2000 years old. Still no good answer. I feel like we are moving towards a very specific answer, without much thought: AI.
8.
▲
by
richardfey
1mo ago
Did anyone calculate this for OpenAI?
9.
▲
by
richardfey
1mo ago
Looking forward to giving this a try with llama.cpp. I’m watching the open-weights competition with high expectations.
10.
▲
by
richardfey
1mo ago
I think that would spoil a lot of the fun. I would favour visual and audio feedback loops to interactively learn how things work under the hood.
11.
▲
by
richardfey
1mo ago
Sounds like a great way to inflate your success metrics for AI queries
12.
▲
by
richardfey
1mo ago
I disagree. Whether something is manipulation depends on whether you are trying to change someone’s opinion or behavior, not on whether the manipulator has “good” or “bad” intentions, since those judgments are not objectively universal.
13.
▲
by
richardfey
1mo ago
> Part of this can be good (you talk about what they care about, where 90% of broadcast messaging might not apply) and part of it can be bad (manipulation.) Side note: it's manipulation either ways because you chose what to talk abo
14.
▲
by
richardfey
2mo ago
Serious question: how do we verify claims like these on the effectiveness of a harness?
15.
▲
by
richardfey
2mo ago
Has anyone tried Kimi K3 against gpt-5.6-sol on real projects?
16.
▲
by
richardfey
2mo ago
Exactly; this is a no-go for me, I will wait for an independent provider to sell the service, which is possible thanks to the open weights.
17.
▲
by
richardfey
2mo ago
I understand you're trying to be funny, but my point is that with novel technology there are novel ways to claim innocence in courts because of the legislative void.
18.
▲
by
richardfey
2mo ago
They could use an agent to summarise the source material, and then train models on those summaries, and claim that some sort of clean-room training has happened?
19.
▲
by
richardfey
2mo ago
> Or are you referring to the claims from Meta, Google, and Adobe -- which failed to hold up under independent evaluation. This. > However, he demonstrated a clear lack of understanding regarding what makes the bits "independent&
20.
▲
by
richardfey
2mo ago
I can't find on their website some indication of what kind of usage I can get out it, otherwise I'd be interested.
21.
▲
by
richardfey
2mo ago
This is a great statistical analysis and it was a pleasure to read, but I wasn't expecting the claims to be so poorly supported. There's also a reply from one of the Meta authors there, worth checking out.
22.
▲
by
richardfey
2mo ago
I remember hearing this perspective when I first started in the software industry, and I agreed with it for quite some time. But frankly, we’ve never been further from it.
23.
▲
by
richardfey
3mo ago
The more I read into it, the more pain memory flashbacks I got. Bravo
24.
▲
by
richardfey
3mo ago
I'm trying it right now for a side project of mine, compared to Opus it is effectively better at following instructions and somehow has better "depth" when reasoning on complex tasks. However, if it will not be part of subscr
25.
▲
by
richardfey
3mo ago
I don't know what I am doing right, or wrong, but I have access to claude and codex and I find myself giving the more serious work to codex recently. I tend to trust it more. I might try again Fable when it's back, but this Sonnet
26.
▲
by
richardfey
3mo ago
How did you give LLMs tool use?
27.
▲
by
richardfey
3mo ago
I could spot numerous bugs in code written recently and less recently, by me or colleagues. I was not angry but grateful and I knew there was no way back!
28.
▲
by
richardfey
3mo ago
You mean bad because they could have used a larger memory module and thus higher resolution sound samples?
29.
▲
by
richardfey
4mo ago
> To do this, your device is shouting to the world a ton of your personal information in something called a probe packet. A probe packet contains the MAC address as well as the list of all the past Wi-fi networks that your device has tri
30.
▲
by
richardfey
4mo ago
I'm going to give a try to piclaw as I want to get into the mindset of author, thanks!
More ›