Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
timfsu
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
13 ms
·
1.
▲
by
timfsu
5d ago
I'm not an Anthropic fan, but it's worth asking - why is it always OpenAI?
2.
▲
by
timfsu
23d ago
Subscriptions are, and will likely remain, the best deal in town. Unfortunately, larger companies aren't able to do that. When your monthly token costs are in the $5-10k range, the local inference starts to look a lot more attractive
3.
▲
by
timfsu
1mo ago
Fascinating article. I daily catch LLMs in “lies” like: “I found the root cause of the bug” or “this approach is twice as fast”. It’s hard to say what causes this uninformed certainty - is it intrinsic to being trained on human writing, or
4.
▲
by
timfsu
1mo ago
More accurate to say restaurants are about people, not food. May be an apt analogy - chefs and cooks spend their time in the kitchen preparing food for a diner they don’t see.
5.
▲
by
timfsu
1mo ago
This makes me think - should we be using non-monospace fonts for coding agent chat to improve readability? Wonder if we can get Claude and Codex to do that automatically in supported terminals
6.
▲
by
timfsu
2mo ago
In my testing, models don't really know their own name - I would be very suspicious of --model actually revealing the real model name
7.
▲
by
timfsu
2mo ago
I had exactly this use case - in a grocery store in the Alps, no internet, fired up a local LLM on my phone to figure out what to cook and what to buy
8.
▲
by
timfsu
3mo ago
Wow, this is pretty scary. LLMs have made phishing attempts look so much more legit, and the damage they can do so much greater.
9.
▲
by
timfsu
3mo ago
This is big, but until we have policy clarity we can’t trust it. I’ve always migrated all of our agents to Pi SDK, we aren’t going back
10.
▲
by
timfsu
3mo ago
In contrast, I’m on the $200 max plan for Codex and I hit the 5 hour limit near daily at work. I typically am having it work on about 5 tasks an hour. I’ve never hit a 5 hour limit on Claude on the $200 plan but I have hit my weekly limit.
11.
▲
by
timfsu
3mo ago
Yeah it’s hard to call that cheating from a model. Maybe “disqualifying” is more accurate
12.
▲
by
timfsu
3mo ago
We saw this too with Gemini specifically. My favorite example - we built a hallucination detector (given the input, does the output make any false claims) in Gemini, and after the Seahawks won the Superbowl in February, it would consistentl
13.
▲
by
timfsu
4mo ago
Imagine you try two products you’ve never heard of. You prefer one over the other. Was it marketing? That’s what’s happening here. Marketing can get you to try something you wouldn’t have otherwise, and it may suggest benefits you’d get if
14.
▲
by
timfsu
4mo ago
This is dope! We basically built something very similar internally for our team and it's been a very natural and intuitive way to manage agents (as opposed to having a bunch of terminals to track). Not every task/conversation can
15.
▲
by
timfsu
4mo ago
Yes, but they expire in ways that unions don’t
16.
▲
by
timfsu
4mo ago
PSA - you can run something like `npm install -g npm@11.10.0; npm config set min-release-age=3` to update to a version of npm that supports the min-release-age configuration
17.
▲
by
timfsu
4mo ago
Pnpm - installs are faster to boot. We haven’t missed anything
18.
▲
by
timfsu
4mo ago
Did not know this was a thing, kudos to her for speaking out!
19.
▲
by
timfsu
5mo ago
This is a really neat bridge between “looks cool” and “feels like you’re there”. Inferring real life properties like lighting is a cool trick and just the beginning I’m sure. I’m excited to explore new and dynamic worlds and bring the AAA e
20.
▲
by
timfsu
5mo ago
I might call it a few different things, but spyware seems disingenuous until we learn that it’s actually spying…
21.
▲
by
timfsu
5mo ago
The question is - if the SOTA model disappear - do these follow-on models have the ability to improve themselves without distillation?
22.
▲
by
timfsu
7mo ago
I for one enjoyed this very long essay. It should've been a lot shorter, but you also didn't have to read it, it says right there in the title :)
23.
▲
by
timfsu
7mo ago
Love this idea. Working with AI assistants, I find it easier to push to GitHub to look at the changes, rather than use my IDE. I wish that wasn’t the case, so this makes a ton of sense.
24.
▲
by
timfsu
7mo ago
Fascinating article. Also extremely confusing (though probably not unexpected) that an important health researcher is named Nestle.
25.
▲
by
timfsu
7mo ago
I for one would love this - if it’s done well - except that it would presumably be locked in to OpenAI agents
26.
▲
by
timfsu
7mo ago
These narratives are so strange to me. It's not at all obvious why the arrival of AGI leads to human extinction or increasing our lifespan by thousands of years. Still, I like this line of thinking from this paper better than the doome
27.
▲
by
timfsu
8mo ago
I get the appeal, but it seems too early for one AI tool to be able to do "everything". I'm guessing there's some company out there trying to automate each of these tasks and dealing with the attendant complexity that co
28.
▲
by
timfsu
9mo ago
In my opinion, what you need is a person (or three), not a book :) Someone who relies on you, whatever the context, is some of the greatest motivation out there.
29.
▲
by
timfsu
9mo ago
Possibly best thing ever on Hacker News. There is something quite appealing about the simplicity of Boléro
30.
▲
by
timfsu
9mo ago
It's not clear how much ChatGPT is investing in the discovery part of the app store experience, so this seems like mostly a way for users to install apps they're already familiar with and use them from inside a chat. For now, it s
More ›