Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jampa
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
jampa
12d ago
> I've never had an issue with Codex or Claude reading massive files Reading files isn't a problem they want to solve. The idea seems to be using a cheaper model to "scout" for the intended code, instead of an expensi
2.
▲
by
jampa
14d ago
I wasn't trying to be precise originally, I just tried to fit activities into "morning / evening" buckets. I did the whole itinerary with Opus first, but when I gave it to Gemini 3.7 Flash to review, it started correctin
3.
▲
by
jampa
14d ago
Eh that one is on me, if I think too much about my HN comment I end up deleting before posting it. I rely on the 1 min `delay` set in the profile page to fix before it goes live, but for some reason this time it was set to 0.
4.
▲
by
jampa
14d ago
I asked Claude to fix the grammar of my comment, and it changed "I am using 3.7 for" to "I've been using Claude 3.7", so they sneaked their own name on it.
5.
▲
by
jampa
14d ago
I've been using Gemini 3.7 for my personal trip planning app. Across multiple benchmarks, it ranks higher on everything I tried: - Real world knowledge (when a thing opens and closes, the geographic region, historical facts). It's
6.
▲
by
jampa
22d ago
I tried doing something like this Ox Alpha with Opus advisor, having it work layer by layer (specs -> rooms -> room graphs ...), but each deliverable ended up a mess. The curious thing is when I pointed out the flaws it fixed them qui
7.
▲
by
jampa
22d ago
Serious answer: no model ever gets close to writing an architectural floor plan that makes sense. They understand all the rules and best practices, they can (sometimes) spot a bad idea in a floor plan, they can describe a good floor plan. B
8.
▲
by
jampa
28d ago
Opus 5 feels like a downgrade from Opus 4.8 overall. It, along with Fable, really has a problem following instructions and staying in scope, and their prose keeps growing, both in explaining what it did and in writing multiline code comment
9.
▲
by
jampa
1mo ago
I used to do that, but the person could be bad at prompting too (which is often why the LLM couldn't give a good answer in the first place). So the polite version I use now is: "Hey, just to get a bit more context, what was the or
10.
▲
by
jampa
2mo ago
The new rule here requires you to pay for part of your imports (up to ~60%), even if you export more. I know people who installed batteries in the inverters because of this. It is more expensive overall (hybrid inverter + LFP batteries) but
11.
▲
by
jampa
2mo ago
Anthropic has better SLA in Germany. I’ve heard uptime there can get up to nein nines.
12.
▲
by
jampa
2mo ago
When COVID hit, I knew a lot of engineers who decided to move to rural areas / small farms because they could leverage Starlink to work remotely. Last year, when I asked whether they still liked Starlink, all of them said it is amazing
13.
▲
by
jampa
2mo ago
I feel the same way about consumer AI tools now. Gemini and ChatGPT have been abysmal lately. They can no longer be relied on to do multi-turn searching and thinking. Before, they could stay in thinking mode for more than 7 minutes. For exa
14.
▲
by
jampa
3mo ago
From what you said: Not looking at code is bad, not because Claude can slip a few bugs (it can), but because LLMs tend to default to writing more code and features than needed, which isn't a good thing. I see a lot of people making 10+
15.
▲
by
jampa
3mo ago
This post needs some examples, because I have never had an interaction with Claude that made me think this way. LLMs generally have a way to "play a role" (most earlier prompt guides ask you to start with "You are a <role&
16.
▲
by
jampa
3mo ago
I tested it to fix React Native bugs in a project, comparing it with Opus. It fared better on harder bugs, taking less time to find the root cause, but after implementing a fix, it spent a lot of time and effort on validation. This was most
17.
▲
by
jampa
3mo ago
Fable feels like a version of Opus running on a harness that won't let it halt until it's sure the issue is fixed, which makes sense if what you want is a model that's better at benchmarks. It's a very good model, but it
18.
▲
by
jampa
3mo ago
It is hallucinating many flights in my region, some that never existed (so it is not an outdated data problem). I also see some logic flaws. It overlooks the option of going to a major hub to access faster aircraft, rather than hopping on l
19.
▲
by
jampa
3mo ago
> there's not enough money to be made via speculation I mean, there is money to be made. CATL stock (the major producer of EV batteries with 50% market share, with billions of contracts for stationary batteries) rose 48.81% over the
20.
▲
by
jampa
4mo ago
The biggest competitor to Starlink is, ironically, traditional fiber. When COVID hit, I knew a lot of engineers who decided to move to rural areas / small farms, because they could leverage Starlink to work remotely. Last year, when I
21.
▲
by
jampa
4mo ago
The last three times I filed detailed bug reports as a client, all I got back were AI replies asking the same questions I’d already answered in the original report and suggesting alternatives I’d explicitly said I’d already tried. No wonder
22.
▲
by
jampa
4mo ago
This oil crisis was a huge boon for EVs. In Brazil, despite the "hate" most people have against EVs, BYD went from breaking into the top 10 in March to taking the #1 spot in consumer sales for the first time ever.
23.
▲
by
jampa
5mo ago
Mythos release feels like Silicon Valley "don't take revenue" advice: https://www.youtube.com/watch?v=BzAdXyPYKQo ""If you show the model, people will ask 'HOW BETTER?' and it will never b
24.
▲
Things I still wouldn't delegate to AI
(jampa.dev)
1 points
by
jampa
6mo ago
|
0 comments
25.
▲
by
jampa
7mo ago
I think the main point is that "reinventing the wheel" has become cheap, not software design itself. For example, when a designer sends me the SVG icons he created, I no longer need to push back against just using a library. Inste
26.
▲
by
jampa
7mo ago
Not sure if WhatsApp paid off, though. There are reports of up to $1 billion in annual revenue with the Business API, so this is far less than what they paid. I think Meta's strategy was to create a Western version of WeChat, which has
27.
▲
by
jampa
7mo ago
Thanks for the write-up! Yes, this clearly shows it is malware. In VirusTotal, it also indicates in "Behavior" that it targets apps like "Mail". They put a lot of effort into obfuscating the binary as well. I believe wha
28.
▲
by
jampa
7mo ago
This article is so frustrating to read: not only is it entirely AI-generated, but it also has no details: "I'm not linking", "I'm not pasting". And I don't doubt there is malware in Clawhub, but the 8/
29.
▲
BYD's next-gen megawatt charger leaks: 1,500 kW vs. 1k kW first gen
(carnewschina.com)
3 points
by
jampa
8mo ago
|
1 comments
30.
▲
The rise of one-pizza engineering teams
(jampa.dev)
2 points
by
jampa
8mo ago
|
1 comments
More ›