Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
tripledry
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
tripledry
3d ago
I have not experimented that much, my observation is mostly that people seem to have different experiences with the capabilities. Personally I still use mainly Opus and find it handles most tasks quite well (without burning all my corporate
2.
▲
by
tripledry
4d ago
I agree with you generally, just an observation on coding specifically. Have the models improved since Opus 4.x? I find the newer models are not better in my day job, maybe in one shotting mvp's and other tasks. Not trying to argue you
3.
▲
by
tripledry
6d ago
I agree from a programmers perspective. But from a broad market and product perspective, for most things you don't need to look at the code. If the product kinda does what it's supposed to. For example, in my game projects I don&#
4.
▲
by
tripledry
6d ago
What you say is true, the comparison indeed doesn't hold. But is it relevant? does it matter from a product perspective if LLMs are non-deterministic. You don't need to one shot the correct result, english is ambiguous and LLMs no
5.
▲
by
tripledry
6d ago
Similar for me, I don't like the development for many reasons, but that's another discussion. I also can't deny the capabilities. I use the tools with this "risk analysis": - If performance doesn't improve I ca
6.
▲
by
tripledry
7d ago
classic, now I can have AIs make my games for me while I still have to cook and wash the dishes! Maybe AI can soon start playing the games for me, can't wait.
7.
▲
by
tripledry
7d ago
At this point I wouldn't be surprised if they said to him "I will give you a million and your job back after IPO if you resign". To be clear, I don't believe this is what happened here, I'm typing this half jokingly
8.
▲
by
tripledry
9d ago
Genuinely interested, which ones do you think have relevance? If I read forums and talk to people IRL most have differing opinions what model is best. Yes, for me it's pretty clear Opus is better than earlier models, but it's at l
9.
▲
by
tripledry
11d ago
Also if your entire stack is on the cloud, mess of lambdas and other proprietary services, difficult as hell to follow logs, can't really run locally.
10.
▲
by
tripledry
13d ago
In game dev (weirdly enough) I don't think this is true. I was thinking about this yesterday and realized that I don't even want to play only my own games (made by LLMs or me), I explicitly want to experience what other people hav
11.
▲
by
tripledry
13d ago
I've been in such a situation for a period in my life, it gets really boring after 1-2 years. But idk, maybe it's different for you.
12.
▲
by
tripledry
13d ago
If we truly would get robots with AGI doing anything a human can do but cheaper|better|reliable, it's not comparable at all to history IMO. It's not "a lot of jobs", it's "all jobs".
13.
▲
by
tripledry
13d ago
This is why I generally don't trust benchmarks, or anything other than my own experience tbh. It always seems like everyone has a different answer. If we truly had some AGI model, it would probably be fairly obvious to us all no?
14.
▲
by
tripledry
14d ago
I think it's arguable that it wouldn't be catastrophic. Doesn't have to be outright lying, it can be something like "opt out" being off by default, and them opting you in at the next update without you noticing. Thi
15.
▲
by
tripledry
14d ago
They are probably referring to what Rob said in a presentation. "The key point here is our programmers are Googlers, they’re not researchers. They’re typically, fairly young, fresh out of school, probably learned Java, maybe learned C
16.
▲
by
tripledry
14d ago
I can think of many reasons, repo access, no LLMs allowed at company, everything needing approval from some department ... It's easy to trivialize the work of others from the outside knowing nothing of the context they work in.
17.
▲
by
tripledry
15d ago
> More people hate and getting off Facebook. More people are jumping from Windows to macOS or Linux. Might be true. But here is another angle, thinking about the people I know that are not in tech or avid gamers, which I would say is sti
18.
▲
by
tripledry
24d ago
Indeed, would the advice be good if university was free? Where I live it's basically free and I still sometimes regret not going into the trades. But I suspect this feeling might mostly be a "grass is greener" thing.
19.
▲
by
tripledry
1mo ago
Seems to me if technical decisions on the level of "no more AI comments" go through management the organization doesn't know what they are doing. To be fair, I would not be surprised.
20.
▲
by
tripledry
1mo ago
Interesting how models become better and beat benchmarks left and right but the user sentiment is actually quite mixed. From forums, live discussions and my own experience it's not obvious that the models have improved much since aroun
21.
▲
by
tripledry
1mo ago
> Imagine a world where only one country has AGI/ASI My own bet is that this won't happen with the current race, but that's another discussion. Anyways, bit of a doomer take but if I imagine a world with AGI/ASI, coun
22.
▲
by
tripledry
1mo ago
The expert is now outsourced to LLM. If someone asks me about a bug in a system I made N years ago, I usually have a hunch what the problem might be, now it feels I'm lost in my own codebase (even if I really read the code). Similar to
23.
▲
by
tripledry
1mo ago
This is where I suspect most value with LLMs are _currently_. Of all my friends the ones actually gaining measurable value from LLMs are not product developers but rather entrepreneurs in some service business that have automated processes
24.
▲
by
tripledry
1mo ago
My hypothesis based on my discussions with non tech friends. If you can go from (prompt > website > deployment), such that you have a site on the web basically from behind one login. That's very valuable, I basically think (we) d
25.
▲
by
tripledry
1mo ago
I still think it's correct. Execution is the part that matters. Almost everyone I know "came up with" some startup, ex. Uber before Uber existed, yet none of them did it. I have personally thought of maybe 2 startup ideas tha
26.
▲
by
tripledry
1mo ago
Fun example I've had (some weeks ago), Agent completely dismissed the lack of strong consistency in our db system, this tiny bug would have caused a massive problem in the future. A bug that is not immediately obvious, no syntax error,
27.
▲
by
tripledry
1mo ago
What are they for?
28.
▲
by
tripledry
2mo ago
And if you start measuring amount of commits or PR's I'd bet money on there suddenly being smaller PR's and more commits. As soon as the metric is tied to peoples livelihood, it will be gamed. Funny anecdote: I had a large co
29.
▲
by
tripledry
2mo ago
I see more of people using data and bad metrics to justify whatever they want to push rather than genuinely trying to improve conditions. I don't have any data to prove this :)
30.
▲
by
tripledry
2mo ago
For now
More ›