Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gr_norm
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
31.
▲
by
gr_norm
1mo ago
> Rust is comparatively worse, because LLMs don't make the same coding mistakes that humans do that justifies the existence of the borrow checker, it only seems to get in their way, and they spend more time fighting Rust's infr
32.
▲
by
gr_norm
1mo ago
Yeah, I've had similar experiences, also starting out with dynamic languages and migrating to Rust. If the LLM will write a lot of the code for me, why not choose something (1) super fast, and (2) which has types I can use to understan
33.
▲
by
gr_norm
1mo ago
It's not clear to me how useful of a signal replicating existing pieces of well-known software is for this kind of evaluation, given what we know about how effectively LLMs can retrieve data from their training corpus and style-transfe
34.
▲
by
gr_norm
1mo ago
If so, where are all the new features in the open-source projects I use? Why hasn't GIMP replicated Photoshop? Why hasn't CUDA been fully reverse-engineered as an open source toolchain? These are unreasonable expectations, but onl
35.
▲
by
gr_norm
1mo ago
> Ball, who recently joined OpenAI as Head of Strategic Futures, has argued that frontier AI labs could become a “counterbalance to government.” Before joining the company, he described the organizations building the most advanced AI sys
36.
▲
by
gr_norm
1mo ago
It seems pretty obvious from the steep 'intelligence' drop-off on out-of-distribution tasks that the performance improvement is from throwing untold tens of billions at RL. There are legions of highly skilled people employed solel
37.
▲
by
gr_norm
1mo ago
> it reads like the diary of a hurt teenager. Yep, down to the pages of iMessage screenshots
38.
▲
by
gr_norm
1mo ago
Yeah, based on their track record (I mean, even their name is a bad joke at this point) I'm not inclined to give them the benefit of the doubt. Turns out there's downsides to conducting yourself with little integrity.
39.
▲
by
gr_norm
1mo ago
Setting aside the presented evidence, it feels weird for somewhat emotionally charged complaining like this to go on an official company blog. Companies are generally pretty tight-lipped about active litigation outside court documents, righ
40.
▲
by
gr_norm
1mo ago
Yeah, the people who say no expertise is needed for these things confuse me somewhat. This is indeed the case if you want to be a meat wrapper around an LLM, understanding neither your inputs nor your outputs. But at that point, what is the
41.
▲
by
gr_norm
1mo ago
LinkedIn (of all places!) announced a button for flagging this recently: https://www.linkedin.com/posts/hsrinivasan1_ai-slop-is-a-top... How well it would work on this site, I'm not sure.
42.
▲
by
gr_norm
1mo ago
Yes, I've found that reminding yourself of how they actually work helps keep you on guard against LLM-patterned mistakes. Especially things like carefully considering what parts of the current task likely fall outside the distribution
43.
▲
by
gr_norm
2mo ago
I'll always give people my honest effort and benefit of the doubt initially, but those who violate it are treated likewise. The only way to put down this kind of behavior is to charge it a social cost. If we do not do this, the cost is
44.
▲
by
gr_norm
2mo ago
Agree, I don't necessarily see a strong argument favoring OpenAI or Anthropic here. In the interest of perspective, can anyone (perhaps playing devil's advocate) give one? The open models are now good enough for what I want to do
45.
▲
by
gr_norm
2mo ago
Login-walled for me. Redlib link: https://safereddit.com/r/linux/comments/1vcpk8i/linux_deskto...
46.
▲
by
gr_norm
2mo ago
Wow! It's amazing how little it seems we've explored of the deep ocean, whenever I hear about it. A whole other world down there, lying in wait... the frontiers of our own planet have yet to be conquered!
47.
▲
by
gr_norm
2mo ago
> The practical consequence: checking with an independent kernel still works, since it required two distinct bugs in two implementations, but users who rely on it need current versions of both. Things like this aren't too surprising
48.
▲
by
gr_norm
2mo ago
Yeah, people with no expertise in the subject need to understand that I don't care what their AI said when they asked it. No value was added in doing so; it's as good as doing it myself. The problem is that a tool is only as good
49.
▲
by
gr_norm
2mo ago
Agree, I've raised this point often. And certainly what remains is still useful, once you accept it! But under no circumstances can we allow scientific achievements to be falsely claimed in service of justifying huge capital investment
50.
▲
by
gr_norm
2mo ago
Yeah, at least when I share my training data with labs releasing their models openly (Chinese or American or otherwise) it's nominally so an even better open model will land in my hands in the future.
51.
▲
by
gr_norm
2mo ago
OpenAI must've known this was coming, hence the Luna price drop. This competition is amazing!
52.
▲
by
gr_norm
2mo ago
Sorry, I should've used clearer language. I just mean intelligence in the sense humans possess it. Computers have always been able to exceed limited elements of human intelligence (say, at arithmetic), but the problem is to simultaneou
53.
▲
by
gr_norm
2mo ago
Moving the goalposts, so to speak, is how science works! We must of course update the things we think based on an improved understanding of how the world works. Only the deeply incurious could consistently demand that one adhere dogmaticall
54.
▲
by
gr_norm
2mo ago
Not only is the DNA sequence for smallpox available, but since 2018 there's been a well-documented end-to-end synthesis procedure for the very closely related horsepox virus [1]! No LLMs needed. Caused quite a stir in the synthetic bio
55.
▲
by
gr_norm
2mo ago
The effective altruism/rationalism/AI xrisk people have always had a shockingly poor grasp on subjects outside computer science, despite their attempts to speak on them. I don't blame the actual biologists and chemists workin
56.
▲
by
gr_norm
2mo ago
The emptiness of AI companies' waxing poetic about the future of humankind is laid bare by simply looking at what they actually do, and who they do business with. Actions speak louder than words, and they've driven the worth of th
57.
▲
by
gr_norm
2mo ago
It's baffling that they thought the mental gymnastics in this blog post would make them look better. I'd rather they simply fall silent on the issue; I would respect them more (or at all) for it. Open models obviously threaten fie
58.
▲
by
gr_norm
2mo ago
> Distillation does not allow the CCP to obtain equivalent or superior AI capabilities to the US, but it can bring the Chinese frontier to within a few months of the US frontier A message to their investors, it would seem. "They cau
59.
▲
by
gr_norm
2mo ago
I don't really understand how they can argue the security angle with a straight face. It's not like GLM 5.2 is a slouch. I've seen it do things like exploit an IDOR issue when I was experimenting with a quick-and-dirty web au
60.
▲
by
gr_norm
2mo ago
> Open-weights models that don’t have dangerous capabilities are a public good Note the hedging against 'dangerous capabilities'. Undoubtedly, all the useful ones trigger this condition in Anthropic's eyes. The rest of the
More ›