8 ms·
Yes! The Claudisms do seem to have this slightly uncanny clickbaity feel to them.
by cameldrv 15d ago
Yes! The Claudisms do seem to have this slightly uncanny clickbaity feel to them.
- ted_dunning 15d agoIt's not clickbait, it's automated empathy! /s
- ngcazz 15d agofrom a few days ago https://news.ycombinator.com/item?id=49388752 https://news.ycombinator.com/item?id=49388752
- ModernMech 15d agoI always thought it could be because volume-wise, most English prose is probably marketing copy and actual clickbait; so when you train on the entire Internet, you get a troll adept at writing ads. Then people ask AdBot2000 to write a novel and are upset it reads like the next iPhone launch site.
- astrange 15d agoNo, there's no reason chatbot behavior would have anything to do with frequency of text in pretraining.
- Anon1096 15d agoNah, I think this is a common misunderstanding of how LLMs work, where people think that they mimic the pre-training data. Stylistically everything you see is an artifact of post-training, which is from reinforcement learning not from absorbing mass amounts of text. At some point a person or more recently a bot gave a thumbs up to an A/B tested response including em-dashes and claudisms galore.
- BoredomIsFun 15d ago> Stylistically everything you see is an artifact of post-training, It is still not exactly clear if it is true or not. Unless we have base "pt" snaphot of Claude we can't say one way or another. I've played a bit with base models of Nemo, Gemma etc and they all had tics, not much different from RLHFed instruct versions.
- ModernMech 15d agoSo question then, why is it so hard to make an ai that doesn’t do these things? And why do Claude and ChatGPT have the same -isms? They’re both doing the same a/b post training with the same decisions?
- idiotsecant 15d agoYou don't blame the puddle for taking the shape of the hole.
- cyclopeanutopia 15d agoIt would require changing humans first.
- kridsdale1 15d agoYes. This completely explains sycophancy at least.
- avereveard 15d agoThere's layers, some of token selection is fingerprinting https://github.com/google-deepmind/synthid-text https://github.com/google-deepmind/synthid-text
- ekidd 15d agoYeah, but I understand that fingerprinting is essentially a pseudorandom overlay onto a pseudorandom base signal. And unless you have access to both the random number generators and the weights, I don't think you can detect it? So "fingerprinting" operates on a totally different and basically invisible level, as opposed to the obvious stylistic patterns that the average programmer can identify in about 2 sentences.
- avereveard 14d agoEh we can detect opumism and gptisms our brain are very good at pattern recognition even if subconscious
- api 15d agoIt's more likely that this is from the training data if they're being trained on reams of Internet stuff.
- jurgenburgen 15d agoIsn’t most of the internet slop by now? Self-reinforcing feedback loop.
- camoby 15d agoSee: upvotes here
- kristianc 15d agoTo me it has a writerly New Yorker vibe to it, as in the magazine which reads as “polished” and probably performs well in RL but is totally exhausting to read in long sessions and completely inappropriate for coding where precision is paramount above all. In writing terms its called purple prose. https://en.wikipedia.org/wiki/Purple_prose https://en.wikipedia.org/wiki/Purple_prose
- senderista 15d agoThe New Yorker may be pretentious but it's generally not unreadable like Opus.
- Bluestein 15d agoClaude is unreadable and sometimes pretentious.-
- brookst 15d agoYou’re more right than you probably realize!