6 ms·
I've started feeling slightly physically ill when I read Opus output for hours straight. This article rings very true for me. I've started complaining about it
by block_dagger 2mo ago
I've started feeling slightly physically ill when I read Opus output for hours straight. This article rings very true for me. I've started complaining about it with my team; at least have a personal style guide in your agent rules that eliminates emdashes, the "it's not X, it's Y"s, the long lists of modifiers before the noun, using the word "land" to mean finish, etc. I hope this is just a phase of adolescent LLMs.
- helloplanets 2mo agoIt's kind of offputting how much Anthropic models these days keep repeating "real", "genuine" and "honest". They've RL'd that way over the top.
- vinodc 2mo ago"genuinely load-bearing" is the one that triggers me the most now.
- block_dagger 2mo ago"You're right for pointing this out. Honestly, your comment raises a real concern — they genuinely RL'd that over the top." - Claude
- dpkirchner 2mo agoI had Claude make a world cloud from its responses because I was curious to see how big "honest" would be. It barely showed up, so I asked it to just give me the counts and it responded telling me it was trained not to use the word "honest" much because it makes people distrust responses (in addition to showing me the counts).
- helloplanets 2mo agoI just did this on one .claude directory and >20% of the answers there included some variation of "real", "actual", "exact", "honest", "genuine", "valid", "true". ~15% in that directory contain some variation of "real", "genuine" or "honest". This is excluding thinking tokens, sub-agent output, etc.
- burningChrome 2mo agoThis is one of those things I barely noticed because I tend to read fast and skim. Someone pointed the over use of these terms and now its like hitting a set of spike strips every time I'm reading the output from any given model. Its like when someone points something out a in picture you never saw and now you cannot "unsee" it ever again.
- strken 2mo ago`arc land` is burnt into my brain by Phabricator, so I'm aware that the term predates LLMs, but it still drives me nuts. It's impossible to undo some of these linguistic wobbles. Even if you could filter out 100% of LLM input, the humans themselves are learning to say "land" at a higher frequency now.
- mattas 2mo agoI was describing this exact feeling today. I haven't quite been able to put it into words but I do get slightly physically ill. Almost similar to mild trypophobia?
- alexchantavy 2mo agoVoice really matters in writing. If everyone uses Opus to write without editing, then it all sounds the same regardless of who it came from.
- digitaltinfoil 2mo agoI had to tell it never to say "hand-wavey" ever again to me. But I agree, I hate the way LLMs phrase sentences.
- deleted 2mo ago[deleted]
- throwaway_7274 2mo agoMe too. It feels like I’m taking psychic damage from reading so much of this stuff. Contrary to the theory that it’s “just the contract workers’ Nigerian English,” I think the models are developing an ultra-terse hyper-stylized dialect of their own under RL pressure. They seem to be writing increasingly in _code_, and I don’t mean computer code. The words don’t mean quite what they mean to humans.
- nprateem 2mo agoOver the last few days with fable I've found it at times incomprehensible, terse word salad. It also invents phrases assuming I'll understand (but that could be because it's reusing terms in the codebase I no longer remember). I've often had to paste its output back in to ask it what it actually means. Weird. I think the main thing is just fatigue. There's so little variety. Each model has its preferred idiolect which everyone becomes tired of due to ubiquity. That's the worst part. It's like always eating fast food.
- hotcrossbunny 2mo ago"It's like always eating fast food". That right there
- cpt_sobel 2mo agoMy non-English-native-speaker head of development, to whom I report, does 100% of his work using LLMs and doesn't even check if the code compiles, but somehow this isn't my biggest problem with it – it's the botspeak in the PR comments (or answers to my PR comments) that are so clearly not written by him, and the documentation that makes absolutely 0 sense sometimes even if I break it down. Just a word salad of "robust", "maintainable", "smoke test" that amount to absolutely nothing. And the "You're absolutely right, I fixed it" responses (narrator: he didn't fix it). I used to have a lot of fatigue due to it until I stopped caring.
- throwaway_7274 2mo agocomment reads clean. drafting response when it lands. *onanizing…
- ssl-3 2mo ago"That's such a clever way to see things! Let's delve into that!" The bots (all of them) seem to show patterns of overuse of specific phrases, words, and punctuation. Some of those are the ones you mentioned. Another that I've been seeing lately is overuse of the term "gate", wherein: As a human, I know what a gate is. A gate is a thing that can be open, or that can be closed. It might be locked or unlocked. The path beyond the gate may be passable or impassable or nonexistent. The gate is just a gate, and the presence of the gate doesn't imply whether it is open or closed. But in bot-speak, a gate only refers to a hard block -- an impassable construct. Like a fence or a wall, or even a lava-filled moat. But while a lava-filled moat is intended to be impassable, the bot uses "gate" -- a thing that is designed to be passed -- to describe that same kind of obstacle. That's misuse of the term, I think, based on decades of dealing with gates in reality: Usually when I encounter a gate that is closed, I just open it and walk through. I do have instructions that tell the bot to avoid that usage of the word and it ignores them sometimes anyway. But "gate" is just today's problem-word that comes to mind as I write this. Yesterday, it was something different. Tomorrow, it will be something else entirely. The overall pattern here is that of gratingly-repetitive bullshit-grade jargon that doesn't fit to begin with. "And that's the real, no-nonsense truth!"
- stcg 2mo agoI like the word "botspeak". Another example of typical botspeak is "smoke test". Why not just say "test"? It feels like a way of downplaying the ability to detect problems.
- ssl-3 2mo agoOne thing I did recently with the bot definitely involved actual smoke tests, though: I was working with real hardware that can blow up in real ways, with the bot doing all of the circuit design work and coding based on my goals while I just distantly commanded the show from On-High and plugged shit into a breadboard. (The project works well and I consider it to be Good Enough; I might go back and polish it more later. There was no smoke, but there could have been.)
- zahlman 2mo ago
- pacifika 2mo agoKey points only. Anything written for humans should be written by humans.
- vinodc 2mo agoThis has been my experience as well. It's incredibly grating to repeatedly read "genuinely load-bearing", "honestly?", etc. I've tried to get Opus to stop using these phrases via an entry in its memory, with mediocre results.