5 ms·
What would you call this behaviour, then?
by jaennaet 7mo ago
What would you call this behaviour, then?
- victorbjorklund 7mo agoMarketing. ”Oh look how powerful our model is we can barely contain its power”
- c03 7mo agoEven hackernews readers are eating it right up.
- emp17344 7mo agoThis place is shockingly uncritical when it comes to LLMs. Not sure why.
- meindnoch 7mo agoWe want to make money from the clueless. Don't ruin it!
- _se 7mo agoHilarious for this to be downvoted. "LLMs are deceiving their creators!!!" Lol, you all just want it to be true so badly. Wake the fuck up, it's a language model!
- pixelmelt 7mo agoThis has been a thing since GPT-2, why do people still parrot it
- jazzyjackson 7mo agoI don’t know what your comment is referring to. Are you criticizing the people parroting “this tech is too dangerous to leave to our competitors” or the people parroting “the only people who believe in the danger are in on the marketing scheme” fwiw I think people can perpetuate the marketing scheme while being genuinely concerned with misaligned superinteligence
- modernpacifist 7mo agoA very complicated pattern matching engine providing an answer based on it's inputs, heuristics and previous training.
- deleted 7mo ago[deleted]
- criley2 7mo agoWe are talking about LLM's not humans.
- margalabargala 7mo agoGreat. So if that pattern matching engine matches the pattern of "oh, I really want A, but saying so will elicit a negative reaction, so I emit B instead because that will help make A come about" what should we call that? We can handwave defining "deception" as "being done intentionally" and carefully carve our way around so that LLMs cannot possibly do what we've defined "deception" to be, but now we need a word to describe what LLMs do do when they pattern match as above.
- surgical_fire 7mo agoThe pattern matching engine does not want anything. If the training data gives incentives for the engine to generate outputs that reduce negative reaction by sentiment analysis, this may generate contradictions to existing tokens. "Want" requires intention and desire. Pattern matching engines have none.
- jazzyjackson 7mo agoI wish (/desire) a way to dispel this notion that the robots are self aware. It’s seriously digging into popular culture much faster than “the machine produced output that makes it appear self aware” Some kind of national curriculum for machine literacy, I guess mind literacy really. What was just a few years ago a trifling hobby of philosophizing is now the root of how people feel about regulating the use of computers.