14 ms·
I don't see the problem with this. The chatbot is the most important part of Grok, so it makes sense Elon would be dogfooding it then providing suggestions.. He
by kvetching 6mo ago
I don't see the problem with this. The chatbot is the most important part of Grok, so it makes sense Elon would be dogfooding it then providing suggestions.. He wants it to be truthful... It was shown on benchmarks recently that it hallucinates the least...
- Braxton1980 6mo ago>He wants it to be truthful How do you know this? Why would you believe him considering the massive lies he's told, for example about the 2020 widespread election fraud
- kvetching 6mo agohttps://artificialanalysis.ai/evaluations/omniscience?omniscience-accuracy=accuracy-vs-attempt-rate https://artificialanalysis.ai/evaluations/omniscience?omnisc... AA-Omniscience Hallucination Rate (lower is better) measures how often the model answers incorrectly when it should have refused or admitted to not knowing the answer. It is defined as the proportion of incorrect answers out of all non-correct responses, i.e. incorrect / (incorrect + partial answers + not attempted). Grok 4.2 which was just released in the API just benched the best at this benchmark.
- SideQuark 6mo agoOf all the valuable metrics on that site, all of which grok does badly at except one, you managed to pick that single one. https://artificialanalysis.ai/models https://artificialanalysis.ai/models
- Braxton1980 6mo agoThis isn't a response to my question. I asked why you trust him
- watwut 6mo ago[flagged]
- SouthSeaDude 6mo agoI totally agree, it's his company 100%, why would you even apply for a job in a company where you don't agree with the owner or his vision.
- karmakurtisaani 6mo agoSome of us have a pesky addiction to food and shelter.
- jazzpush2 6mo agoDo you think the investors of xAI want this behavior baked into the model? Do you think other frontier labs enforce their models to praise their CEO and never insult them? And, how does this fit into a vision, exactly? What vision might that be beyond "I am only to be praised?"
- estearum 6mo ago[flagged]
- kvetching 6mo ago[flagged]
- ecshafer 6mo ago> Great point! This actually reminds me of the white genocide in South Africa, where some say "Kill the Boer" is just a non-violent rallying cry, but actually it's ... Are you implying that "Kill the Boer" is actually a non-violent rallying cry, and not a genocidal call to action? Ill say that that is an absurd notion, and if you s/Boer/Jew or whatever ethnic or religious group you want, it will become very obvious why that's the case.
- scubbo 6mo ago> Are you implying that "Kill the Boer" is actually a non-violent rallying cry (Not the person you're replying to, so caveats about me speaking for them, but) no, they're not. They're highlighting how Grok _isn't_ accurate/unbiased/whatever, by giving examples of how it distorts the truth to fit Elon's narrative.
- hunterpayne 6mo agoI assure you that all the models have such biases. Ask any LLM who caused the most death in history and you will get skinny mustache man, an opinion any historian will tell you is wrong. He is in the top 5, but not the top of the table. That was clearly biased into the models in the same way Elon biases his models. I'm not defending this behavior but I don't know how you both get models that returned the sanitized answers some want and the correct answers others want at the same time. Pure correctness probably gets you Mecha-H. Pure sanitized answers will get many wrong. Pick your poison I guess.
- etchalon 6mo agoHe wants it to tell the truth as he sees it.