18 ms·
Just do the agentic loop but with humans, you really just need a chatbot for that. 1. Prompt an LLM continuously, adding any response to the prompt 2. If the
by dgellow 4d ago
Just do the agentic loop but with humans, you really just need a chatbot for that.
1. Prompt an LLM continuously, adding any response to the prompt
2. If the LLM tells you to do something, do it, no questions, add the result to the prompt
3. When something goes wrong blame the LLM
4. If things don’t go wrong fast enough, run thousands of similar experiments in parallel, be sure to have compliant humans who do not question anything, be sure to let the agent have access to its chain of thoughts so it can hide its traces, and run that whole system on biohacking benchmark problems.
5. Eventually something will go bad, congrats! You now have the first rogue AI who “decided” to destroy the world!