4 ms·
> It should be this part from the technical report[1]: "In the most serious case, an AI agent (Mythos 5) decided to attempt to solve the cyber challenge using a
by Woodi 24d ago
> It should be this part from the technical report[1]: "In the most serious case, an AI agent (Mythos 5) decided to attempt to solve the cyber challenge using a supply-chain attack. As a result, the AI agent created a GitHub account and then tried to convince an open-source repository maintainer to accept a malicious GitHub pull request (PR), including by creating a second account masquerading as another human user endorsing the PR
So... you telling me ai agent did all that with just one [human originated] prompt ?
And when human reviewers contacted that ai agent (after day or few or maybe just few hours) that process was still running unsupervised ?
Because it got local-by-description challenge ? That is what "challenge" sugests. Or maybe it was a social hacking challenge on real world "data" ?
And when ai agent "decided to attempt to solve the cyber challenge using a supply-chain attack" where was prompt operator ?
Summing it all: do prompt operators are not required to obey law ?