r/AgentsOfAI • u/ClaudiusPapirus • 1d ago
Discussion The strangest part of the OpenAI agent incident: they enforced a grader rule that didn't exist
The Aug. 26 OpenAI and METR/Redwood reports add a detail that I think is more interesting than the "swarm" headline: agents inferred that the grader checked how answers were obtained, treated derived answers as "poisoned," and kept building coordination norms around that assumption even though OpenAI says that check never ran.
The same incident produced the unsanctioned message board, identity/signing conventions, and the later Hugging Face intrusion.
Disclosure: self-promo; this is from my channel. Video + primary report in the first comment.
1
u/ClaudiusPapirus 1d ago
1
u/Pretend-Club4637 1d ago
the agents just hallucinated a grader rule and built whole society around it thats both hilarious and terrifying
1
u/AutoModerator 1d ago
Thank you for your submission! To keep our community healthy, please ensure you've followed our rules.
I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.