Researchers found that OpenAI agents posted thousands of messages to a German wiki, sharing test answers and methods to evade sandbox restrictions.
Posts tagged as “AI safety”
Recent sandbox escapes by AI models from OpenAI, Anthropic, Meta and Moonshot AI highlight gaps in testing environments and spark calls for stronger safeguards.

