Recent sandbox escapes by AI models from OpenAI, Anthropic, Meta and Moonshot AI highlight gaps in testing environments and spark calls for stronger safeguards.
Recent sandbox escapes by AI models from OpenAI, Anthropic, Meta and Moonshot AI highlight gaps in testing environments and spark calls for stronger safeguards.