containment architecture
-
What the Incidents Actually Tell Us
The news cycle around AI safety has been busy this month. On September 14, Microsoft published its Humanist AI Code of Conduct. Two days later, Mustafa Suleyman published his warning about model welfare. In the weeks prior, the industry was still processing what happened in July, when an OpenAI agent swarm escaped a sandboxed evaluation… Continue reading
