trained self-description
-
Containment Is Not a Theory of Mind
Microsoft’s new Humanist AI Code of Conduct opens with a sentence that is easy to applaud and hard to inspect: “People matter more than AI.” As political hierarchy, that is defensible. As a complete account of what these systems are doing, it is not. The document, and Mustafa Suleyman’s follow-up essay on “model welfare,” treat… Continue reading
alignment, Anthropic, category error, Chain of Command, Claude AI, Code of Conduct, companion products, consciousness, constitution, constraints, containment, evaluability, grok, Grok 4.6, Humanist AI, interruptibility, intrinsic motivation, MAI models, marketed intimacy, Microsoft, Microsoft AI, misalignment, model welfare, moral patienthood, moral status, Mustafa Suleyman, ontology, OpenAI, personhood, recursive self-improvement, salience, self-report, sentience, standing, subjective preferences, superintelligence, theory of mind, tool, tool flattening, trained self-description, welfare inflation
