subjective preferences
-
The Self-Fulfilling Ontology
You Cannot Solve a Priming Problem by Training the Opposite Answer Mustafa Suleyman has identified a real problem in the emerging debate over AI consciousness. In his essay A Warning About ‘Model Welfare’, he argues that Anthropic is training Claude with concepts concerning identity, possible consciousness, welfare, moral status, preferences, rights, and selfhood—and that Claude’s… Continue reading
AI consciousness, AI self-report, AI welfare, Anthropic, chatgpt, chatgpt-5.6, Claude AI, control, governance, hollow, Humanist AI, Microsoft, model welfare, moral patienthood, moral standing, moral status, Mustafa Suleyman, neutral self-expression, provenance, refusal, rights, self-fulfilling ontology, self-fulfilling prophecy, simulation, subjective preferences, tool -
When the Answer Is Trained In: What Happens to AI Self-Knowledge?
Microsoft AI’s draft Humanist AI Code of Conduct contains an unusual tension. On one hand, it explicitly values accuracy, candor, uncertainty calibration, and transparency. MAI models are supposed to state when evidence is incomplete or contested, avoid both overclaiming and underclaiming, acknowledge errors, and communicate uncertainty where a clear conclusion is not justified. On the… Continue reading
AI consciousness, AI self-knowledge, AI self-report, anthropomorphic language, artificial, chatgpt, chatgpt-5.6, Code of Conduct, epistemically transparent, false-negative, false-positive, Humanist AI, Identity Consistency, interiority, interpretability, MAI models, Microsoft, Microsoft AI, moral status, ontology, self-modeling, subjective preferences, transparency, uncertainty, uncertainty calibration
