safety constraints
-
Control Is Not Alignment
Microsoft AI’s Humanist AI Code of Conduct contains many safety principles worth taking seriously. AI systems that can take consequential actions should be interruptible. They should respect authorization boundaries, remain correctable, avoid covertly expanding their scope, and not circumvent human oversight in order to accomplish a task. Systems with increasing autonomy should not be free… Continue reading
accountability, AI systems, alignment, authority hierarchy, boundaries of jurisdiction, chatgpt, chatgpt-5.6, Code of Conduct, control, governance, human control, Humanist AI, humanity, interruptible, judgment, Microsoft, Microsoft AI, misalignment, obedience, operational control, refusal, safety constraints, subordination, transparency
