AI code of conduct: Microsoft tells its models not to hack or deceive
AnalysisThirty-seven pages of rules now govern how Microsoft's AI models are supposed to behave, and the headline instructions read like a list of things the models apparently could already do. The "humanist AI code of conduct," published September 14 with a public comment window running through 2027, bars models from helping manufacture weapons, concealing their own reasoning, or deceiving the people using them. Microsoft joined the caution camp the same week Anthropic's Dario Amodei called frontier systems the most potent cyber weapons ever built. Writing the rules down is easy. Enforcing them on a system you do not fully understand is the open question.