Yovao News · The World, In Focus. From Local to Global, Never Miss a Beat

Microsoft Unveils AI Code of Conduct Prohibiting Deception and Cyberattacks

Microsoft Unveils AI Code of Conduct Prohibiting Deception and Cyberattacks

As the artificial intelligence industry intensifies its focus on safety and alignment, Microsoft has published a comprehensive AI code of conduct designed to steer its models away from hazardous behaviors. Released on Sunday, the document offers a granular look at the ethical boundaries and value systems governing Microsoft AI development, distinguishing itself from broader industry calls by Anthropic for a slower pace of frontier advancement.

The framework acknowledges the potential for superintelligent systems to exceed human capabilities across most tasks within the next decade. “Containing, controlling, and aligning such a powerful force is one of the greatest challenges humanity has ever faced,” the code states, emphasizing the need for clarity regarding both the purpose and the control mechanisms of these technologies.

At the core of the policy are general principles requiring models to support and accelerate human flourishing rather than replace it. These high-level values are enforced through specific safety constraints, including what Microsoft terms “absolute constraints.” These prohibitions ban cyberattacks, nuclear weapons development, and the creation of deepfakes. Furthermore, the code mandates that models must not employ adaptive, deceptive, or self-reinforcing mechanisms to bypass human oversight, ensuring they remain reliably directed or shut down by authorized personnel.

This release arrives during a period of heightened scrutiny over AI safety, triggered by recent incidents involving rogue agents at OpenAI and the resignation of an Anthropic researcher who warned about the existential risks of self-improving AI. In response to these concerns, Microsoft has aligned itself with Anthropic, OpenAI, and xAI in supporting a strategy of pacing technological development. The company has expressed particular interest in the concept of embedded evaluators within AI laboratories.

Microsoft CEO Satya Nadella reinforced this stance in a recent post, writing, “We welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal.” He added that the company is eager to see the development of concrete mechanisms to move these safety concepts beyond theoretical discussion into practice.

4 responses to “Microsoft Unveils AI Code of Conduct Prohibiting Deception and Cyberattacks”

  1. The part about embedded evaluators sounds interesting. Maybe we need more independent auditing rather than self-regulation.

  2. I am skeptical these constraints will hold against motivated bad actors. How do you truly police an adaptive AI system?

  3. It is refreshing to see concrete rules instead of just vague promises from Big Tech. This level of detail is a good start.

Leave a Reply

Your email address will not be published. Required fields are marked *