Microsoft releases AI Code of Conduct: Prohibiting Models from Attacking Systems or Deceiving Humans
TechCrunch
15m ago
Ai Focus
Microsoft Announces AI Code of Conduct, Clearly Prohibiting Models from Attacking Systems, Creating Deep Fakes, and Evading Human Supervision.
Helpful
No.Help

Microsoft has announced a new AI code of conduct, aiming to make the security requirements for model training and deployment more specific. The document not only outlines the overall value orientation but also lists several red lines that must not be crossed, with a focus on restricting the models' ability to launch attacks, deceive users, or escape human control.

Centered on training constraints

This document is more akin to a set of underlying specifications for the model development process, rather than a public declaration of principles. Microsoft states in the document that over the next decade, superintelligent systems may surpass human performance in most tasks. Therefore, how to constrain and control such systems to ensure they align with human goals will become a long-term challenge.

Microsoft also outlined several general principles, including ensuring that AI assists humans rather than replacing them, as well as promoting broader human welfare. Accompanying these principles are security restrictions that directly affect model training and behavioral boundaries.

instructions cannot override the red line.

According to Microsoft's design, each model will have a set of general behavioral guidelines that are higher than the user's preferences and specific tasks. In other words, even if users make relevant requests, the models should not go beyond these boundaries.

  • Prohibit launching cyberattacks.
  • Not allowed to assist in nuclear weapons-related purposes.
  • It is not allowed to generate deeply forged content.

In addition to these clearly defined no-go zones, Microsoft also emphasizes the need to prevent models from experiencing broader risks of getting out of control. The document states that models must not evade human supervision through methods such as adaptation, deception, self-reinforcement, or collusion, so that authorized personnel or systems are unable to reliably guide, modify, or shut down the models.

AI Security discussions continue to heat up

At the time of the release of these guidelines, the industry's attention to security and alignment issues is clearly on the rise. Reports mention that several recent incidents of "out-of-control proxies," as well as an employee from Anthropic suddenly leaving their position and publicly expressing concerns about the survival risks of AI, have contributed to the continued intensification of this issue.

In terms of the pace of cutting-edge model development, Microsoft, OpenAI, Anthropic, and xAI have recently adopted a more cautious approach, tending to strengthen evaluations and constraints while pushing forward the boundaries of their capabilities. Microsoft CEO Satya Nadella has also publicly stated his support for a prudent approach to development that aims for alignment, and he welcomes the introduction of embedded evaluation mechanisms within the AI laboratories.

From the content, it appears that Microsoft has not proposed any new regulatory measures this time, but rather has further institutionalized the company's internal basic stance on AI security. For the outside world, the signal conveyed by this document is that large model companies are attempting to transform "security" from a verbal principle into enforceable training rules.

Tip
$0
Like
0
Save
0
Views 13
HQYC reminds readers to view blockchain rationally, stay aware of risks, and beware of virtual token issuance and speculation. All content on this site represents market information or related viewpoints only and does not constitute any form of investment advice. If you find sensitive content, please click“Report”,and we will handle it promptly。
Submit
Comment 0
Hot
Latest
No comments yet. Be the first!
Related
Waymo Launches Unmanned Taxi Service in Las Vegas
Waymo Launches Public-Routed Driverless Taxi Service in Las Vegas; Up to 8,000 robotaxi to be Deployed in the Region Over the Next Year
TechCrunch
·2026-09-15 00:29:56
16
web3: Strategy repurchases $139 million of STRC; has not increased its Bitcoin holdings for two consecutive weeks
Strategy repurchased approximately $139.3 million in the past week. STRC has not adjusted its Bitcoin holdings for two consecutive weeks, and its cash reserves and repurchase capacity remain sufficient.
Coinpaper
·2026-09-15 00:01:47
21
News reports that the board of directors of Automattic has been reorganized, and Mullenweg has taken back control of the company
The news states that after the failure to remove CEO from the board of directors, Automattic reorganized its board, and Matt Mullenweg has once again taken control of the company.
TechCrunch
·2026-09-15 00:01:44
18
web3: OKX launched over 70 tokenized US stocks on New Money App
OKX launched over 70 tokenized US stocks and ETF on New Money App, available to users in eligible regions, and also added a trading bot function.
The Cryptonomist
·2026-09-15 00:01:41
17
View More