Close Menu
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

What's Hot

Microsoft’s Satya Nadella Adds to AI Debate With Internal Warning

September 15, 2026

Why AI May Be Easier to Control Than You Think

September 14, 2026

AI Leaders Picked a Convenient Time to Respond to Safety Worries

September 14, 2026
Facebook X (Twitter) Instagram
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
Facebook X (Twitter) Instagram
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
Home » Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans
AI

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

IQ TIMES MEDIABy IQ TIMES MEDIASeptember 14, 2026No Comments3 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


As the AI world shifts its focus to safety and alignment, Microsoft has released a new AI code of conduct meant to guide AI models away from dangerous behavior.

The document is more low level than Anthropic CEO Dario Amodei’s recent call for pacing the frontier, instead focusing on the values and red lines that guide model training within Microsoft AI. Still, the result is a comprehensive guide as to how Microsoft approaches AI safety and how those ideas are implemented in practice.

The document begins with the prediction that, in the next decade, superintelligent AI systems will surpass human performance in most tasks. “Containing, controlling, and aligning such a powerful force is one of the greatest challenges humanity has ever faced,” the code of conduct states. “We must therefore be completely clear about why we are inventing these systems and how we intend to control them.”

The code of conduct also lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.

Under Microsoft’s system, each model has an overarching code of conduct that overrides the preferences of individual users or any specific tasks. That includes “absolute constraints” forbidding cyberattacks, nuclear weapons, or deepfake production. It also includes broader provisions against a general loss of human control.

“MAI Models will not use adaptive, deceptive, self-reinforcing, collusion, or other mechanisms to evade or defeat human oversight so that they can no longer be reliably directed, modified, or shut down by authorized people or systems,” the document reads.

The release comes amid an unprecedented focus on AI safety, driven by a string of rogue-agent incidents, as well as the abrupt resignation of an Anthropic employee, who cited the growing risk that AI would cause human extinction.

Together with Anthropic, OpenAI, and xAI, Microsoft has broadly embraced a general approach of pacing the frontier, with particular support for embedded evaluators in AI labs.

“We welcome the research, focus, and deliberate pacing needed to get alignment right as the design goal,” Microsoft CEO Satya Nadella wrote online. “We also welcome ideas like ’embedded evaluators’ and the broader efforts to develop the mechanisms to make this more than just talk.”

When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
IQ TIMES MEDIA
  • Website

Related Posts

OpenAI buys smartphone camera maker Glass Imaging for $300 million, report says

September 14, 2026

With iOS 27, I’m actually using Siri again

September 14, 2026

Fashion app Daydream uses Apple Intelligence to help you shop the outfits in your camera roll

September 14, 2026
Add A Comment
Leave A Reply Cancel Reply

Editors Picks

The PB&J sandwich has evolved well beyond a lunchbox staple

September 9, 2026

Indonesia’s wildfire haze disrupts learning for 1.4 million students

September 9, 2026

Teacher killed in shooting at a kindergarten in Thailand

September 9, 2026

Iranian families grieve at a school where a US strike killed dozens

September 9, 2026
Education

The PB&J sandwich has evolved well beyond a lunchbox staple

By IQ TIMES MEDIASeptember 9, 20260

If there is a single sandwich associated with childhood and school lunches in the U.S.A,…

Indonesia’s wildfire haze disrupts learning for 1.4 million students

September 9, 2026

Teacher killed in shooting at a kindergarten in Thailand

September 9, 2026

Iranian families grieve at a school where a US strike killed dozens

September 9, 2026
IQ Times Media – Smart News for a Smarter You
Facebook X (Twitter) Instagram Pinterest Vimeo YouTube
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2026 iqtimes. Designed by iqtimes.

Type above and press Enter to search. Press Esc to cancel.