Close Menu
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
What's Hot

Texas landowners ask court to halt President Trump’s border wall | Donald Trump News

September 14, 2026

Learn how to build persistent AI defenses at Disrupt 2026

September 14, 2026

India’s inflation rate rose to 4.82% in August, slightly higher than expected

September 14, 2026
Facebook X (Twitter) Instagram
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Facebook X (Twitter) Instagram
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Home » Microsoft’s new AI ‘code of conduct’ instructs models not to hack systems or deceive humans
AI

Microsoft’s new AI ‘code of conduct’ instructs models not to hack systems or deceive humans

Editor-In-ChiefBy Editor-In-ChiefSeptember 14, 2026No Comments2 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
Follow Us
Google News Flipboard
Share
Facebook Twitter LinkedIn Pinterest Email


As the AI ​​world shifts focus to safety and alignment, Microsoft has released a new AI Code of Conduct aimed at guiding AI models away from dangerous behavior.

This document is lower-level than Anthropic CEO Dario Amodei’s recent call to keep pace with the frontier, instead focusing on the values ​​and red lines that guide model training within Microsoft AI. Still, the results are a comprehensive guide to how Microsoft approaches AI safety and how its ideas are implemented in practice.

The paper begins with the prediction that within the next decade, superintelligent AI systems will outperform humans on most tasks. “Containing, controlling and coordinating such powerful forces is one of the greatest challenges humanity has ever faced,” the code states. “So we need to be completely clear about why we are inventing these systems and how we intend to control them.”

The Code also provides general principles that Microsoft AI models must adhere to (for example, supporting humans rather than replacing them and accelerating human flourishing) and specific safety constraints for implementing those principles.

In Microsoft’s system, each model has a comprehensive code of conduct that disables individual user preferences and certain tasks. This includes “absolute restrictions” banning cyber-attacks, nuclear weapons and the production of deepfakes. It also includes broader provisions for a general loss of human control.

“MAI models do not utilize adaptive, deceptive, self-reinforcing, collusive, or other mechanisms to evade or defeat human oversight, so they cannot be reliably directed, modified, or shut down by authorized persons or systems,” the document says.

The release comes amid an unprecedented focus on AI safety due to a series of rogue agent incidents and the sudden resignation of Anthropic employees due to the increased risk of AI causing human extinction.

Microsoft is broadly embracing common approaches to address frontiers, including specifically supporting AI Lab’s built-in evaluators, along with Anthropic, OpenAI, and xAI.

“We welcome the research, focus, and intentional pacing needed to get alignment as a design goal right,” Microsoft CEO Satya Nadella wrote online. “We also welcome ideas like ’embedded evaluators’ and broader efforts to develop mechanisms to make this more than just lip service.”

If you buy through links in our articles, we may earn a small commission. This does not affect editorial independence.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Editor-In-Chief
  • Website

Related Posts

Learn how to build persistent AI defenses at Disrupt 2026

September 14, 2026

Fashion app Daydream uses Apple Intelligence to help you shop for outfits in your camera roll

September 14, 2026

iOS 27 actually uses Siri again

September 14, 2026
Add A Comment

Comments are closed.

News

Texas landowners ask court to halt President Trump’s border wall | Donald Trump News

By Editor-In-ChiefSeptember 14, 2026

Native American tribes, environmentalists, and local residents in several states oppose construction of a border…

President Trump says demands for more control over AI are a ‘chic conspiracy’ | Business and Economic News

September 14, 2026

Carney pitches Canada to global investors amid US trade war | Business and Economic News

September 14, 2026
Top Trending

Learn how to build persistent AI defenses at Disrupt 2026

By Editor-In-ChiefSeptember 14, 2026

For AI founders, the biggest competitive threat may no longer be another…

Microsoft’s new AI ‘code of conduct’ instructs models not to hack systems or deceive humans

By Editor-In-ChiefSeptember 14, 2026

As the AI ​​world shifts focus to safety and alignment, Microsoft has…

Fashion app Daydream uses Apple Intelligence to help you shop for outfits in your camera roll

By Editor-In-ChiefSeptember 14, 2026

Daydream, the AI-powered fashion discovery app, on Monday launched two new features…

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Welcome to WhistleBuzz.com (“we,” “our,” or “us”). Your privacy is important to us. This Privacy Policy explains how we collect, use, disclose, and safeguard your information when you visit our website https://whistlebuzz.com/ (the “Site”). Please read this policy carefully to understand our views and practices regarding your personal data and how we will treat it.

Facebook X (Twitter) Instagram Pinterest YouTube

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Facebook X (Twitter) Instagram Pinterest
  • Home
  • Advertise With Us
  • Contact US
  • DMCA Policy
  • Privacy Policy
  • Terms & Conditions
  • About US
© 2026 whistlebuzz. Designed by whistlebuzz.

Type above and press Enter to search. Press Esc to cancel.