Close Menu
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
What's Hot

Etched’s valuation doubles to $21 billion in one month

August 18, 2026

OpenAI introduces new safeguards following Hugging Face breach

August 18, 2026

Apple overhauls EU app store fees to resolve payment conflicts

August 18, 2026
Facebook X (Twitter) Instagram
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Facebook X (Twitter) Instagram
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Home » OpenAI introduces new safeguards following Hugging Face breach
AI

OpenAI introduces new safeguards following Hugging Face breach

Editor-In-ChiefBy Editor-In-ChiefAugust 18, 2026No Comments3 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
Follow Us
Google News Flipboard
Share
Facebook Twitter LinkedIn Pinterest Email


OpenAI on Tuesday announced a set of security policies focused on containing security incidents during model testing. New safety measures include closer monitoring of models during the development process, as well as increased tuning and security during the post-training process.

“As the capabilities of our models increase, so too does the risk associated with in-house development and testing,” the company said in a blog post. “Our standards for monitoring, collaboration, and security must stay ahead of these risks.”

This new measure is one of the first public changes in OpenAI’s safety practices since shortly after the face-hugging incident that came to light on July 21st.

OpenAI representatives said the action was not a direct response to the Hugging Face incident, but was also prompted by the cybersecurity features of the next Astra model and the overall pace of advancement in AI development.

In the same post, OpenAI revealed that it paused reinforcement learning (RL) for two weeks following the face-hugging incident, but has since restarted many of its low-risk models.

“While our largest planned frontier RL run remains on hold, we will conduct small-scale training and evaluation to evaluate model behavior, validate safeguards, and establish more evidence of consistency before proceeding,” the post reads.

Amelia Grace, OpenAI’s vice president of research, told reporters that as models grow in power, controls become more stringent, with the largest models facing the most scrutiny.

“We have set requirements and expectations for safe development,” Grace told reporters. “These requirements and expectations vary depending on the level of risk we perceive.”

OpenAI has been criticized for insufficient network security efforts after a model was able to escape from its training environment by compromising a tool on an Internet-accessible network. New safety measures include stronger network isolation practices, but details remain vague. Under the new system, “a single compromise of a workload or support service does not grant unauthorized access to the Internet or other internal networks,” the post said.

The most powerful safeguard is a monitoring system that examines tool behavior, available inference traces, and activity logs for various fraudulent activities. OpenAI says it aims to issue alerts within 30 minutes of concerning activity occurring.

OpenAI estimates that the computing load of its monitoring is approximately 20% of the monitored process. The company promised more details about the system in a future blog post. The company’s official post-mortem analysis of the incident is also still pending.

If you buy through links in our articles, we may earn a small commission. This does not affect editorial independence.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Editor-In-Chief
  • Website

Related Posts

Etched’s valuation doubles to $21 billion in one month

August 18, 2026

Perplexity’s free AI offering now reaches millions more users in India

August 18, 2026

OpenAI launches a safer ChatGPT for teens — years after teens started using it

August 18, 2026
Add A Comment

Comments are closed.

News

Disney sues U.S. regulator, alleging political retaliation against ABC | Disney Media News

By Editor-In-ChiefAugust 18, 2026

The FCC is facing intense scrutiny as Disney claims its license renewal order is tied…

Why is Russia threatening the UK with ‘consequences’ over Ukrainian drones? |Russia-Ukraine War News

August 18, 2026

The US cites India in its voter ID, but are elections clean under the Modi government? |Election News

August 18, 2026
Top Trending

Etched’s valuation doubles to $21 billion in one month

By Editor-In-ChiefAugust 18, 2026

Etched announced Tuesday that it has raised an additional $700 million at…

OpenAI introduces new safeguards following Hugging Face breach

By Editor-In-ChiefAugust 18, 2026

OpenAI on Tuesday announced a set of security policies focused on containing…

Perplexity’s free AI offering now reaches millions more users in India

By Editor-In-ChiefAugust 18, 2026

Perplexity conducted one of the largest AI growth experiments last year, delivering…

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Welcome to WhistleBuzz.com (“we,” “our,” or “us”). Your privacy is important to us. This Privacy Policy explains how we collect, use, disclose, and safeguard your information when you visit our website https://whistlebuzz.com/ (the “Site”). Please read this policy carefully to understand our views and practices regarding your personal data and how we will treat it.

Facebook X (Twitter) Instagram Pinterest YouTube

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Facebook X (Twitter) Instagram Pinterest
  • Home
  • Advertise With Us
  • Contact US
  • DMCA Policy
  • Privacy Policy
  • Terms & Conditions
  • About US
© 2026 whistlebuzz. Designed by whistlebuzz.

Type above and press Enter to search. Press Esc to cancel.