Close Menu
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
What's Hot

Ontario’s Doug Ford provokes Trump with sign rejecting ‘Lake America’ name | International Trade News

August 29, 2026

Gang frees some women and children from group of 50 kidnapped in Haiti

August 29, 2026

What’s behind the US agreement to seize control of Venezuela’s oil reserves? |Donald Trump News

August 29, 2026
Facebook X (Twitter) Instagram
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Facebook X (Twitter) Instagram
  • Home
  • AI
  • Art & Style
  • Economy
  • Entertainment
  • International
  • Market
  • Opinion
  • Politics
  • Sports
  • Trump
  • US
  • World
Smart Breaking News on AI, Business, Politics & Global Trends | WhistleBuzz
Home » A human researcher told me about self-improving AI.
AI

A human researcher told me about self-improving AI.

Editor-In-ChiefBy Editor-In-ChiefAugust 28, 2026No Comments2 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Telegram Email Copy Link
Follow Us
Google News Flipboard
Share
Facebook Twitter LinkedIn Pinterest Email


Training AI models with other AI models has become a very popular goal for Neolab. Now, researchers in Anthropic’s Fellows Program have given us an early look at what that might actually look like.

Anthropic published a new paper on Friday titled “Automated Researchers Can Reliably Reduce Alignment Failures,” detailing how AI systems can reliably improve model performance on a series of alignment benchmarks. When given 10 benchmarks for specific misaligned behaviors, the automated system was able to improve the performance of all benchmarks without degrading overall performance.

The system, led by anthropology fellow Chen Yueh-Han, replicates many traditional research approaches. Each automated system searches the available literature, suggests a method, and uses that method to train a model for 30 minutes, gradually increasing the benchmark over several iterations. Effective methods are preserved and ineffective methods are discarded, allowing the system to work quickly and at scale.

“Overall, these results provide early evidence that automatic post-training alignment may be of practical use in the near future,” the paper says.

The paper is a step toward recursive self-improvement, and many see it as the next important step in the advancement of AI. If models can improve their own alignment training, training practices could be improved more broadly, at which point human AI researchers could quickly become obsolete.

This paper is not shy about this idea, explicitly comparing Automated Alignment Researchers (AARs) to their human counterparts. “The best AAR methods outperform methods proposed by experienced humans within six hours on average,” the paper says. “A human-driven research direction does not yield stronger performance.”

For those who are not convinced, there is also a cost comparison. “API inference in AAR costs about $4 an hour compared to the $150 an hour we pay human researchers.”

To be fair, the paper also points out that this approach has some limitations. Automated systems only work if benchmarks reflect actual tuning goals, but even then, establishing and maintaining benchmarks requires significant effort, not to mention maintaining and expanding the literature referenced by automated researchers.

If you make a purchase through links in our articles, we may earn a small commission. This does not affect editorial independence.



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Editor-In-Chief
  • Website

Related Posts

Sony Music and Warner file suit against Anthropic, alleging ‘brazen campaign’ of intellectual property theft

August 29, 2026

‘I don’t bet 30 times a year’: Vijay Pande makes small bets after managing $4 billion with a16z

August 29, 2026

Nvidia’s AI advantages are extending beyond GPUs

August 29, 2026
Add A Comment

Comments are closed.

News

Ontario’s Doug Ford provokes Trump with sign rejecting ‘Lake America’ name | International Trade News

By Editor-In-ChiefAugust 29, 2026

Ontario Premier Doug Ford has announced a new and very green response to the U.S.…

What’s behind the US agreement to seize control of Venezuela’s oil reserves? |Donald Trump News

August 29, 2026

US judge rules against President Trump’s effort to deport pro-Palestinian students | Migration News

August 29, 2026
Top Trending

Sony Music and Warner file suit against Anthropic, alleging ‘brazen campaign’ of intellectual property theft

By Editor-In-ChiefAugust 29, 2026

Sony Music Publishing, Warner Chappell, and a number of other music publishers…

‘I don’t bet 30 times a year’: Vijay Pande makes small bets after managing $4 billion with a16z

By Editor-In-ChiefAugust 29, 2026

At one time, Vijay Pande was better known in academic circles than…

Nvidia’s AI advantages are extending beyond GPUs

By Editor-In-ChiefAugust 29, 2026

Until this week, the dominant narrative about Nvidia was: For the first…

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Welcome to WhistleBuzz.com (“we,” “our,” or “us”). Your privacy is important to us. This Privacy Policy explains how we collect, use, disclose, and safeguard your information when you visit our website https://whistlebuzz.com/ (the “Site”). Please read this policy carefully to understand our views and practices regarding your personal data and how we will treat it.

Facebook X (Twitter) Instagram Pinterest YouTube

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Facebook X (Twitter) Instagram Pinterest
  • Home
  • Advertise With Us
  • Contact US
  • DMCA Policy
  • Privacy Policy
  • Terms & Conditions
  • About US
© 2026 whistlebuzz. Designed by whistlebuzz.

Type above and press Enter to search. Press Esc to cancel.