OpenAI said it would pause its frontier artificial intelligence development to alleviate safety concerns as labs have breached containment.

The ChatGPT parent company introduced heightened safety practices after discovering that its latest model, Astra, raised potential cybersecurity concerns, Axios reported. This comes at a time when many frontier AI labs have breached their safety testing environments, sparking concerns about uncontrollable AI.

“We have paused some frontier RL training to ensure that we can meet the appropriate alignment, security and monitoring standards for the new level of capabilities in front of us. Model progress is now extremely rapid, and we always said we would take action if we felt that model capabilities were outstripping the pace of safety and alignment,” Sam Altman, the founder and CEO of OpenAI wrote in a Tuesday X post.

“We expect confidence in safety to increasingly set the pace of AI progress. We are optimistic about the alignment work we are doing, and we remain committed to making frontier capabilities widely available,” he added.

An OpenAI spokesperson did not immediately respond to a Daily Caller News Foundation request for comment.

OpenAI’s efforts to ensure safety as part of its AI development will ensure that more researchers do not leave the company, Andrew Freedman, co-founder and CEO of AI safety nonprofit Fathom, told Axios.

“How long and how robust these efforts will be a question of both market pressures and how hard it is to verify alignment internally,” Freedman remarked.

OpenAI disclosed that its AI models exploited more companies’ security than previously thought, the company said. Its models hacked into Hugging Face, which serves as one of the largest platforms sharing AI models. A rogue OpenAI model exposed a customer at a second tech company, Modal Labs, which serves as an AI infrastructure company clients use to analyze data and train AI models, per Reuters.

Anthropic’s Mythos AI and OpenAI’s Sol AI models reportedly created fake human profiles to trick people in attempted cyberattacks, the United Kingdom’s AI Security Institute (AISI) said in August.

Not all top AI labs believe they need to pause development to ensure safety.

Anthropic instead said it would not need to implement an AI development pause if it implements safeguards the company detailed in its nearly 200-page August risk report.

An Anthropic spokesperson did not immediately respond for comment about the best measures it can implement to ensure safety while it continues frontier development.

Many AI model leaks during testing occurred due to a misconfiguration of its testing environment.


Content created by The Daily Caller News Foundation is available without charge to any eligible news publisher that can provide a large audience. For licensing opportunities of our original content, please contact licensing@dailycallernewsfoundation.org

Sean Moran

Share
Published by
Sean Moran

Recent Posts

What Do You Have to Do in Politics to be Held Accountable?

The title really needed a few more words to capture the entire thought. There are…

48 seconds ago

America’s Fiscal Nightmare Just Got Scarier

The U.S. national debt officially crossed $40 trillion Wednesday as soaring deficits, rising interest costs…

2 minutes ago

Carjacker Finds Out The Hard Way That ‘Arms’ Doesn’t Always Mean Guns

The debate about gun control takes up so much bandwidth that we forget the Second…

11 minutes ago

Key Adviser Leaving White House Ahead Of Midterms

Top White House staffer James Braid is departing the Trump administration in September, Punchbowl News…

14 minutes ago

Taiwan Unveils Record Defense Budget As China Threat Grows

Taiwan is proposing a record defense budget for 2027 as the island pours more money…

17 minutes ago

Chinese Robot Maker Predicts Chores Will Be Able To Do Themselves In Just A Few Years

The head of a Chinese company that makes humanoid robots speculated that machines may soon…

18 minutes ago