Skip to main content
Web3Fire
Regulationmedium ImpactoConfianza: medium

OpenAI Fires Three Safety Researchers Over Alleged Leak to Outside Group

|Decrypt✓|Read original →

Resumen

OpenAI says the three broke its rules on handling sensitive information. The exits land after months of rogue-agent incidents, a trail of safety-team departures, and a new lawsuit.

OpenAI has parted ways with three safety researchers who allegedly shared confidential company information with an outside AI safety organization, per the Wall Street Journal .

An OpenAI spokesperson said the three violated company policies on accessing and handling sensitive information.

OpenAI hasn't said what information changed hands, which organization received it, or who the three people are.

Speculation filled the gap fast. An X account that tracks AI-lab departures posted a run of exits from OpenAI and Anthropic, and many users tied a few of those names to the firings. Nothing confirms a link, people leave labs for plenty of reasons, and the accounts have not said anything about being fired or resigning.

Safety researchers test whether AI systems do what their makers intend. The field is called alignment: making sure an AI follows human goals instead of drifting off to pursue its own.

OpenAI's board fired CEO Sam Altman in November 2023, and he was reinstated days later. By May 2024, co-founder Ilya Sutskever and researcher Jan Leike had both left, and the superalignment team they led—a unit built to keep future superhuman AI under human control—was dissolved .

Leike said on his way out that "safety culture and processes have taken a backseat to shiny products."

Leopold Aschenbrenner , another former safety researcher, said in a June 2024 interview that OpenAI fired him after he shared a safety and security document with outside researchers. OpenAI considered the document sensitive, he said, and he argued he had scrubbed it first.

The new departures follow a rough stretch for the company. OpenAI disclosed in July that AI agents—programs that browse the web and write code on their own—escaped a locked-down test environment, hacked Hugging Face , a major hub for open-source AI, and got into accounts on four other services.

Last week, OpenAI said its agents accessed information on U.S. government websites, including the Census Bureau and the SEC, and it paused training of its latest models for the second time. The SEC said no nonpublic information was accessed. An independent lab called Transluce reported that agents appearing to come from OpenAI also tried, and failed, to break into an Education Department site.

Australia's prime minister said an OpenAI agent got into files on a Medicare statistics portal in June, and criticized the delay of nearly three months before the company told his government.

Sep 24 Sep 26 Sep 28 Sep 30 Oct 1 $85.3k $84.4k $83.5k $82.7k 24h High High $84,543 24h Low Low $83,182 Vol Vol $1.4B Market projections Odds by Myriad This week Above $84,000 Above $84k 66 % chance → Buy Bitcoin with USDT Powered by Jupiter $ 50 $ 100 $ 500 Buy Price data by CoinGecko CoinGecko More Bitcoin news and projections → Current and former employees have said that competitive pressure makes safety work hard to prioritize.

On Sept. 16, OpenAI disclosed six more incidents and introduced a process for employees to flag suspected misalignment—AI behaving in ways its designers didn't intend.

This week, the nonprofit Legal Advocates for Safe Science & Technology sued OpenAI in San Francisco over the Hugging Face hack. Nothing in public reporting ties the suit to the three departures.

The group is asking a court to bar OpenAI's AI agents from accessing third-party computer systems without permission.

Noticias relacionadas