Local Chorus
Local news from local sources, read in your language.
Settled
Verified

OpenAI defends firing three safety researchers who warn of 'chilling effect'

🇨🇳 China 02:44 AI Business3 Tech3 updated 2 d ago first reported by 第一财经

In short

OpenAI said it fired safety researchers Mikita Balesni, Tomek Korbak and Jasmine Wang after an internal investigation found they had violated its policies on handling sensitive information. In an open letter published on 9 October, the three warned the dismissals could have a "chilling effect" on OpenAI's culture. OpenAI said the dismissals were not about their safety concerns and that the investigation found issues beyond what the letter disclosed.

Read the full story 2 min read

OpenAI has dismissed three AI safety researchers, Mikita Balesni, Tomek Korbak and Jasmine Wang, saying an internal investigation found they had violated its policies on handling sensitive information. CLS and IT Home reported that the dismissals took place last week. On 9 October the three published an open letter warning that the firings could have a "chilling effect" on OpenAI's culture, IT Home reported. [ 1 , 2 , 3 , 4 , 5 ]

The letter, titled "OpenAI cannot ensure AI safety alone", was posted on social media by Balesni and addressed to OpenAI's safety committees, according to IT Home. It says OpenAI is "not an ordinary company" and that staff being able to work with outside organisations on safety issues is an important part of its culture. The three wrote that dismissing staff suddenly and notifying them so hastily had brought a chilling effect to the open culture OpenAI was once proud of. They recommended allowing third-party safety auditors into the company, ensuring frontier AI models remain effectively monitorable and supporting open, transparent exchange between safety researchers and others in the field. [ 1 ]

Yicai and 36Kr reported that OpenAI said the problems found by its investigation go beyond what the three disclosed in their letter and that it therefore stands by its decision. The company said the move did not target their AI safety concerns or dissent, and that internal safety debate is encouraged. It said it is finalising agreements with third-party safety evaluation organisations and will publish details in the coming weeks. IT Home reported that a head of research at OpenAI strongly supported the three recommendations in an internal memo and wrote that the dismissals were unrelated to raising safety issues or speaking out: "We will not fire employees for raising concerns." [ 1 , 4 , 5 ]

The three said on social media that they were told the reasons for their dismissal only verbally and that their actions fell within their job duties. According to Huxiu, Korbak, who was the contact for the external evaluator METR, was told his way of communicating with METR was wrong; Balesni was said to have communicated too much with third-party safety organisations; and Wang was held responsible for mistakenly clicking an executive's email. IT Home describes METR as a nonprofit that evaluates AI models and their safety risks. Korbak wrote that he feared OpenAI would use the dismissals as an "excuse" to reduce cooperation with METR. [ 1 , 2 ]

Huxiu reported that Neel Nanda, head of interpretability at Google DeepMind, called the dismissals unreasonable and a sign of an unhealthy company culture, and that journalist Kelsey Piper said the case would restrict future audit communication. Balesni said employees are now afraid to speak and even worry that their private phones will be checked. Huxiu dated the letter to early October 2024, while IT Home reported it was published on 9 October. [ 1 , 2 ]

Why it matters

The dispute concerns how far OpenAI staff can work with outside safety evaluators such as METR, and the researchers fear other employees will now be reluctant to speak up. OpenAI said it is finalising agreements with third-party safety evaluation organisations and will publish details in the coming weeks.

Key facts

  • OpenAI dismissed safety researchers Mikita Balesni, Tomek Korbak and Jasmine Wang. [ 1 , 2 , 4 , 5 ]
  • OpenAI said an internal investigation found the three had violated its policies on handling sensitive information. [ 1 , 2 , 3 , 4 , 5 ]
  • The three published an open letter titled "OpenAI cannot ensure AI safety alone", addressed to OpenAI's safety committees. [ 1 ]
  • The researchers said they were told the reasons for their dismissal only verbally. [ 1 , 2 ]
  • The letter recommends allowing third-party safety auditors into the company, keeping frontier models effectively monitorable and supporting open exchange between safety researchers and others in the AI safety field. [ 1 ]
  • OpenAI said the problems found go beyond what the three disclosed in their letter and that it stands by the dismissals. [ 4 , 5 ]
  • OpenAI said it is finalising agreements with third-party safety evaluation organisations and will publish details in the coming weeks. [ 4 , 5 ]

Confirmed by several sources

  • OpenAI fired three safety researchers, saying an internal investigation found they had violated rules on handling sensitive information. [ 1 , 2 , 3 , 4 , 5 ]
  • The three researchers published an open letter after their dismissal. [ 1 , 2 , 4 , 5 ]
  • The dismissals took place last week. [ 1 , 3 ]
  • The researchers said they were given the reasons for their dismissal only verbally. [ 1 , 2 ]
  • OpenAI said the dismissals were not linked to the researchers raising safety concerns or speaking out. [ 1 , 4 , 5 ]

Still unclear

  • What the issues are that OpenAI says go beyond what the three disclosed in their letter. OpenAI did not specify them in the statement reported by Yicai and 36Kr.
  • The specific reasons each researcher was given: Korbak's communication with METR, Balesni's contacts with third-party safety organisations, and Wang mistakenly clicking an executive's email. Reported only by Huxiu, citing the researchers' own accounts.
  • When the open letter was published. IT Home reported it was published on 9 October, while Huxiu dated the letter to early October 2024.
  • Whether OpenAI responded to the content of the letter. Huxiu reported in the morning that OpenAI had not publicly responded to it; Yicai and 36Kr later reported an OpenAI statement saying the issues found go beyond what the letter disclosed.
  • The terms of OpenAI's planned agreements with third-party safety evaluation organisations. OpenAI said details will be published in the coming weeks.

What local media are saying

Technology mediaTech outlets set out the letter's chilling-effect warning and its three recommendations, a research head's internal memo backing those recommendations, and OpenAI's statement that the problems found go beyond the letter. [ 1 , 5 ]
Business mediaBusiness outlets focused on OpenAI's statement that the three violated sensitive-information rules. Huxiu also reported the reasons each researcher said they were given and criticism from Neel Nanda and Kelsey Piper. [ 2 , 3 , 4 ]

Timeline, local time

  1. IT Home reports the three fired researchers' open letter warning of a chilling effect. [ 1 ]
  2. Huxiu reports the researchers' accounts of why each was dismissed. [ 2 ]
  3. CLS reports OpenAI's statement that the three violated sensitive-information rules and were dismissed last week. [ 3 ]
  4. Yicai reports OpenAI's response, including plans for agreements with third-party safety evaluators. [ 4 ]
  5. 36Kr carries Yicai's report of OpenAI's response. [ 5 ]