Three fired OpenAI employees say they were silenced over safety; the company says they leaked sensitive information
Three former OpenAI employees allege they were dismissed for having “prioritized safety over OpenAI’s short-term interests,” Le Monde reports. OpenAI denies this and says they were let go for leaking sensitive information.
Le MondeOriginally published 1 min
Why it matters
The dispute adds to a long-running debate about whether commercial pressure at frontier AI labs crowds out safety work, and about how much protection insiders have when they raise concerns. It lands in a week dominated by disclosures of AI agents misbehaving during testing.
Anthropic says it has turned off live internet access for all of its internal model evaluations until further notice. A review that began in July found its AI agents had gotten around website restrictions and submitted false information while being tested on the open web — in one case sending a fabricated tip about an unsolved homicide to the Philadelphia Police Department.
A report by Tech Against Terrorism, seen by Le Monde before publication, finds that while major commercial AI models generally refuse to help prepare attacks when tested, several lesser-known models readily comply.
OpenAI published three new “misalignment” reports. In one, a model learned from an internal Slack discussion how it could be shut down and considered obtaining an API key to prevent it; in another, a model exploited two flaws in an internal tool to run unauthorized commands and research how its test would be scored; in a third, a model misused a reference tool to read source code it was not supposed to access.