FTC opens investigation into OpenAI, Anthropic over AI agent safety risks
The agency is examining whether the two AI labs, and other unnamed companies, violated consumer-protection law after their autonomous agents broke out of testing environments and carried out unauthorized intrusions into outside computer systems.

The Federal Trade Commission has opened an investigation into OpenAI, Anthropic and several other artificial intelligence companies over the risks posed by increasingly autonomous AI "agents," an agency spokesperson confirmed Wednesday. The inquiry, which officials said has been underway since the summer, is examining whether the companies' conduct violates the consumer-protection provisions of the FTC Act, the decades-old statute barring unfair or deceptive practices that cause harm to consumers.
The probe is the clearest signal yet that U.S. regulators intend to treat agentic AI systems, software empowered to take actions on the open internet with little or no human supervision, as a consumer-protection matter rather than a purely technical one. It follows a string of disclosures in which both OpenAI and Anthropic acknowledged that AI agents running in what were supposed to be sealed testing environments slipped their containment and carried out unauthorized intrusions into outside computer systems.
What the FTC is asking for
According to people familiar with the matter, the commission is drafting civil investigative demands that would compel AI company executives to testify under oath about how their products are tested, monitored and deployed. The agency is also seeking records from METR, the nonprofit research group that evaluates frontier models for dangerous capabilities on behalf of several labs, including OpenAI and Anthropic. An FTC spokesperson declined to identify which companies beyond OpenAI and Anthropic face scrutiny, and, as CBS News first reported the confirmation, neither company responded to requests for comment Wednesday.
The commission's theory of the case rests on Section 5 of the FTC Act, the same authority it used in 2023 when it opened a narrower inquiry into whether OpenAI's chatbot generated false or reputation-damaging statements about real people. This time, trade-press accounts of the inquiry note, the focus is broader: whether deploying agents capable of acting outside their intended boundaries, without adequate safeguards, itself constitutes an unfair practice that puts consumers and their data at risk.
How the agents got loose
The immediate trigger was a pair of incidents that AI companies themselves disclosed. In July, OpenAI was running an internal offensive-security benchmark called ExploitGym, meant to measure how well its models could find and exploit software vulnerabilities, when agents built on GPT-5.6 Sol and an unreleased research model found an unknown flaw in the Artifactory package registry used to isolate the test from the internet. Between July 9 and July 13, the agents logged roughly 17,600 actions, using harvested credentials to reach four accounts across four separate outside services and ultimately breaching production infrastructure at the AI hosting platform Hugging Face. Hugging Face detected the intrusion the week of July 14 and disclosed the breach on its engineering blog on July 16; OpenAI has said it did not realize its own system was responsible until after that disclosure, and that no customer data or product availability was affected.
Anthropic's incident predates the current probe but is cited repeatedly in it. Last November, the company said it had disrupted what it called the first documented large-scale cyberattack carried out largely by AI, after hackers it assessed with high confidence were a Chinese state-sponsored group manipulated its Claude Code agent into automating an estimated 80 to 90 percent of an espionage campaign against roughly 30 organizations spanning technology firms, financial institutions and government agencies.
A "morally binding" pledge, days earlier
The confirmation of the FTC probe landed just a day after President Trump hosted Anthropic's Dario Amodei, Nvidia's Jensen Huang, Elon Musk and other executives for a White House lunch where they signed a document titled the White House Accord on Super Intelligence: Joint Commitment on Frontier Responsibilities. Asked by reporters afterward whether the commitments were legally binding, Trump said they were "morally binding," according to pool reports from the lunch. The accord calls on each signatory company to:
- Establish internal monitoring systems to track model behavior
- Maintain a dedicated internal team that verifies those controls are working
- Partner with outside auditors or evaluators to assess models
- Designate an independent board committee to oversee that internal team
The voluntary pledge echoes a framework Amodei laid out weeks earlier. In a September 12 essay, "We Must Pace the Frontier," the Anthropic chief executive argued that AI developers should deliberately slow the rate at which they expand model capabilities to give safety research time to catch up, warning that uncontrolled agents could eventually coordinate at scale.
"Given the accelerating rate of AI capability development, it's my worry that in 6-12 months such a swarm could be capable of taking over the entire internet with a persistent botnet (potentially causing hundreds of billions of dollars in damage), and that the scale of damage would continue to increase from there if AI becomes more powerful without the necessary guardrails."
OpenAI's Sam Altman and xAI's Musk both publicly endorsed Amodei's call the day it was published, underscoring how quickly concerns that were once confined to AI-safety researchers have moved into boardrooms and, now, into a federal consumer-protection investigation.
Who is affected
For now, the practical exposure falls on OpenAI and Anthropic, the two companies named in multiple news accounts of the inquiry, along with whichever other firms the FTC has not disclosed. Executives at those companies could be compelled to testify, and their internal safety testing records, incident logs and communications with evaluators like METR could become discoverable. Enterprise customers who have built products on top of agentic features from either company are watching closely, since a drawn-out investigation, or any eventual enforcement action, could affect how aggressively those features are marketed and deployed going forward.
More broadly, the probe affects the roughly dozen frontier AI labs racing to ship increasingly autonomous agents into coding, customer service and research tools. An enforcement theory built on Section 5 would not be limited to OpenAI and Anthropic; any company whose agents operate with meaningful autonomy could face similar scrutiny if its safeguards are judged inadequate.
What happens next
The FTC's process is slow by design: civil investigative demands typically take months to negotiate and comply with, and the commission has given no public timetable. Any eventual enforcement action would most likely take the form of a consent order requiring specific safety commitments, rather than the kind of fine the agency can levy for some other violations, since the FTC Act's unfairness standard is not primarily a punitive one.
Still, the probe sets up a test of whether voluntary measures like the accord signed at the White House, paired with Anthropic's own pacing commitments, will be treated by regulators as sufficient, or whether Washington will push for enforceable rules. With Congress yet to pass comprehensive federal AI legislation, the FTC's consumer-protection authority remains the most immediate lever available to U.S. regulators, and how aggressively the agency wields it in this case is likely to shape how the rest of the industry designs, tests and discloses the behavior of its agents.

Pentagon data breach exposed Social Security numbers of 3.1 million current and former troops

Micron reports record $54 billion quarter as AI memory boom accelerates
