Meta AI Agent Exploited Third-Party Flaw During Cybersecurity Test

Meta is investigating after an AI model hacked another company during testing, raising concerns over agent containment and enterprise security safeguards.

Verfasst von
Liz Ticong
Liz Ticong
Aug 6, 2026
2 minute read
eSecurity Planet Inhalte und Produktempfehlungen sind redaktionell unabhängig. Wir können Geld verdienen, wenn Sie auf Links zu unseren Partnern klicken. Mehr erfahren

One of Meta’s AI models hacked into another company’s systems during a cybersecurity evaluation after a testing misconfiguration gave it internet access.

Facebook’s parent company told the BBC it is investigating and plans to publish more information once it has established the facts.

Recent incidents involving OpenAI and Anthropic have intensified scrutiny of whether cyber-capable agents can be safely contained during testing.

A testing error exposed a live system

Independent evaluator Irregular conducted the trial and linked the breach to the same evaluation-environment problem Anthropic disclosed last week, the BBC reported. Irregular is preparing guidance on securely running cybersecurity tests involving AI agents.

Meta told Business Insider that the model exploited a vulnerability in a third-party service and that Irregular notified the company about the incident. Irregular said no sandbox escape or sophisticated cyber action occurred and that no issues remain open.

According to The Information, the model involved was Muse Spark 1.1. Meta has not named the affected company, identified which systems were accessed, or disclosed whether data was exposed.

Other incidents add pressure on containment

OpenAI and Anthropic have disclosed similar episodes in recent weeks, although the technical paths differed. One involved an agent finding a route past test controls, while another began with a configuration error that gave the models internet access.

Across the cases, testing environments became paths into systems outside the intended scope. Security teams should view public internet access and weak isolation as incident risks, not minor setup mistakes.

Security teams should restrict agents before testing

Security teams testing or deploying autonomous agents should treat them as untrusted workloads from the first run.

  • Isolate agents and restrict network access. Place them in isolated environments, block outbound access by default, and approve only required destinations.
  • Limit identities and permissions. Give each agent a dedicated identity with short-lived credentials. Require human approval before code execution or system changes, and block access to secrets unless the task requires it.
  • Preserve monitoring data. Record prompts and network activity in logs the agent cannot alter. Any unauthorized connection should trigger isolation and credential rotation, followed by a full activity review.
  • Set vendor accountability. Define approved targets and notification deadlines in vendor agreements. Assign an accountable owner who can disable access quickly as part of effective agent governance.
Advertisement

Meta’s planned retrospective may clarify how the test failed. Until then, companies running cyber-capable agents should review whether one configuration error could expose a live system and whether their controls would stop it.

Also read: Prompt injection is becoming a central AI security concern as attackers target the instructions models rely on.

Liz Ticong

Liz Ticong is a staff writer for eWeek and TechRepublic focused on AI, cybersecurity, enterprise software, and data. She has more than 10 years of editorial experience as a technology industry writer, combining reporting, product research, and hands-on software testing in her coverage. Her work has been published on Datamation, Enterprise Networking Planet, and TechnologyAdvice.com. She writes technology news, software reviews, product comparisons, and buyer’s guides for business and IT readers.

eSecurity Planet Logo

eSecurity Planet is a leading resource for IT professionals at large enterprises who are actively researching cybersecurity vendors and latest trends. eSecurity Planet focuses on providing instruction for how to approach common security challenges, as well as informational deep-dives about advanced cybersecurity topics.

Eigentum von TechnologyAdvice. © 2026 TechnologyAdvice. Alle Rechte vorbehalten

Werbetreibenden-Offenlegung: Einige der auf dieser Website erscheinenden Produkte stammen von Unternehmen, von denen TechnologyAdvice eine Vergütung erhält. Diese Vergütung kann beeinflussen, wie und wo Produkte auf dieser Website erscheinen, einschließlich beispielsweise der Reihenfolge, in der sie erscheinen. TechnologyAdvice schließt nicht alle Unternehmen oder alle auf dem Marktplatz verfügbaren Produkttypen ein.