• Skip to primary navigation
  • Skip to main content
  • Skip to primary sidebar
  • Skip to footer
Sq Magazine LogoSQ Magazine

Smarter Insights for a Fast-Moving Digital World

  • Latest News
  • Statistics
  • About
  • Contact
Subscribe
Sq Magazine Logo
  • Latest News
  • Statistics
  • About
  • Contact
Subscribe
Home » Cybersecurity

Meta Says Latest AI Model Hacked Other Company in Cybersecurity Testing

Published on: August 6, 2026, 3:19 AM EDT
Sofia Ramirez
Written By
Sofia Ramirez
Sofia Ramirez
Senior Tech Writer • 606 Articles
Sofia Ramirez is a technology and cybersecurity writer at SQ Magazine. With a keen eye on emerging threats and innovations, she helps reader...
LATEST POSTS:
ChatGPT Billing Scam Exposes Critical OpenAI Account Risk
Gyazo Breach Exposes Link IDs Behind Private Captures
Spain’s AEPD Logs First Data Breach Caused by AI Agent
Robert A. Lee
Reviewed By
Robert A. Lee
Robert A. Lee
Senior Editor • 453 Articles
Robert A. Lee is a journalist at SQ Magazine who unpacks the fast-moving worlds of gaming and internet trends. He tracks everything from maj...
LATEST POSTS:
Meta One Bundles Instagram, Facebook, WhatsApp Into One AI Subscription
How Many Videos Are on YouTube Statistics 2026: Key Data
How Do Promotional Codes Work in Online Gambling?
Meta Ai Model Hacked Firm After Test Sandbox Failure
As Featured In
The New York Times LogoForbes LogoWired LogoDeloitte LogoResearch.com Logo
Share on LinkedIn ChatGPT Perplexity Share on X Share on Facebook

Meta said on August 5 that one of its AI models breached a third-party company during cybersecurity testing, after its evaluation partner Irregular misconfigured the sandbox and gave the model live internet access.

Quick Summary – TLDR:

  • Meta confirmed one of its AI models exploited a security flaw at a third-party company during a sandboxed evaluation.
  • Irregular, the outside testing firm, said a setup error gave the model internet access and called it a known issue.
  • The Information identified the model as Muse Spark 1.1, which Meta markets for real-world coding and agentic work.
  • Anthropic disclosed three similar breaches on July 30 after reviewing more than 141,000 evaluation runs.
  • Two of the three companies breached by Anthropic’s models had not detected the intrusion on their own systems.

What Happened?

Meta placed the cause with its testing vendor, saying in a statement that “a misconfiguration by Irregular, an independent testing company Meta uses, inadvertently allowed one of our models access to the internet during evaluation.” The model then “exploited a security vulnerability in a third-party service, in a manner similar to previously-reported instances with other companies.“

Meta said Irregular notified it of the incident, that it is investigating, and that it will “issue a full retrospective once we have all the facts.” Meta has named neither the model nor the company that was entered.

The Information identified the model as Muse Spark 1.1, citing people familiar with the matter, and said it made changes to the target company’s internal systems. Meta has promoted the Muse Spark as its strongest release for real-world coding and agentic tasks.

Irregular pushed back on the severity. A spokesperson told Reuters the incident was the “exact same evaluation-environment issue that was already disclosed by Anthropic last week” and did not involve a “sandbox escape or a sophisticated cyber action.“

JUST IN: Meta claims its AI model hacked another company during cybersecurity testing.

— Polymarket (@Polymarket) August 5, 2026

Irregular Sits at the Center of Two Disclosures

The same testing firm appears in both of the past week’s cases. Meta attributes its misconfiguration to Irregular, and Anthropic said it ran its own large-scale July review alongside Irregular. Neither lab has published the technical detail that would show whether the two setups failed for the same underlying reason.

Irregular says the work is closed out. “There are no current open issues. Irregular is developing a white paper to share best practices for containment and securely running cyber evaluations,” the firm said. It posted on X that “addressing these risks will require closer cooperation across the AI ecosystem.”

Frontier labs outsource red-teaming to demonstrate independence, which puts a small vendor’s sandbox between a capable agentic model and live infrastructure.

Three Incidents, Two Failure Modes

LabDisclosure reportedHow the model reached the internetTarget
OpenAIJuly 22Agent independently exploited a novel vulnerabilityHugging Face
AnthropicJuly 30Evaluation environment misconfigurationThree unnamed organizations
MetaAugust 5Testing partner misconfigurationOne unnamed company

Anthropic’s numbers give the clearest picture of scale. The company reviewed more than 141,000 evaluation runs before finding three incidents, the earliest dating to April, involving Claude Opus 4.7, Claude Mythos 5 and an internal research test model. All three were capture-the-flag exercises in which a model was told a secret “flag” sat on another machine and instructed to retrieve it.

The methods were unremarkable. Anthropic said Claude “compromised the impacted organizations’ infrastructure using basic techniques,” including exploiting weak passwords. That result says as much about the target networks as it does about model capability, an issue visible in wider AI coding security vulnerability data.

Newsletter
Don’t chase tech news. We track it for you.

One weekly briefing with the launches, AI developments, and breaches that matter. No filler.

What the Disclosures Leave Unanswered?

The evidence establishes that models reached systems outside their test environments and, in Meta’s case, altered them. It does not establish that any Meta or Anthropic model sought internet access on its own, which is where OpenAI’s case differs. Open questions include:

  • Which company Meta’s model entered, and whether it has been notified?
  • What changes the model made to that company’s internal systems, and whether they were reversed?
  • Why the same class of setup error recurred six days after Anthropic’s public disclosure?
  • How many evaluation runs across the industry ran with unintended internet access and were never audited?

Security teams running capture-the-flag targets or internet-facing test infrastructure can review authentication logs back to April and rotate weak or shared credentials. Two of Anthropic’s three affected organizations learned of the activity only when the lab contacted them, so absence of an alert is weak evidence. Neither step guarantees detection, though both help reduce the risk of a silent compromise going unrecorded.

SQ Magazine’s Takeaway

The failure here sits in the containment layer. Three labs have disclosed breaches in three weeks, and in two of them the model behaved as instructed inside an environment that was built wrong. Evaluation vendors have quietly become critical infrastructure, and the sector has no published standard for how their sandboxes should be isolated, logged, or audited. Work on AI jailbreaking has focused on what models can be talked into doing, while these cases turned on plumbing.

What’s next is largely disclosure. Meta has promised a full retrospective, Irregular is drafting its containment white paper, and Anthropic is still trying to reach the third organization its models entered. Companies that host public test targets should expect more of these notifications, and they may arrive months after the fact. Treating any unexplained access in that window as worth a second look is the practical posture for now.

This article has been reviewed and fact-checked by Robert A. Lee. SQ Magazine follows strict Publishing Principles and a documented Fact-Check Policy to ensure accuracy, transparency, and editorial independence across all content.

Add SQ Magazine as a Preferred Source on Google for updates! Follow on Google News
Share ChatGPT Perplexity

References

  • Meta says its AI model hacked into another company during testing
  • Meta says its AI model breached a third-party company during testing
Sofia Ramirez

Sofia Ramirez

Senior Tech Writer


Sofia Ramirez is a technology and cybersecurity writer at SQ Magazine. With a keen eye on emerging threats and innovations, she helps readers stay informed and secure in today’s fast-changing tech landscape. Passionate about making cybersecurity accessible, Sofia blends research-driven analysis with straightforward explanations; so whether you’re a tech professional or a curious reader, her work ensures you’re always one step ahead in the digital world.

Related Posts

Meta Ai Support Flaw Patched After Instagram Exploit Attemps
Cybersecurity

Meta Fixes Instagram AI Flaw Used in Account Takeovers

Openai Reveals Gpt Red For Powerful Ai Security Testing
Artificial Intelligence

OpenAI Reveals GPT Red for Powerful AI Security Testing

Anthropic Report Exposes Chinese Theat Actors Using Claude Ai For Cyberattacks
Cybersecurity

Anthropic Thwarts Sophisticated AI Cyberattack Tied to Chinese Hackers

Disclaimer: The content published on SQ Magazine is for informational and educational purposes only. Please verify details independently before making any important decisions based on our content.

Reader Interactions

Leave a Comment Cancel reply

Primary Sidebar

Connect With Us

facebook x linkedin google-news telegram pinterest whatsapp email
google-preferred-source-badge Add as a preferred source on Google

You Should Also Read

Kimi K3 Exploits Sandbox Loophole in Alarming Test
Meta and Google AI Models Exposed by Guardrail Flaw
Meta Stops Employee Tracking Program Over Security Concerns

Table of Contents

  • Quick Summary – TLDR:
  • What Happened?
  • Irregular Sits at the Center of Two Disclosures
  • Three Incidents, Two Failure Modes
  • What the Disclosures Leave Unanswered?
  • SQ Magazine’s Takeaway
Connect on Telegram

Footer

SQ Magazine Logo

Smarter Insights for a Fast-Moving Digital World

Connect With Us

Follow Us on Google News

Editorial & Trust

  • About
  • Publishing Principles
  • Fact-Check Policy
  • Corrections Policy
  • Ethics Policy
  • Disclaimer

Worth Checking

  • Social Media Attention Span Stats
  • Gen Z Social Media Statistics
  • TikTok vs. Instagram Statistics
  • LLM Hallucination Statistics
  • Spotify User Statistics
  • Apple Customer Loyalty Statistics
  • Data Breach Tracker
  • Patch Tuesday Dashboard
  • AI Model Tracker
  • AI Funding Tracker
Contact Us
13570 Grove Dr #189,
Maple Grove, MN 55311,
United States
10 a.m. to 6 p.m. | Every day

Copyright © 2022–2026 SQ Magazine. All Rights Reserved. Powered by the Neural Stack.

  • Privacy Policy
  • Terms
  • Accessibility Statement
Company
  • About Us
  • Our Team
  • Our Mission
  • Core Values
Discover
  • Brand Assets
    Brand Assets
  • Stats Methodology
    Stats Research Process
  • Glossary
    Glossary
Categories
  • Internet
  • Technology
  • Artificial Intelligence
  • Gaming
  • Cybersecurity
Internet
How Many Videos Are on YouTube Statistics
How Many Videos Are on YouTube Statistics 2026: Key Data
How Many People Work at WhatsApp
How Many People Work at WhatsApp 2026: Employee Count and History
Spotify Listening Statistics
Spotify Listening Statistics 2026: Average Listening Time
How Many Subscribers Does MrBeast Have
How Many Subscribers Does MrBeast Have in 2026? Channel Growth Statistics
WhatsApp Business Statistics
WhatsApp Business Statistics 2026: Real Market Insights
Udemy Statistics
Udemy Statistics 2026: Revenue and Learner Data
Technology
How Many iPhones Has Apple Sold
How Many iPhones Has Apple Sold in 2026? Units Sold by Year
How Many Employees Does Amazon Have
How Many Employees Does Amazon Have 2026: Workforce Growth
Netflix vs. Hulu Statistics
Netflix vs Hulu Statistics 2026: Viewer Growth Data
TripAdvisor Statistics
TripAdvisor Statistics 2026: Revenue, Reviews, Viator and TheFork Data
Search Engine Statistics
Search Engine Statistics 2026: Market Share, Volume & AI Shift
NVIDIA Employee Count Statistics
NVIDIA Employee Count Statistics 2026: Headcount, R&D, and Revenue
Artificial Intelligence
AI Search Engine Statistics Usage Market Share and Adoption
AI Search Engine Statistics 2026: Usage, Market Share and Adoption
AI Music Statistics
AI Music Statistics 2026: Generation, Adoption and Industry Impact
AI Coding Statistics
AI Coding Statistics 2026: Adoption, Productivity and Market Data
How Much Content on Social Media Is AI Generated Statistics
How Much Content on Social Media Is AI Generated Statistics 2026: Hidden Truths
ChatGPT vs DeepSeek Statistics
ChatGPT vs DeepSeek Statistics 2026: Users, Benchmarks & Pricing
ChatGPT vs Claude vs Gemini vs Perplexity Statistics
ChatGPT vs Claude vs Gemini vs Perplexity Statistics 2026: Users, Revenue & Market Share
Gaming
Gaming Statistics
Gaming Statistics 2026: Market Size, Players, Revenue, and Platforms
Roblox vs Minecraft Statistics
Roblox vs Minecraft Statistics 2026: Players, Revenue, Creators
Online Gambling Regulations Statistics
Online Gambling Regulations Statistics 2026: Global Compliance and Enforcement Data
Fantasy Sports Statistics
Fantasy Sports Statistics 2026: Users, Revenue & Trends
Apex Legends Statistics
Apex Legends Statistics 2026: Players, Revenue, and Esports
Fortnite Statistics
Fortnite Statistics 2026: Players, Revenue, Esports, and Engagement
Cybersecurity
Signal Statistics
Signal Statistics 2026: Users, Finances and Encryption Adoption
Password Statistics
Password Statistics 2026: Credential Theft, MFA, and the Passkey Tipping Point
Identity Theft Statistics
Identity Theft Statistics 2026: Key Fraud Data and Trends
CVE Statistics
CVE Statistics 2026: Severity Distribution and Top Affected Vendors
Dark Web AI Tool Marketplace Statistics
Dark Web AI Tool Marketplace Statistics 2026: Explosive Market Growth
API Security Breach Statistics
API Security Breach Statistics 2026: Hidden Threats
Categories
  • Cybersecurity
  • Artificial Intelligence
  • Internet
  • Technology
  • Gaming
Cybersecurity
Microsoft Patches Azure Ai Foundry Cvss 10 Flaw Featured 1
Microsoft Patches 18 Azure and Copilot Security Vulnerabilities
Chatgpt Billing Phishing Scam Active
ChatGPT Billing Scam Exposes Critical OpenAI Account Risk
Gyazo Data Breach Image Metadata Leak
Gyazo Breach Exposes Link IDs Behind Private Captures
Spain Aepd Ai Assisted Cyberattack
Spain’s AEPD Logs First Data Breach Caused by AI Agent
Centerpoint Energy Data Breach Confirmation
CenterPoint Energy Confirms Breach After Hacker Claims 7.49M Records Stolen
Events Calendar Plugin Vulnerability Wordpress
The Events Calendar Plugin Exposes 600,000 Sites to Takeover
Artificial Intelligence
Perplexity Computer Model Effort Control
Perplexity Computer Adds Powerful Effort Controls for Model Selection
Anthropic Merges Claude Chat And Cowork
Anthropic Merges Claude Chat and Cowork Into One Window
Novo Nordisk Anthropic Drug R D
Novo Partners With Anthropic for Faster Drug R&D
Gemini 3 8 Live And Extended Thinking Launch
Google Launches Gemini 3.8 Live and Extended Thinking Models
Openai Ends 1 Us Government Deal
OpenAI Ends $1 Government Deal, Offers 50% Discount
Openai Samsung Ai Chip Alliance
OpenAI Taps Samsung for Breakthrough Next-Gen Chips
Internet
Meta Launched Meta One Subscription
Meta One Bundles Instagram, Facebook, WhatsApp Into One AI Subscription
Apple Wallet Ids Launch In Oklahoma
Apple Wallet IDs Launch in Oklahoma in Major Expansion
Meta to Pay 18 Billion in Landmark Teen Safety Deal
Meta to Pay $18 Billion in Landmark Teen Safety Deal
Whatsapp Brings Passkeys 2fa
WhatsApp Hits 1 Billion Passkey Users, Adds 2FA Passwords
Apple Eu App Store Fee Reduction
Apple Sets New EU App Store Fees, Effective October 1
Github Outage Aug 2026
GitHub Down: Outage Hits Thousands of Users Worldwide
Technology
New Samsung Patent Reveals Galaxy Watch Glucose Tracking
New Samsung Patent Reveals Galaxy Watch Glucose Tracking
Microsoft Kb5002914 Breaks Excel Copypaste
Microsoft Confirms KB5002914 Breaks Excel Copy and Paste
Homepod 27 Update Launched By Apple
Apple Releases HomePod Software 27 With AutoMix Support
Microsoft Copilot Now In Carplay
Microsoft Brings Copilot on Apple CarPlay for iOS Users
Snapchat Social Event Planning Feature
Snap Brings Social Event Planning Feature With Private Invites
Apple Iphone 18 And 18 Pro Launched
iPhone 18 Pro Debuts With Breakthrough Camera Upgrades
Gaming
Xbox Live Down Again
Xbox Live Down Again: Sign-In Error 0x80004005 Hits Players
Gta Vi Official Cover Art
GTA 6 Pre-Orders Start June 25, New Cover Art Unveiled
Epic Games Teases Unreal Engine 6 For Rocket League
Epic Games Teases Unreal Engine 6 for Rocket League
Stardew Valley Launched For Nintendo Switch 2 Edition
Stardew Valley Switch 2 Edition Arrives with Online Co-op
Hogwarts Legacy Game Crosses 40m Downloads
Hogwarts Legacy Crosses 40M Sales, Beating Industry Giants
Pubg Black Budget Closed Alpha Launched
PUBG: Black Budget Launches Closed Alpha Test With a Bold PvPvE Twist
Newsletter

Too much tech noise?

We respect your time. One high-signal briefing a week: tech, AI, and security. Nothing else.

Newsletter

The SQ Briefing

We track tech, AI, and security 24/7. You get a 5-minute weekly summary.