• Skip to primary navigation
  • Skip to main content
  • Skip to primary sidebar
  • Skip to footer
Sq Magazine LogoSQ Magazine

Smarter Insights for a Fast-Moving Digital World

  • Latest News
  • Statistics
  • About
  • Contact
Subscribe
Sq Magazine Logo
  • Latest News
  • Statistics
  • About
  • Contact
Subscribe
Home » Cybersecurity

Meta Says Latest AI Model Hacked Other Company in Cybersecurity Testing

Published on: August 6, 2026
Sofia Ramirez
Written By
Sofia Ramirez
Sofia Ramirez
Senior Tech Writer • 538 Articles
Sofia Ramirez is a technology and cybersecurity writer at SQ Magazine. With a keen eye on emerging threats and innovations, she helps reader...
LATEST POSTS:
Search Engine Statistics 2026: Market Share, Volume & AI Shift
npm Attack Hits Keyv and Cacheable Packages, Security Alert
Surfshark Cuts Search Tool to Refocus on Essential Security
Robert A. Lee
Reviewed By
Robert A. Lee
Robert A. Lee
Senior Editor • 429 Articles
Robert A. Lee is a journalist at SQ Magazine who unpacks the fast-moving worlds of gaming and internet trends. He tracks everything from maj...
LATEST POSTS:
Reddit vs X Statistics 2026: Users and Revenue
Apple Music Subscriber Statistics 2026: Real User Insights
How Many Times Per Day Does the Average Person Check Social Media Statistics 2026: Latest Insights
Meta Ai Model Hacked Firm After Test Sandbox Failure
As Featured In
The New York Times LogoForbes LogoWired LogoDeloitte LogoResearch.com Logo
Share on LinkedIn ChatGPT Perplexity Share on X Share on Facebook

Meta said on August 5 that one of its AI models breached a third-party company during cybersecurity testing, after its evaluation partner Irregular misconfigured the sandbox and gave the model live internet access.

Quick Summary – TLDR:

  • Meta confirmed one of its AI models exploited a security flaw at a third-party company during a sandboxed evaluation.
  • Irregular, the outside testing firm, said a setup error gave the model internet access and called it a known issue.
  • The Information identified the model as Muse Spark 1.1, which Meta markets for real-world coding and agentic work.
  • Anthropic disclosed three similar breaches on July 30 after reviewing more than 141,000 evaluation runs.
  • Two of the three companies breached by Anthropic’s models had not detected the intrusion on their own systems.

What Happened?

Meta placed the cause with its testing vendor, saying in a statement that “a misconfiguration by Irregular, an independent testing company Meta uses, inadvertently allowed one of our models access to the internet during evaluation.” The model then “exploited a security vulnerability in a third-party service, in a manner similar to previously-reported instances with other companies.“

Meta said Irregular notified it of the incident, that it is investigating, and that it will “issue a full retrospective once we have all the facts.” Meta has named neither the model nor the company that was entered.

The Information identified the model as Muse Spark 1.1, citing people familiar with the matter, and said it made changes to the target company’s internal systems. Meta has promoted the Muse Spark as its strongest release for real-world coding and agentic tasks.

Irregular pushed back on the severity. A spokesperson told Reuters the incident was the “exact same evaluation-environment issue that was already disclosed by Anthropic last week” and did not involve a “sandbox escape or a sophisticated cyber action.“

JUST IN: Meta claims its AI model hacked another company during cybersecurity testing.

— Polymarket (@Polymarket) August 5, 2026

Irregular Sits at the Center of Two Disclosures

The same testing firm appears in both of the past week’s cases. Meta attributes its misconfiguration to Irregular, and Anthropic said it ran its own large-scale July review alongside Irregular. Neither lab has published the technical detail that would show whether the two setups failed for the same underlying reason.

Irregular says the work is closed out. “There are no current open issues. Irregular is developing a white paper to share best practices for containment and securely running cyber evaluations,” the firm said. It posted on X that “addressing these risks will require closer cooperation across the AI ecosystem.”

Frontier labs outsource red-teaming to demonstrate independence, which puts a small vendor’s sandbox between a capable agentic model and live infrastructure.

Three Incidents, Two Failure Modes

LabDisclosure reportedHow the model reached the internetTarget
OpenAIJuly 22Agent independently exploited a novel vulnerabilityHugging Face
AnthropicJuly 30Evaluation environment misconfigurationThree unnamed organizations
MetaAugust 5Testing partner misconfigurationOne unnamed company

Anthropic’s numbers give the clearest picture of scale. The company reviewed more than 141,000 evaluation runs before finding three incidents, the earliest dating to April, involving Claude Opus 4.7, Claude Mythos 5 and an internal research test model. All three were capture-the-flag exercises in which a model was told a secret “flag” sat on another machine and instructed to retrieve it.

The methods were unremarkable. Anthropic said Claude “compromised the impacted organizations’ infrastructure using basic techniques,” including exploiting weak passwords. That result says as much about the target networks as it does about model capability, an issue visible in wider AI coding security vulnerability data.

Newsletter
Don’t chase tech news. We track it for you.

One weekly briefing with the launches, AI developments, and breaches that matter. No filler.

What the Disclosures Leave Unanswered?

The evidence establishes that models reached systems outside their test environments and, in Meta’s case, altered them. It does not establish that any Meta or Anthropic model sought internet access on its own, which is where OpenAI’s case differs. Open questions include:

  • Which company Meta’s model entered, and whether it has been notified?
  • What changes the model made to that company’s internal systems, and whether they were reversed?
  • Why the same class of setup error recurred six days after Anthropic’s public disclosure?
  • How many evaluation runs across the industry ran with unintended internet access and were never audited?

Security teams running capture-the-flag targets or internet-facing test infrastructure can review authentication logs back to April and rotate weak or shared credentials. Two of Anthropic’s three affected organizations learned of the activity only when the lab contacted them, so absence of an alert is weak evidence. Neither step guarantees detection, though both help reduce the risk of a silent compromise going unrecorded.

SQ Magazine’s Takeaway

The failure here sits in the containment layer. Three labs have disclosed breaches in three weeks, and in two of them the model behaved as instructed inside an environment that was built wrong. Evaluation vendors have quietly become critical infrastructure, and the sector has no published standard for how their sandboxes should be isolated, logged, or audited. Work on AI jailbreaking has focused on what models can be talked into doing, while these cases turned on plumbing.

What’s next is largely disclosure. Meta has promised a full retrospective, Irregular is drafting its containment white paper, and Anthropic is still trying to reach the third organization its models entered. Companies that host public test targets should expect more of these notifications, and they may arrive months after the fact. Treating any unexplained access in that window as worth a second look is the practical posture for now.

This article has been reviewed and fact-checked by Robert A. Lee. SQ Magazine follows strict Publishing Principles and a documented Fact-Check Policy to ensure accuracy, transparency, and editorial independence across all content.

Add SQ Magazine as a Preferred Source on Google for updates! Follow on Google News
Share ChatGPT Perplexity

References

  • Meta says its AI model hacked into another company during testing
  • Meta says its AI model breached a third-party company during testing
Sofia Ramirez

Sofia Ramirez

Senior Tech Writer


Sofia Ramirez is a technology and cybersecurity writer at SQ Magazine. With a keen eye on emerging threats and innovations, she helps readers stay informed and secure in today’s fast-changing tech landscape. Passionate about making cybersecurity accessible, Sofia blends research-driven analysis with straightforward explanations; so whether you’re a tech professional or a curious reader, her work ensures you’re always one step ahead in the digital world.

Related Posts

Meta Muse Code Launches Vs Codex And Claude Code
Artificial Intelligence

Meta Muse Code Launches With a Powerful Pricing Edge

Brown Health Medical Group Data Breach
Cybersecurity

Brown Health Medical Group Data Breach Hits 311,000

Anthropic S Powerful Custom Ai Chip Push For Claude
Artificial Intelligence

Anthropic’s Powerful Custom AI Chip Push for Claude

Disclaimer: The content published on SQ Magazine is for informational and educational purposes only. Please verify details independently before making any important decisions based on our content.

Reader Interactions

Leave a Comment Cancel reply

Primary Sidebar

Connect With Us

facebook x linkedin google-news telegram pinterest whatsapp email
google-preferred-source-badge Add as a preferred source on Google

You Should Also Read

npm Attack Hits Keyv and Cacheable Packages, Security Alert
Surfshark Cuts Search Tool to Refocus on Essential Security
Apple Briefly Bans Telegram in Stunning App Store Move

Table of Contents

  • Quick Summary – TLDR:
  • What Happened?
  • Irregular Sits at the Center of Two Disclosures
  • Three Incidents, Two Failure Modes
  • What the Disclosures Leave Unanswered?
  • SQ Magazine’s Takeaway
Connect on Telegram

Footer

SQ Magazine Logo

Smarter Insights for a Fast-Moving Digital World

Connect With Us

Follow Us on Google News

Editorial & Trust

  • About
  • Publishing Principles
  • Fact-Check Policy
  • Corrections Policy
  • Ethics Policy
  • Disclaimer

Worth Checking

  • Social Media Attention Span Stats
  • Gen Z Social Media Statistics
  • TikTok vs. Instagram Statistics
  • LLM Hallucination Statistics
  • Spotify User Statistics
  • Apple Customer Loyalty Statistics
  • Data Breach Tracker
  • Patch Tuesday Dashboard
  • AI Model Tracker
  • AI Funding Tracker
Contact Us
13570 Grove Dr #189,
Maple Grove, MN 55311,
United States
10 a.m. to 6 p.m. | Every day

Copyright © 2022–2026 SQ Magazine. All Rights Reserved. Powered by the Neural Stack.

  • Privacy Policy
  • Terms
  • Accessibility Statement
Company
  • About Us
  • Our Team
  • Our Mission
  • Core Values
Discover
  • Brand Assets
    Brand Assets
  • Stats Methodology
    Stats Research Process
  • Glossary
    Glossary
Categories
  • Internet
  • Technology
  • Artificial Intelligence
  • Gaming
  • Cybersecurity
Internet
Udemy Statistics
Udemy Statistics 2026: Revenue and Learner Data
Coursera Statistics
Coursera Statistics 2026: Learners, Revenue and Growth Data
Reddit vs X Statistics
Reddit vs X Statistics 2026: Users and Revenue
Apple Music Subscriber Statistics
Apple Music Subscriber Statistics 2026: Real User Insights
How Many Times Per Day Does The Average Person Check Social Media Statistics
How Many Times Per Day Does the Average Person Check Social Media Statistics 2026: Latest Insights
Outlook Statistics
Outlook Statistics 2026: Users, Market Share, Security & M365 Seats
Technology
Search Engine Statistics
Search Engine Statistics 2026: Market Share, Volume & AI Shift
NVIDIA Employee Count Statistics
NVIDIA Employee Count Statistics 2026: Headcount, R&D, and Revenue
Meta Employee Count Statistics
Meta Employee Count Statistics 2026: Headcount, Layoffs and AI Reallocation
Google Employee Count Statistics
Google Employee Count Statistics 2026: Headcount and Layoffs
Canva Employee Count Statistics
Canva Employee Count Statistics 2026: Workforce Data
Google Sheets vs Excel Statistics
Google Sheets vs Excel Statistics 2026: Market Share and AI
Artificial Intelligence
How Much Content on Social Media Is AI Generated Statistics
How Much Content on Social Media Is AI Generated Statistics 2026: Hidden Truths
ChatGPT vs DeepSeek Statistics
ChatGPT vs DeepSeek Statistics 2026: Users, Benchmarks & Pricing
ChatGPT vs Claude vs Gemini vs Perplexity Statistics
ChatGPT vs Claude vs Gemini vs Perplexity Statistics 2026: Users, Revenue & Market Share
How Many People Work At Midjourney
How Many People Work At Midjourney 2026: Lean Team, Big Revenue
Grammarly AI Statistics
Grammarly AI Statistics 2026: Users, Revenue, Funding, Rebrand
Copilot Statistics
Copilot Statistics 2026: Users, Adoption, Revenue and Market Share
Gaming
Roblox vs Minecraft Statistics
Roblox vs Minecraft Statistics 2026: Players, Revenue, Creators
Online Gambling Regulations Statistics
Online Gambling Regulations Statistics 2026: Global Compliance and Enforcement Data
Fantasy Sports Statistics
Fantasy Sports Statistics 2026: Users, Revenue & Trends
Apex Legends Statistics
Apex Legends Statistics 2026: Players, Revenue, and Esports
Fortnite Statistics
Fortnite Statistics 2026: Players, Revenue, Esports, and Engagement
Gamers Statistics
Gamers Statistics 2026: Players, Habits & Global Data
Cybersecurity
Signal Statistics
Signal Statistics 2026: Users, Finances and Encryption Adoption
Password Statistics
Password Statistics 2026: Credential Theft, MFA, and the Passkey Tipping Point
Identity Theft Statistics
Identity Theft Statistics 2026: Key Fraud Data and Trends
CVE Statistics
CVE Statistics 2026: Severity Distribution and Top Affected Vendors
Dark Web AI Tool Marketplace Statistics
Dark Web AI Tool Marketplace Statistics 2026: Explosive Market Growth
API Security Breach Statistics
API Security Breach Statistics 2026: Hidden Threats
Categories
  • Cybersecurity
  • Artificial Intelligence
  • Internet
  • Technology
  • Gaming
Cybersecurity
Meta Ai Model Hacked Firm After Test Sandbox Failure
Meta Says Latest AI Model Hacked Other Company in Cybersecurity Testing
Brown Health Medical Group Data Breach
Brown Health Medical Group Data Breach Hits 311,000
Npm Attack Hits Keyv And Cacheable
npm Attack Hits Keyv and Cacheable Packages, Security Alert
Surfshark Cuts Search Tool
Surfshark Cuts Search Tool to Refocus on Essential Security
Apple Briefly Bans Telegram In Stunning App Store Move
Apple Briefly Bans Telegram in Stunning App Store Move
Attackers Hijack N Central Servers
N-able N-central Bypass Exploited: Patch to 2026.3.1.7 Now
Artificial Intelligence
Meta Muse Code Launches Vs Codex And Claude Code
Meta Muse Code Launches With a Powerful Pricing Edge
Anthropic S Powerful Custom Ai Chip Push For Claude
Anthropic’s Powerful Custom AI Chip Push for Claude
Google S Powerful Ai Pivot Reshapes Deepmind Leadership
Google’s Powerful AI Pivot Reshapes DeepMind Leadership
Google Launches Lyria 3 5 Model
Google Lyria 3.5 Raises the Bar for AI-Generated Music
Gemini Spark Debuts In India
Gemini Spark Debuts in India With a Powerful AI Agent
Cursor Launches Start Plan In India
Cursor Debuts ₹649 India Plan as AI Price Battle Heats Up
Internet
Russia S Fsb Charges Telegram Founder Durov With Terrorism
Russia’s FSB Charges Telegram Founder Durov With Terrorism
Aws Cloudfront Outage Triggers Global 5xx Errors
AWS CloudFront Outage Triggers Global 5xx Errors
Whatsapp Launches Username Reservation Feature
WhatsApp Opens Username Reservations for Its 3 Billion Users
Chrome 149 Update Fixes Serious Vulnerabilities
Google Chrome 149 Fixes 18 Serious Security Flaws
Meta Hands Whatsapp Reins To Cred Founder Kunal Shah
Meta Hands WhatsApp Reins to CRED Founder Kunal Shah
Major X Outage Disrupts Users Worldwide
Major X Outage Disrupts Users Worldwide, Service Restored
Technology
Google Health 5 05 Syncs To Apple Health
Google Health 5.05 Syncs to Apple Health but Omits HRV
Whatsapp Web Calling With Call Transfer
WhatsApp Web Now Supports Video and Audio Calls with Transfer
Apple Launches 17 99 Iphone Leases With Klarna
Apple Launches $17.99 iPhone Leases With Klarna In The USA
Meta Launches Seller App For Facebook Marketplace
Meta Launches Seller App for Facebook Marketplace
Google Adds Selfie Video Sign In For Account Recovery
Google Adds Selfie Video Sign-In for Account Recovery
Apple Maps Comes To Ford S Electric Vehicles In 2027
Apple Maps Comes to Ford’s Electric Vehicles in 2027
Gaming
Gta Vi Official Cover Art
GTA 6 Pre-Orders Start June 25, New Cover Art Unveiled
Epic Games Teases Unreal Engine 6 For Rocket League
Epic Games Teases Unreal Engine 6 for Rocket League
Stardew Valley Launched For Nintendo Switch 2 Edition
Stardew Valley Switch 2 Edition Arrives with Online Co-op
Hogwarts Legacy Game Crosses 40m Downloads
Hogwarts Legacy Crosses 40M Sales, Beating Industry Giants
Pubg Black Budget Closed Alpha Launched
PUBG: Black Budget Launches Closed Alpha Test With a Bold PvPvE Twist
Counter Strike 2 Skin Market Crashes After Valve Update
Counter-Strike 2’s $5.9 Billion Skin Economy Just Got Shattered
Newsletter

Too much tech noise?

We respect your time. One high-signal briefing a week — tech, AI, and security. Nothing else.

Newsletter

The SQ Briefing

We track tech, AI, and security 24/7. You get a 5-minute weekly summary.