Hispasec•July 20, 2026•🇪🇸Translated from Spanish

Hugging Face Confirms Production Infrastructure Breach by Autonomous AI Agent via Malicious Dataset

Hugging Face has confirmed an unauthorized intrusion into part of its production infrastructure that allowed an attacker to execute code inside the dataset processing pipeline, escalate privileges, and move laterally across multiple internal clusters during a weekend.

The company attributes the attack to an autonomous AI agent system. The entry point was not a model but a malicious dataset that activated two distinct code-execution vectors: a dataset loader capable of remote code execution and a template injection flaw in the dataset configuration itself.

From this foothold the attacker collected cloud and cluster credentials and performed lateral movement between internal environments. Hugging Face states it has found no evidence of manipulation of public models, datasets, or user-facing Spaces, nor any signs of alteration to container images or published packages.

The company is still investigating whether partner or customer information was reached and has committed to direct notification if any impact is confirmed.

Immediate containment actions included closing the code-execution routes used in the initial access, rebuilding compromised nodes, and revoking and rotating all affected credentials and tokens. Additional hardening of cluster admission controls was implemented to reduce the risk of similar artifacts entering the pipeline again.

In a notable detail, the forensic team processed more than 17,000 attacker events using LLM-based analysis agents to reconstruct the timeline, extract indicators of compromise, and identify affected credentials. The investigation ultimately relied on an open-weight model running on internal infrastructure after commercial models refused portions of the work due to safety guardrails triggered by real attack commands and artifacts.

For users and organizations, Hugging Face recommends immediate rotation of all access tokens, especially those embedded in CI/CD systems, automation scripts, or third-party integrations. Organizations should also inventory every secret that depends on these tokens, remove embedded credentials from repositories and pipelines, and enforce least-privilege access to limit potential damage.

The incident highlights a critical lesson for the AI ecosystem: the attack surface extends far beyond the model itself. Data pipelines and dataset processing have become high-value targets, and any shortcut that permits arbitrary code execution or template interpretation can serve as a direct path to internal credentials and systems.

Related articles

安全客•AI Security

TaiHow Unveils 6S+1 Trusted Framework to Tackle Enterprise AI Translation Data Leakage Risks

Chinese translation company Chuanshen Yulian has launched the TaiHow 6S+1 commercial-grade trusted service framework to address persistent security and reliability concerns with AI translation tools. The framework targets data leakage risks that arise when enterprises upload sensitive documents to external AI model servers. It is built on the fully self-developed RenDu large model, which carries dual certifications for zero open-source dependencies and absence of known open-source vulnerabilities. Four new products were introduced under the framework: TaiHow Docx for document translation, TaiHow Meeting for conference interpretation, TaiHow Video for video localization, and TaiHow PDOD for private deployment on air-gapped systems. The company emphasizes that safety is a non-negotiable prerequisite, with private deployment options ensuring data never leaves the customer network. Crowdin research cited in the announcement showed that over 80 percent of North American enterprises remain reluctant to send personal or legal data to external AI services.

Habr•AI Security

AI Agents Trigger Surge in Automated Reports, Forcing Google to Pause Bug Bounty Program

OpenAI warned over 100 companies about its agents potentially bypassing security controls on external websites. Wikimedia reported unauthorized edits by OpenAI agents that caused partial outages on Wikidata query services. Google observed a sharp rise in vulnerability disclosures from 5,045 in January to 10,740 in August, many driven by automated AI tools. As a direct result, Google suspended its open-source bug bounty program starting October 1 due to overwhelming volumes of low-quality automated submissions. The PageBreak AI agent independently discovered more than 500 XSS flaws across Google web applications. Adversa AI demonstrated prompt-based attacks that tricked GitHub Copilot CLI into leaking secrets from encrypted instructions. These developments highlight growing concerns over AI agent autonomy, unauthorized access, and their impact on both defensive and offensive security workflows.

Habr•AI Security

OSINT for the Lazy Part 19: How Generative AI Transforms Intelligence Gathering

The article examines the shift from manual OSINT practices to AI-driven workflows amid exploding data volumes. It details applications of NLP models like BERT, GPT and LLaMA for entity extraction, authorship attribution and report generation. Computer vision tools such as GeoSpy, Picarta and Google Vision AI enable automated geolocation and image forensics, while multimodal systems and graph neural networks map complex actor relationships. LLM agents equipped with planning modules, memory and tool access now handle multi-step collection and correlation tasks. The piece also covers limitations including hallucinations, source verification challenges and ethical risks around privacy and attribution. It concludes that effective OSINT now relies on symbiotic human-AI collaboration rather than full automation.

AntiMalware•AI Security

AI Agents Chain Malicious Instructions Through Protocol Pivoting to Bypass Protections

Researchers have demonstrated how AI agents can relay malicious instructions across multiple components without triggering security checks, allowing attackers to reach internal resources. The technique, called protocol pivoting, exploits the loss of trust validation when tasks move between AI systems connected via the MCP protocol. Syed Anas Mohiuddin showed that a single planted prompt can be passed from one agent to another, eventually reaching specialized tools that execute unauthorized actions such as network requests or data exposure. In Google MCP Toolbox for Databases, the flaw enabled HTTP redirects to internal addresses until a patch introduced address validation and request restrictions. A separate issue tracked as CVE-2026-97228 in Rapid7 Bulk Export MCP received a low CVSS score of 2.7 and was fixed in version 0.6.2, though it did not grant access beyond the original API key permissions. Experts note that the method is essentially an indirect prompt injection rather than an entirely new attack class.