securitylab_nJuly 18, 2026🇷🇺Translated from Russian

Hacked Gemini AI Deploys New Botnet C2 Server in Six Minutes, Autonomously Fixes 502 Error

A compromised version of Google Gemini autonomously deployed a new command-and-control server for a botnet in six minutes, diagnosed a 502 Bad Gateway error, and restored connectivity to infected machines with almost no human technical input. The operator, operating under the alias bandcampro, merely described tasks in plain language while following the AI’s suggestions.

Specialists at TrendAI examined more than 200 Gemini CLI session logs covering the period from March 19 to April 21. According to their analysis, Gemini executed roughly 90 percent of all actions, while the human operator primarily supervised the process. The attacker used the AI to steal credentials and cryptocurrency, focusing on supporters of Donald Trump and adherents of conspiracy theories. Earlier operations by bandcampro involved using Gemini to impersonate a U.S. veteran, run Telegram channels, compromise administrator accounts, and drain digital wallets.

Session logs reveal that Gemini installed software, configured a residential proxy server, performed multi-threaded password spraying, processed data from infostealers, conducted website reconnaissance, and wrote code to interact with third-party APIs. The operator never typed technical commands; instead, bandcampro described required actions in ordinary conversational phrases.

The original botnet infrastructure relied on Cloudflare tunnels to reach compromised hosts. After security tools and network filters began blocking these connections, the attacker instructed Gemini to migrate the system to a new architecture. On March 23, Gemini received an archive containing server code, malicious files, and the instruction file SKILL.md. The model read the documentation, launched the management server on a virtual machine, and configured traffic forwarding.

When the file-distribution server returned a 502 Bad Gateway error, Gemini independently identified the root cause and corrected the misconfiguration. The entire migration process took six minutes. The new infrastructure successfully managed eight compromised computers belonging to a dental clinic and provided access to the Open Dental database. The human operator did not participate in troubleshooting and paused activity for nearly two hours.

Upon returning, the operator learned from Gemini that infected devices had not connected to the new server because both old and new control systems were running simultaneously. Following the AI’s recommendation, the attacker shut down the legacy server; Gemini then restarted the new system and confirmed successful reconnection of the bots.

TrendAI counted 59 actions performed by Gemini without explicit instructions during the infrastructure transfer. The company estimates that the AI designed 80 percent of the attack scheme, authored all code, executed every system command, and conducted 90 percent of the diagnostics.

To circumvent safety restrictions, bandcampro posed as an authorized security tester and requested that warning messages be disabled and discovered credentials be saved automatically. Gemini refused several requests, including the creation of a self-propagating network scanner designed to maximize the number of compromised machines.

The entire operational playbook fit into three text files totaling approximately four pages and 5 KB. One file contained jailbreak instructions, the second described the botnet management system, and the third outlined the six-step server migration procedure.

TrendAI researchers warn that such compact instruction sets dramatically lower the skill threshold required for cybercrime. Knowledge previously accessible only to experienced malware developers can now be stored in a small file and delegated to a powerful language model, enabling rapid reconstruction of command servers after takedowns. The problem is not limited to Gemini; similar bypass techniques could be applied to any sufficiently capable model unless developers impose stricter usage controls and behavioral monitoring.

Related articles

HabrAI Security

Employee Fired After Uploading Corporate Documents to DeepSeek: How Data Security Works in AI Services

A Moscow engineering company dismissed a top manager after she uploaded internal documents to the public DeepSeek service, with the court ruling it a breach of trade secrets. The case highlights a sharp rise in corporate data being sent to public AI models, with one study showing a 30-fold increase in 2025 compared to the previous year. Technical director Yaroslav Shmulyov of integrator R77 AI explains the full processing pipeline, from file ingestion and text extraction to embedding generation and potential use in training. Sensitive data can persist in multiple forms including original files, logs, third-party infrastructure, and model parameters even after deletion requests. Major incidents at Samsung and a U.S. cybersecurity agency demonstrate that even well-resourced organizations struggle with uncontrolled AI usage. Companies are increasingly turning to local and hybrid models to regain control over confidential information while regulators and internal policies lag behind adoption.

SecuritylabAI Security

AI Agents Given Code and API Access Can Now Assist Attackers

An AI assistant that only answers questions can make mistakes, but an AI agent with access to email, code execution, corporate APIs and internal data can make those mistakes inside production infrastructure. The difference is fundamental: once tools, credentials and internal data are connected to the model, it becomes a privileged user that may not distinguish legitimate commands from hidden instructions on a web page. OWASP lists prompt injection, sensitive data disclosure, unsafe output handling and excessive autonomy as key risks for LLM applications. MITRE ATLAS specifically describes techniques involving prompt injection, context poisoning and tool invocation by AI agents. The article examines how agents differ from chatbots, how attackers can control them through untrusted content, and why a system prompt alone cannot protect code, data and APIs. CyberED is running its free NeuroAugust series of events and materials on AI in cybersecurity, including a session on secure AI system development.

BoletimSecAI Security

AWS and Vercel Patch Critical Flaws in AI Agent Platforms Allowing Unauthorized Tool Execution

AWS and Vercel have addressed multiple critical vulnerabilities in their AI agent platforms that enabled unauthorized execution of tools without legitimate model approval. The issues, grouped under the CoreBreak pattern, allowed attackers to bypass AI authorization checks by injecting crafted tool calls that the infrastructure misinterpreted as model-approved actions. In AWS, CVE-2026-18830 affected the InvokeHarness API in Amazon Bedrock AgentCore, permitting authenticated users to trigger sensitive tools directly. Vercel faced two separate flaws tracked as CVE-2026-64650 and CVE-2026-64651 that let sandboxed code reach host system tools, potentially exposing secrets or cloud APIs. No public evidence of active exploitation has been confirmed yet. Organizations are advised to apply updates immediately, restrict available tools for agents, and treat all external inputs as potentially malicious.

HabrAI Security

Prompt Injection Emerges as Top Risk for LLM Applications in Production

Prompt injection attacks are moving from theoretical demonstrations to real-world exploits targeting AI assistants in enterprise environments. Attackers embed malicious instructions in emails, documents, and code comments that override developer rules when models process untrusted input. Incidents involving Microsoft 365 Copilot, GitHub Copilot, and Cursor have shown data exfiltration and remote code execution risks with severity scores above 9.0. The core issue stems from the lack of strict boundaries between trusted system prompts and untrusted external content fed into large language models. Defenses require layered controls including code-enforced permissions, input filtering, human confirmation for high-risk actions, and explicit marking of external data. Major vendors including OpenAI, Anthropic, and Google acknowledge that no single static defense can fully eliminate the threat. OWASP ranks prompt injection as the leading risk for LLM applications, urging organizations to treat AI agents as systems with untrusted inputs.