A Russian-speaking threat actor using the handle Trim reportedly bypassed safety controls on public frontier AI models and turned the resulting jailbreak methods into an offensive cyber platform. Researchers said Trim first posted a guide on a Russian-language cybercrime forum in March describing six techniques for coercing Anthropic Claude Opus into generating malicious outputs, while also naming fallback models and offering access to black-market API keys. The activity did not rely on exploiting software flaws in the models themselves; instead, it abused public API access, leaked system prompts, and prompt-engineering techniques to defeat guardrails.
By June, Trim had allegedly commercialized the approach as AI Pentest Checker, a service that automates web reconnaissance, vulnerability validation, exploitation reporting, and PDF report generation. The platform reportedly combined AI models including Opus 4.8 and GLM-5 with established offensive tools such as Nuclei, ffuf, katana, subfinder, and gitleaks, with one modified system prompt said to be derived from a leaked Fable 5 configuration. Researchers warned the case shows how quickly cybercriminals can operationalize and monetize AI-assisted offensive workflows, increasing the speed and accessibility of attack capabilities.

Mallory correlates global threat intelligence with your attack surface — know if you’re exposed before adversaries strike.
2 events from the most recent confirmed update back to the earliest known activity.
By June 21, 2026, Trim had turned the jailbreak techniques into a commercialized platform called AI Pentest Checker. The service reportedly combined frontier AI models with tools including Nuclei, ffuf, katana, subfinder, and gitleaks to automate reconnaissance, vulnerability validation, exploitation reporting, and PDF report generation.
On March 31, 2026, the Russian-speaking threat actor "Trim" appeared on a Russian-language cybercrime forum and shared six named techniques for bypassing Claude Opus safety controls to obtain malicious outputs. The post also referenced fallback model options and access to black-market API keys.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
5 references tracked. Mallory keeps watching after this page renders.
cyberveille.ch
Open sourcecyberaccord.com
Open sourcecybersecuritynews.com
Open sourcecatonetworks.com
Open sourcedarkreading.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.