Role-Specific Attacks Turn Agent Communication into Trading Exploit Surface
Inter-agent communication in multi-agent trading systems creates a direct attack surface: corrupted signals propagate across roles and translate into...
Created by P Tracey
Hands‑on AI jailbreaks, exploit writeups, and defensive mitigation guides
Explore the latest content tracked by AI Red Teaming Hub
Inter-agent communication in multi-agent trading systems creates a direct attack surface: corrupted signals propagate across roles and translate into...
Large-scale red teaming with over 15,000 humans and RL-trained attackers reveals jailbreaks transfer from small open models directly to frontier...
Generative AI exploitation in intellectual property theft and privacy breaches further complicates the security landscape, as examined through adversarial attacks on AI models from a cybersecurity perspective.
AI safety evaluators introduce their own failure modes that can shift verdicts even when model responses stay fixed.
Security teams face growing attribution challenges as internal AI agents and malicious automated attackers produce similar behaviors.
Non-reserved placeholder domains like third-party[.]com now serve ClickFix malware to Windows users while showing decoys to others. The domain appears...
A claimed Mexican government hack using jailbroken AI illustrates the headline risk, but the documented shift is obliterating—an open-source technique...
Anthropic's disclosures of bad actors probing models for malware, drones, and infrastructure attacks show misuse risks that already justify a global...
The Manus vulnerability shows how indirect prompt injection—via an email hiding JSFuck-obfuscated commands—bypassed filters to achieve RCE and steal...
The U.S. Army is accelerating AI agents for cyber defense—detecting threats faster and enabling autonomous responses—while requiring human...
AAIGF-E identifies that AI-enabled smart-grid systems face adversarial manipulation, data poisoning, model tampering, and adversarial input attacks.
Blocking public GenAI tools fails to stop data exposure because employees route around controls using personal devices and accounts.
A single operator used three open-source AI agents to compromise 100+ retailers at roughly $25 each, extracting over 600,000 cards with almost no...
AI now generates vulnerabilities faster than teams can handle, cutting critical fix times ~50% while critical backlogs grew 29-fold. The scarce...
AI agents are already executing real attacks— one Chinese hacker used DeepSeek, Kimi and Claude to hit 100 organizations and steal credit-card data...
Enterprise prompt-protection tools must deliver three core capabilities to counter unauthorized AI use.
Efficiency optimizations like token pruning—used to remove redundant tokens—redefine the threat model for vision-language models by enabling...
AI safety fears and rogue agents are pushing investors to fund AI-native security startups at unprecedented levels, as traditional human-in-the-loop...