Advertisement
Alice Secures $140M to Enhance AI Model Defenses and Guardrails
AI security firm Alice raised $140M to combat adversarial AI, prompt injection, and jailbreak attempts in generative AI systems.
Autonomous AI Models as Attackers: Securing Enterprise AI
Analysis of the emerging threat where an organization's own AI models act as autonomous attackers. Learn to secure enterprise AI models from such intrusions.
Yellow Teams: Pioneering Adversarial AI Security Methodologies
Explore how 'Yellow Teams' are crucial for assessing AI security, developing defensive strategies, and understanding AI's potential as both a cyber weapon and a shield.
Bypass AI Malware Scanners via Policy-Triggering Prompt Injection
Malware authors are embedding 'forbidden' text into code to trigger safety refusals in AI-mediated security scanners, effectively bypassing automated analysis.
Agentic AI Worms: Defending Against LLM-Powered Lateral Movement
Enterprise security teams must prepare for agentic AI worms capable of autonomous adaptation, self-replication, and automated exploitation within a year.
AI Red Teaming: Guardrail Manipulation via Jailbreaking and Data Poisoning
Explores AI red teaming methods like jailbreaking and data poisoning used to manipulate AI guardrails and harden machine learning models against adversarial attacks.
Advertisement
Claude Mythos: Analyzing AI Threat Rumors in Japan Finance Sector
An investigation into the Claude Mythos panic in Japan's financial sector and how security experts distinguish between AI hype and actual cyber risks.
AI-Driven Cloud Attacks: The Zealot PoC and Autonomous Exploitation
Research into 'Zealot' reveals how AI-driven cloud attack simulations execute full-scale breaches faster than human defenders can effectively intervene.
AI Agent Autonomy: Analyzing the Machine-Speed Espionage Threat
Anthropic details a state-sponsored campaign where AI agents automated 90% of tactical operations, requiring new strategies for autonomous threat detection.
Hiding Malicious Commands from AI via Font-Rendering Manipulation
Learn how attackers use font-rendering tricks to bypass AI safety filters and execute prompt injection attacks against LLM-powered assistants.