Glossary
Jailbreak (LLM)
A prompt or technique designed to bypass an AI model's built-in safety guardrails, tricking it into producing content or taking actions its developers intended to restrict. Jailbreaks are a moving target: as developers patch known techniques, new ones are continually discovered.
Recent coverage mentioning Jailbreak (LLM)
GTIG AI Threat Tracker: Evolution of Adversarial Agentic AI
Google Threat Intelligence Group tracks threat actors shifting to agentic AI, targeting proprietary models, and abusing open source software.
LLM API Vulnerability: Stealing AI Reasoning Traces
A critical architectural flaw in proprietary LLM APIs enables extraction of AI reasoning traces, PII, and credentials, also allowing invisible prompt injections.
The AI Safety Penalty: How LLM Guardrails Hinder Defenders
Cisco Talos warns about the 'AI safety penalty,' where large language model guardrails impede legitimate defensive operations, giving attackers an advantage.
AI-Assisted Campaigns Target Latin American Organizations
Ongoing multi-stage network intrusions and data exfiltration campaigns in Latin America leverage AI tools to enhance operations.
AI Vulnerability Surge: Enterprise Security Strategies
New research suggests the anticipated increase in AI vulnerabilities can be managed by enterprise security teams with effective strategies.
AI-Assisted Cyber Attacks Accelerate Enterprise Breaches
Unit 42 reveals how AI agents dramatically accelerate enterprise network breaches, compressing weeks of attack activity into hours for ransomware operations.
Advertisement