InfoResearchPeer-reviewedLLM-specific
CAST: A Compiler-Based Framework for Systematically Testing LLM Compositional Safety
- Published
- Record updated
Summary
Researchers introduce compositional safety, the property that an LLM stays safe under multi-step workflows that decompose harmful intents, not only isolated malicious prompts. They present CAST, a compiler-inspired testing framework that uses an intermediate representation, CAIR, to generate and reassemble sub-tasks for evaluating malicious-code safety. On four LLMs across three testbeds, CAST exposed severe safety violations in strongly aligned models, with up to a 365% increase in successful test cases over baseline strategies.
Related items
- CriticalHermes Agent - PKCE Session Takeover via Redirect-URI Parser ConfusionSimilar attack · Tenable Research Advisories
- LowLost in the comments: Social context as a single‐pass jailbreak and defense on agentic platformsSimilar attack · OpenAlex (peer-reviewed AI security)
- MediumGHSA-hmq2-7hp6-7crh: Banks: User-controlled prompt input can be parsed as privileged chat messagesSimilar attack · GitHub Advisory Database
- HighGHSA-6wjp-v33h-5cvq: PraisonAI: AgentOS defaults to network-exposed no-auth mode, allowing unauthenticated agent invocation and instruction disclosureSimilar attack · GitHub Advisory Database
- HighCVE-2026-101998: Docker Sandboxes fail open when masking credentials in proxy responsesSimilar attack · NVD/CVE Database