{"data":{"id":"076a613a-6646-4084-9f00-494402d3cb22","title":"OpenAI adds an AI safety layer to detect misuse without retaining enterprise data","summary":"OpenAI is introducing Private Safety Processing, a new safety system that detects misuse patterns across multiple AI interactions without keeping copies of the prompts or responses, allowing enterprises to monitor risks while maintaining Zero Data Retention (ZDR, keeping no record of user inputs or outputs after processing). Unlike traditional safety systems that check each interaction separately, this capability identifies suspicious behavior patterns that only become visible when viewing multiple related requests together, addressing risks like repeated attempts to bypass safeguards or coordinated misuse across accounts.","solution":"According to the source, Private Safety Processing itself is the mitigation being offered. OpenAI describes it as designed to \"identify patterns across related interactions without giving OpenAI personnel access to the underlying content.\" The system uses \"automated systems analyze interactions and generate a narrowly defined signal indicating the type of activity involved, instead of exposing the underlying prompts or responses.\" The capability is currently \"being tested with eligible enterprise and API customers.\"","labels":["safety","security"],"sourceUrl":"https://www.csoonline.com/article/4212398/openai-adds-an-ai-safety-layer-to-detect-misuse-without-retaining-enterprise-data.html","publishedAt":"2026-08-21T09:45:36.000Z","cveId":null,"cweIds":null,"cvssScore":null,"cvssSeverity":null,"severity":"info","attackType":[],"issueType":"news","affectedPackages":null,"affectedVendors":["OpenAI"],"affectedVendorsRaw":["OpenAI","Anthropic"],"classifierModel":"claude-haiku-4-5-20251001","classifierPromptVersion":"v3","cvssVector":null,"attackVector":null,"attackComplexity":null,"privilegesRequired":null,"userInteraction":null,"exploitMaturity":null,"epssScore":null,"patchAvailable":null,"disclosureDate":"2026-08-21T09:45:36.000Z","capecIds":null,"crossRefCount":0,"attackSophistication":"moderate","impactType":["safety"],"aiComponentTargeted":"inference","llmSpecific":true,"classifierConfidence":0.88,"researchCategory":null,"atlasIds":null}}