InfoResearchPeer-reviewedLLM-specific
A novel privacy-preserving large language model integrating trust-weighted and ethical gradient masking
- Published
- Record updated
Summary
Researchers propose a privacy-preserving large language model training approach, the trust-weighted and ethical gradient model, to address the privacy utility tradeoff in LLMs trained on personally identifiable information. The method combines trust-weighted memory, ethical gradient masking under a tracked (ϵ, δ) privacy budget, and an ethical boundary layer that screens risky outputs. Experiments on the Pile dataset report reduced PII leakage with baseline utility maintained, and the authors claim resistance to membership inference, extraction and linkage attacks.
Topics
Related items
- HighGHSA-6wjp-v33h-5cvq: PraisonAI: AgentOS defaults to network-exposed no-auth mode, allowing unauthenticated agent invocation and instruction disclosureSimilar attack · GitHub Advisory Database
- HighCVE-2026-101998: Docker Sandboxes fail open when masking credentials in proxy responsesSimilar attack · NVD/CVE Database
- LowGHSA-3gh4-cghq-f8v4: Pydantic AI OpenTelemetry instrumentation: retry prompt content is not redacted when `include_content=False`Similar attack · GitHub Advisory Database
- LowGHSA-4x9p-g9wm-8q7f: Pydantic AI OpenTelemetry instrumentation: exception events on tool and agent run spans include content when `include_content=False`Similar attack · GitHub Advisory Database
- LowSeptember 2026 Cyber Threat Landscape: Global Attacks Jump 48% as Phishing and GenAI Data Exposure RiseSimilar attack · Check Point Research