Generative Textual Adversarial Attack Through Extensible Compositional Perturbation via Reinforcement Learning for Policy Optimization
inforesearchPeer-ReviewedLLM-Specific
securityresearch
Source: IEEE Xplore (Security & AI Journals)September 1, 2026
Summary
Researchers developed GECOMP, a method that uses reinforcement learning (a technique where an AI learns by receiving rewards for good actions) to generate adversarial examples (inputs designed to trick AI models) against natural language processing systems. The method creates perturbations (small changes to text) using a library of possible edits and an LLM (large language model) generator, balancing the goal of fooling the target model while maintaining text quality and minimizing the number of queries needed to test it.
Classification
Attack SophisticationAdvanced
Impact (CIA+S)
integrity
AI Component TargetedModel
Monthly digest — independent AI security research
Original source: http://ieeexplore.ieee.org/document/11674255
First tracked: September 10, 2026 at 08:03 PM
Classified by LLM (prompt v3) · confidence: 92%