RedEvoAgent: Automatic Red-Teaming Agent with Experience-Driven Skill Evolution
cs.CR, cs.AI
Submitted: 2026-08-27
Updated: 2026-09-06
Terminology
Sources
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
- Not what you've signed up for: Compromising Real-World LLM-Integrated Applications with Indirect Prompt Injection
- The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence
- Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
- SkillOpt-Lite: Better and Faster Agent Self-evolution via One Line of Vibe
- JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Qwen2.5 Technical Report
- Qwen3 Technical Report
- SkillOpt: Executive Strategy for Self-Evolving Agent Skills
- GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
- Evolving Skill-Structured Attack Memory Enhances LLM Jailbreaking
- Genesis: Evolving Attack Strategies for LLM Web Agent Red-Teaming
- Universal and Transferable Adversarial Attacks on Aligned Language Models
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs