SkillShield: Prompt-Space Security Skills for LLM Coding Agents
cs.CR
Submitted: 2026-08-26
Updated: 2026-08-26
Code: https://github.com/anomalyco/opencode
Terminology
Sources
- SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Constitutional AI: Harmlessness from AI Feedback
- DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models
- FATH: Authentication-based Test-time Defense against Indirect Prompt Injection Attacks
- Universal and Transferable Adversarial Attacks on Aligned Language Models
- Jailbreak Attacks and Defenses Against Large Language Models: A Survey
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs