EADC: Evaluation of Advanced and Deep-level Compliance in Large Language Models
cs.AI
Submitted: 2026-09-22
Updated: 2026-09-22
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Constitutional AI: Harmlessness from AI Feedback
- Obfuscated Activations Bypass LLM Latent-Space Defenses
- Knowledge Graph Representations for LLM-Based Policy Compliance Reasoning
- Jailbreaking Large Language Models with Symbolic Mathematics
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- A Different Approach to AI Safety: Proceedings from the Columbia Convening on Openness in Artificial Intelligence and AI Safety
- Gemma 4 Technical Report
- GLM-5: from Vibe Coding to Agentic Engineering
- The Llama 3 Herd of Models
- COMPL-AI Framework: A Technical Interpretation and LLM Benchmarking Suite for the EU Artificial Intelligence Act
- GraphAttack: Exploiting Representational Blindspots in LLM Safety Mechanisms
- Safety Compliance: Rethinking LLM Safety Reasoning through the Lens of Compliance
- Kimi K2.5: Visual Agentic Intelligence
- KG-DF: A Black-box Defense Framework against Jailbreak Attacks Based on Knowledge Graphs
- AIReg-Bench: Benchmarking Language Models That Assess AI Regulation Compliance
- GPT-4 Technical Report
- Red Teaming Language Models with Language Models
- OpenAI GPT-5 System Card
- Safety Assessment of Chinese Large Language Models
- Qwen3 Technical Report
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection