An Open Pipeline and Dashboard for Systemic-Risk Evidence under the EU AI Act's Code of Practice
cs.AI
Submitted: 2026-09-23
Updated: 2026-09-25
Code: https://github.com/jacobemmerson/certificate
Terminology
Sources
- AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
- Time Travel in LLMs: Tracing Data Contamination in Large Language Models
- LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
- EU-Agent-Bench: Measuring Illegal Behavior of LLM Agents Under EU Law
- TruthfulQA: Measuring How Models Mimic Human Falsehoods
- Agentic Misalignment: How LLMs Could Be Insider Threats
- SocialHarmBench: Revealing LLM Vulnerabilities to Socially Harmful Requests
- Bench-2-CoP: Can We Trust Benchmarking for EU AI Compliance?
- Hermes 4 Technical Report
- Sociotechnical Safety Evaluation of Generative AI Systems
- DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks
- Universal and Transferable Adversarial Attacks on Aligned Language Models
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection