Compared to What? A Human-Anchored Security Benchmark for LLM-Generated Infrastructure-as-Code
cs.CR, cs.AI, cs.MA, cs.SE
Submitted: 2026-08-28
Updated: 2026-09-21
Code: https://github.com/AnimeshShaw/GenIaC-SecBench
Terminology
Sources
- Security-First Evaluation of Text-to-Terraform: Benchmarking LLMs and SLMs for Secure IaC Generation
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- A Survey on LLM-as-a-Judge
- Empirical Standards for Software Engineering Research
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs