QBugLM: An Agentic Benchmarking Framework for LLM-based Quantum Software Debugging
cs.SE, cs.ET, quant-ph
Submitted: 2026-06-05
Updated: 2026-06-05
Code: https://github.com/qachub/qbuglm
Terminology
Sources
- Qiskit Code Assistant: Training LLMs for generating Quantum Computing Code
- Qiskit HumanEval: An Evaluation Benchmark For Quantum Code Generative Models
- Leveraging Mutation Analysis for LLM-based Repair of Quantum Programs
- Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
- ReAct: Synergizing Reasoning and Acting in Language Models
- Evaluating Large Language Models Trained on Code
- Asking the Right Questions: Improving Reasoning with Generated Stepping Stones
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties