Structural Enforcement of Statistical Rigor in AI-Driven Discovery: A Functional Architecture
cs.SE, cs.AI
Submitted: 2025-11-10
Updated: 2026-09-25
Code: https://github.com/karsar/aiscientist-guards
Terminology
Sources
- The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
- Agentic AI for Scientific Discovery: A Survey of Progress, Challenges, and Future Directions
- DeepScientist: Advancing Frontier-Pushing Scientific Findings Progressively
- The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
- ToolUniverse: An open platform for democratizing AI scientists
- Curie: Toward Rigorous and Automated Scientific Experimentation with AI Agents
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties