Grounding SWE-Agent Decisions in Architecture-0 Design: Navigating Unknown Unknowns through Physical Mapping
cs.SE, cs.AI
Submitted: 2026-09-15
Updated: 2026-09-15
Comments: 51 pages, 13 figures, 18 tables. Preprint of a manuscript under review at ACM TOSEM
Code: https://github.com/donnemartin/system-design-primer
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Demonstrating specification gaming in reasoning models
- Evaluating Large Language Models Trained on Code
- SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
- Feedback Loops With Language Models Drive In-Context Reward Hacking
- OpenHands: An Open Platform for AI Software Developers as Generalist Agents
- SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
- Grounding LLMs in Scientific Discovery via Embodied Actions
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties