What Do CAE Simulation Agents Really Need Beyond a Generic Harness?
cs.CE, cs.CL, physics.comp-ph
Submitted: 2026-09-03
Updated: 2026-09-03
Code: https://github.com/openai/codex
Terminology
Sources
- SimBench: A Framework for Evaluating and Diagnosing LLM-Based Digital-Twin Generation for Multi-Physics Simulation
- OptMetaOpenFOAM: Large Language Model Driven Chain of Thought for Sensitivity Analysis and Parameter Optimization based on CFD
- MetaOpenFOAM: an LLM-based multi-agent framework for CFD
- MetaOpenFOAM 2.0: Large Language Model Driven Chain of Thought for Automating CFD Simulation and Post-Processing
- ALL-FEM: Agentic Large Language models Fine-tuned for Finite Element Methods
- Fine-tuning a Large Language Model for Automating Computational Fluid Dynamics Simulations
- FEABench: Evaluating Language Models on Multiphysics Reasoning Ability
- RExBench: Can coding agents autonomously implement AI research extensions?
- MechAgents: Large language model multi-agent collaborations can solve mechanics problems, generate new data, and integrate knowledge
- FeaGPT: an End-to-End agentic-AI for Finite Element Analysis
- CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics
- AgenticSciML: Collaborative Multi-Agent Systems for Emergent Discovery in Scientific Machine Learning
- Single-Agent LLMs Outperform Multi-Agent Systems on Multi-Hop Reasoning Under Equal Thinking Token Budgets
- Confucius Code Agent: Scalable Agent Scaffolding for Real-World Codebases
- Integrating Large Language Models for Automated Structural Analysis
- Automating Structural Engineering Workflows with Large Language Model Agents
- VASP Agent: An Agentic Framework for Autonomous First-principles Calculations
- SimuAgent: An LLM-Based Simulink Modeling Assistant Enhanced with Reinforcement Learning
- A Preliminary Assessment of Coding Agents for CFD Workflows
- Dive into Claude Code: The Design Space of Today's and Future AI Agent Systems
Related papers
- Constrained Sensing and Reliable State Estimation with Shallow Recurrent Decoders on a TRIGA Mark II Reactor
- Evidence-Unit Fairness and the Limits of Query-Adaptive Sparse-Dense Fusion in Financial Document Retrieval
- Chemical Chain-of-Thought Functions as a Hallucination-Prone Molecular Scratchpad
- Lightweight Adaptation of EEG Foundation Models for Stroke Motor Imagery Decoding: Domain Shift and Subject-Level Robustness
- RetroDFM-R: Reasoning-Driven Retrosynthesis Prediction with Large Language Models via Reinforcement Learning
- Wildfire Suppression: Complexity, Models, and Instances