Text-Trained LLMs Can Zero-Shot Extrapolate PDE Dynamics, Revealing a Three-Stage In-Context Learning Mechanism
cs.LG
Submitted: 2025-09-08
Updated: 2026-03-11
Journal ref: Mach. Learn.: Sci. Technol. 7 (2026) 055001
Code: https://github.com/Jiajun-Bao/LLM-PDE-Dynamics
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Phi-4 Technical Report
- GPT-4 Technical Report
- Pre-trained Large Language Models Learn Hidden Markov Models In-context
- A Survey on In-context Learning
- The Llama 3 Herd of Models
- CodePDE: An Inference Framework for LLM-driven PDE Solver Generation
- Explain Like I'm Five: Using LLMs to Improve PDE Surrogate Models with Text
- PDE-Controller: LLMs for Autoformalization and Reasoning of PDEs
- Gemma 3 Technical Report
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Large Language Models as Markov Chains
- A Survey of Large Language Models
- Text2PDE: Latent Diffusion Models for Accessible Physics Simulation
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks