p1: Better Prompt Optimization with Fewer Prompts
cs.LG, cs.CL
Submitted: 2026-04-09
Updated: 2026-08-26
Code: https://github.com/huggingface/Math-Verify
Terminology
Sources
- GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
- PRL: Prompts from Reinforcement Learning
- UPRISE: Universal Prompt Retrieval for Improving Zero-Shot Evaluation
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- RLPrompt: Optimizing Discrete Text Prompts with Reinforcement Learning
- Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution
- Prompt Curriculum Learning for Efficient LLM Post-Training
- EvoPrompt: Connecting LLMs with Evolutionary Algorithms Yields Powerful Prompt Optimizers
- DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines
- PRewrite: Prompt Rewriting with Reinforcement Learning
- StablePrompt: Automatic Prompt Tuning using Reinforcement Learning for Large Language Models
- Prompt-R1: Collaborative Automatic Prompting Framework via End-to-end Reinforcement Learning
- Understanding R1-Zero-Like Training: A Critical Perspective
- Automatic Prompt Optimization with "Gradient Descent" and Beam Search
- Generalizing Verifiable Instruction Following
- POLCA: Stochastic Generative Optimization with LLM
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- Qwen3 Technical Report
- One Prompt is not Enough: Automated Construction of a Mixture-of-Expert Prompts
- PromptBridge: Cross-Model Prompt Transfer for Large Language Models
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks