Emergent Abilities in Large Language Models: A Survey
cs.LG, cs.AI, cs.CL
Submitted: 2025-02-28
Updated: 2026-08-26
Code: https://github.com/SajjjadAyobi/PersianQA
Terminology
Sources
- A Theory for Emergence of Complex Skills in Language Models
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- Constitutional AI: Harmlessness from AI Feedback
- Mechanistic Interpretability for AI Safety -- A Review
- User Intent Recognition and Satisfaction with Large Language Models: A User Study with ChatGPT
- Could a Large Language Model be Conscious?
- States Hidden in Hidden States: Implicit Discrete State Representations Emerge in LLMs' Hidden States
- AgentVerse: Facilitating Multi-Agent Collaboration and Exploring Emergent Behaviors
- Scaling Laws for Predicting Downstream Performance in LLMs
- Why Can GPT Learn In-Context? Language Models Implicitly Perform Gradient Descent as Meta-Optimizers
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
- CausalLM is not optimal for in-context learning
- A Survey on In-context Learning
- Understanding Emergent Abilities of Language Models from the Loss Perspective
- Demystifying Prompts in Language Models via Perplexity Estimation
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges
- Deception Abilities Emerged in Large Language Models
- A Theory of Emergent In-Context Learning as Implicit Structure Induction
- Structured Prompting: Scaling In-Context Learning to 1,000 Examples
- Measuring Massive Multitask Language Understanding
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks