Cognitive Expert Language Models Better Align with the Corresponding Brain Systems
cs.CL
Submitted: 2026-09-28
Updated: 2026-09-28
Terminology
Sources
- Phi-4 Technical Report
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- LoRA Learns Less and Forgets Less
- A Dataset for Answering Time-Sensitive Questions
- Training Verifiers to Solve Math Word Problems
- A foundation model of vision, audition, and language for in-silico neuroscience
- The Llama 3 Herd of Models
- Measuring Massive Multitask Language Understanding
- LoRA: Low-Rank Adaptation of Large Language Models
- Branch-Train-Merge: Embarrassingly Parallel Training of Expert Language Models
- Dissociating language and thought in large language models
- Orca-Math: Unlocking the potential of SLMs in Grade School Math
- Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli
- Qwen2.5 Technical Report
- Compressive Transformers for Long-Range Sequence Modelling
- Fine-tuned vs. Prompt-tuned Supervised Representations: Which Better Account for Brain Language Representations?
- ExpertPrompting: Instructing Large Language Models to be Distinguished Experts
- Contravariance Theory: Strong Alignment for Minimal Solutions to Hard Tasks
- Qwen3 Technical Report
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering