Language, Language Models, and What We're Talking About
cs.CL
Submitted: 2026-09-03
Updated: 2026-09-03
Code: https://github.com/nlp-uoregon/mlmm-evaluation
Project page: https://malvinanissim.github.io
Terminology
Sources
- GPT-NL Public Corpus: A Permissively Licensed, Dutch-First Dataset for LLM Pre-training
- IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
- LLaMA: Open and Efficient Foundation Language Models
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Advanced Natural-based interaction for the ITAlian language: LLaMAntino-3-ANITA
- LoRA: Low-Rank Adaptation of Large Language Models
- Mistral 7B
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
- TinyStories: How Small Can Language Models Be and Still Speak Coherent English?
- Alignment Makes Language Models Normative, Not Descriptive
- The Shrinking Landscape of Linguistic Diversity in the Age of Large Language Models
- MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering