Precision over Scale: A Polish-Silesian Benchmark and a Translation System Outperforming Open-Source and Commercial Models
cs.CL
Submitted: 2026-10-01
Updated: 2026-10-03
Terminology
Sources
- QLoRA: Efficient Finetuning of Quantized LLMs
- Beyond English-Centric Multilingual Machine Translation
- Language-agnostic BERT Sentence Embedding
- TranslateGemma Technical Report
- PLLuM: A Family of Polish Large Language Models
- MADLAD-400: A Multilingual And Document-Level Large Audited Dataset
- EuroLLM-9B: Technical Report
- Bielik 11B v3: Multilingual Large Language Model for European Languages
- HPLT 3.0: Very Large-Scale Multilingual Resources for LLMs and MT. Mono- and Bi-lingual Data, Multilingual Evaluation, and Pre-Trained Models
- A Call for Clarity in Reporting BLEU Scores
- CCMatrix: Mining Billions of High-Quality Parallel Sentences on the WEB
- OpenAI GPT-5 System Card
- No Language Left Behind: Scaling Human-Centered Machine Translation
- BOUQuET: dataset, Benchmark and Open initiative for Universal Quality Evaluation in Translation
- Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering