CanLegalRAGBench: Evaluating Retrieval-Augmented Generation on Canadian Case Law
Ethan Zhao, Maksym Taranukhin, Wei Cui, Moira Aikenhead, Vered Shwartz
cs.CL
Submitted: 2026-08-18
Updated: 2026-08-20
Code: https://github.com/NLP-UBC/CanLegalRAGBench
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- The Massive Legal Embedding Benchmark (MLEB)
- DSPy: Compiling Declarative Language Model Calls into Self-Improving Pipelines
- Experimenting with Legal AI Solutions: The Case of Question-Answering for Access to Justice
- LegalBench-RAG: A Benchmark for Retrieval-Augmented Generation in the Legal Domain
- EmbeddingGemma: Powerful and Lightweight Text Representations
- Introducing the A2AJ's Canadian Legal Data: An open-source alternative to CanLII for the era of computational law
- Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering