Wontopos Tablet 2: Measuring Multilingual and Multimodal Memory Retrieval Without Lexical Matching
cs.IR, cs.CL, cs.CV
Submitted: 2026-08-24
Updated: 2026-08-24
Terminology
Sources
- mMARCO: A Multilingual Version of the MS MARCO Passage Ranking Dataset
- M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
- Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
- Language-agnostic BERT Sentence Embedding
- RULER: What's the Real Context Size of Your Long-Context Language Models?
- Unsupervised Dense Information Retrieval with Contrastive Learning
- Dense Passage Retrieval for Open-Domain Question Answering
- Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks
- Lost in the Middle: How Language Models Use Long Contexts
- Evaluating Very Long-Term Conversational Memory of LLM Agents
- MTEB: Massive Text Embedding Benchmark
- MemGPT: Towards LLMs as Operating Systems
- The Great Nugget Recall: Automating Fact Extraction and RAG Evaluation with Large Language Models
- Measuring and Narrowing the Compositionality Gap in Language Models
- Learning Transferable Visual Models From Natural Language Supervision
- Beyond a Million Tokens: Benchmarking and Enhancing Long-Term Memory in LLMs
- BEIR: A Heterogenous Benchmark for Zero-shot Evaluation of Information Retrieval Models
- "Knowing When You Don't Know": A Multilingual Relevance Assessment Dataset for Robust Retrieval-Augmented Generation
- Crossmodal-3600: A Massively Multilingual Multimodal Evaluation Dataset
- Interleaving Retrieval with Chain-of-Thought Reasoning for Knowledge-Intensive Multi-Step Questions
Related papers
- The Price of Isolation: Estimating the Ecosystem Cost of Symmetric Two-Sided A/B Testing
- SCAR: Semantic Continuity-Aware Retrieval for Efficient Context Expansion in RAG
- MixLoRA-DSI: Dynamically Expandable Mixture-of-LoRA Experts for Rehearsal-Free Generative Retrieval over Dynamic Corpora
- RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
- Right Family, Wrong Skill: Evaluating Risk Exposure in Agent Skill Retrieval
- UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG