tau-Rec: A Verifiable Benchmark for Agentic Recommender Systems
cs.IR, cs.AI, cs.CL
Submitted: 2026-06-08
Updated: 2026-07-24
Code: https://github.com/nbharaths/tau-rec
Terminology
Sources
- $\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment
- Toward Safe and Human-Aligned Game Conversational Recommendation via Multi-Agent Decomposition
- Interactive Recommendation Agent with Active User Commands
- Benchmark Leakage Trap: Can We Trust LLM-based Recommendation?
- Instruction-Following Evaluation for Large Language Models
Related papers
- The Price of Isolation: Estimating the Ecosystem Cost of Symmetric Two-Sided A/B Testing
- SCAR: Semantic Continuity-Aware Retrieval for Efficient Context Expansion in RAG
- MixLoRA-DSI: Dynamically Expandable Mixture-of-LoRA Experts for Rehearsal-Free Generative Retrieval over Dynamic Corpora
- RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
- Right Family, Wrong Skill: Evaluating Risk Exposure in Agent Skill Retrieval
- UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG