Scaling Articulated Rationales for MLLM-based Recommendation
cs.IR, cs.AI
Submitted: 2026-09-15
Updated: 2026-09-21
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Qwen2.5-VL Technical Report
- Mind2Web 2: Evaluating Agentic Search with Agent-as-a-Judge
- KuaiLive-M3: A Multi-Modal, Multi-Domain, and Multi-Feedback Dataset for Live Streaming Recommendation
- V-STaR: Training Verifiers for Self-Taught Reasoners
- Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
- G-Eval: NLG Evaluation using GPT-4 with Better Human Alignment
- LLM-Alignment Live-Streaming Recommendation
- JudgeBench: A Benchmark for Evaluating LLM-based Judges
- Gemini: A Family of Highly Capable Multimodal Models
- Qwen3-VL Technical Report
- RecGPT-V2 Technical Report
- OneLive: Dynamically Unified Generative Framework for Live-Streaming Recommendation
- QARM V2: Quantitative Alignment Multi-Modal Recommendation for Reasoning User Sequence Modeling
- MiniCPM4: Ultra-Efficient LLMs on End Devices
- SARM: LLM-Augmented Semantic Anchor for End-to-End Live-Streaming Ranking
- Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
- Agent-as-a-Judge: Evaluate Agents with Agents
Related papers
- The Price of Isolation: Estimating the Ecosystem Cost of Symmetric Two-Sided A/B Testing
- SCAR: Semantic Continuity-Aware Retrieval for Efficient Context Expansion in RAG
- MixLoRA-DSI: Dynamically Expandable Mixture-of-LoRA Experts for Rehearsal-Free Generative Retrieval over Dynamic Corpora
- RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation
- Right Family, Wrong Skill: Evaluating Risk Exposure in Agent Skill Retrieval
- UltRAG: a Universal Simple Scalable Recipe for Knowledge Graph RAG