SignRAG: Unified Retrieval-Augmented Gloss-Free Sign Language Translation
cs.CL, cs.CV, cs.MM
Submitted: 2026-10-08
Updated: 2026-10-08
Code: https://github.com/open-mmlab/mmpose
Terminology
Sources
- Qwen2.5-VL Technical Report
- VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- RAG-RL: Advancing Retrieval-Augmented Generation via RL and Curriculum Learning
- Lost in Translation, Found in Embeddings: Sign Language Translation and Alignment
- R3-RAG: Learning Step-by-Step Reasoning and Retrieval for LLMs via Reinforcement Learning
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning
- Qwen3 Technical Report
- CoCa: Contrastive Captioners are Image-Text Foundation Models
- A Chinese Continuous Sign Language Dataset Based on Complex Environments
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering