Many Dialects, Many Languages, One Cultural Lens: Evaluating Multilingual VLMs for Bengali Culture Understanding Across Historically Linked Languages and Regional Dialects
cs.CL, cs.CV
Submitted: 2026-03-22
Updated: 2026-08-27
Comments: Accepted at EMNLP 2026 (Findings)
Project page: https://labib1610.github.io/BanglaVerse
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- ChitroJera: A Regionally Relevant Visual Question Answering Dataset for Bangla
- IndicVisionBench: Benchmarking Cultural and Multilingual Understanding in VLMs
- CultureVLM: Characterizing and Improving Cultural Understanding of Vision-Language Models for over 100 Countries
- DRISHTIKON: A Multimodal Multilingual Benchmark for Testing Language Models' Understanding on Indian Culture
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering