FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection
cs.CL, cs.AI, cs.CV, cs.MM
Submitted: 2026-05-29
Updated: 2026-08-29
Code: https://github.com/MaartenGr/BERTopic
Terminology
Sources
- Evaluating ChatGPT's Performance for Multilingual and Emoji-based Hate Speech Detection
- QLoRA: Efficient Finetuning of Quantized LLMs
- What matters when building vision-language models?
- In-context Vectors: Making In Context Learning More Effective and Controllable Through Latent Space Steering
- Read as You See: Guiding Unimodal LLMs for Low-Resource Explainable Harmful Meme Detection
- STEMTOX: From Collaborative Tags to Fine-Grained Toxic Meme Detection via Entropy-Guided Multi-Task Learning
- The Enforcement and Feasibility of Hate Speech Moderation
- VL-CheckList: Evaluating Pre-trained Vision-Language Models with Objects, Attributes and Relations
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering