TukaBench: A Culturally Grounded Jailbreak Benchmark for African Languages
cs.CL, cs.AI
Submitted: 2026-05-31
Updated: 2026-09-22
Terminology
Sources
- UbuntuGuard: A Culturally-Grounded Policy Benchmark for Equitable AI Safety in African Languages
- One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety
- FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios
- Do Methods to Jailbreak and Defend LLMs Generalize Across Languages?
- Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting
- International AI Safety Report
- A Cross-Language Investigation into Jailbreak Attacks in Large Language Models
- Boundary Point Jailbreaking of Black-Box LLMs
- BabelSafe: A Policy-Grounded Multilingual Safety Benchmark for LLMs
- Universal and Transferable Adversarial Attacks on Aligned Language Models
- Structured Semantic Cloaking for Jailbreak Attacks on Large Language Models
- ALERT: A Comprehensive Benchmark for Assessing Large Language Models' Safety through Red Teaming
- Low-Resource Languages Jailbreak GPT-4
- AfriqueLLM: How Data Mixing and Model Architecture Impact Continued Pre-training for African Languages
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering