General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level
cs.CL, cs.AI
Submitted: 2023-11-23
Updated: 2026-08-31
Code: https://github.com/BingkangShi/general-phrase-debiaser
Terminology
Sources
- BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
- RoBERTa: A Robustly Optimized BERT Pretraining Approach
- DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
- You Only Compress Once: Towards Effective and Elastic BERT Compression via Exploit-Explore Stochastic Nature Gradient
- IDEAL: Influence-Driven Selective Annotations Empower In-Context Learners in Large Language Models
- HyperTime: Hyperparameter Optimization for Combating Temporal Distribution Shifts
- AutoGen: Enabling Next-Gen LLM Applications via Multi-Agent Conversation
- MathChat: Converse to Tackle Challenging Math Problems with LLM Agents
- Measuring and Reducing Gendered Correlations in Pre-trained Models
- Refined Coreset Selection: Towards Minimal Coreset Size under Model Performance Constraints
- Decoupled Weight Decay Regularization
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering