Backdoor Mitigation in Decentralized LLM Fine-Tuning
cs.CR, cs.LG
Submitted: 2026-09-29
Updated: 2026-09-29
Code: https://github.com/tatsu-lab/stanford_alpaca
Terminology
Sources
- Constitutional AI: Harmlessness from AI Feedback
- The Trigger in the Haystack: Extracting and Reconstructing LLM Backdoor Triggers
- Measuring the Effects of Non-Identical Data Distribution for Federated Visual Classification
- Weight space Detection of Backdoors in LoRA Adapters
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- BloombergGPT: A Large Language Model for Finance
- Qwen3 Technical Report
- Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs