The Privacy Fallacy of Crowdsourced Fine-Tuning: Extracting Proprietary Data via Topic-Based Poisoning
cs.CR, cs.LG
Submitted: 2026-09-27
Updated: 2026-09-27
Code: https://github.com/thu-coai/Backdoor-Data-Extraction
Terminology
Sources
- Qwen Technical Report
- Extracting alignment data in open models
- Training Verifiers to Solve Math Word Problems
- The Llama 3 Herd of Models
- Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs
- Discovering Universal Activation Directions for PII Leakage in Language Models
- Poisoning Attacks on LLMs Require a Near-constant Number of Poison Samples
- Gemini: A Family of Highly Capable Multimodal Models
- Gemma 3 Technical Report
- OpenMathInstruct-2: Accelerating AI for Math with Massive Open-Source Instruction Data
- A Survey on Knowledge Distillation of Large Language Models
- ShareChat: A Dataset of Chatbot Conversations in the Wild
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs