Bergson: An Open Source Library for Data Attribution
cs.LG
Submitted: 2026-06-10
Updated: 2026-09-25
Code: https://github.com/huggingface/peft
Terminology
Sources
- Second-Order Stochastic Optimization for Machine Learning in Linear Time
- Towards Tracing Factual Knowledge in Language Models Back to the Training Data
- Training Data Attribution via Approximate Unrolled Differentiation
- Pythia: A Suite for Analyzing Large Language Models Across Training and Scaling
- Studying Large Language Model Generalization with Influence Functions
- Language Models are Few-Shot Learners
- Shampoo: Preconditioned Stochastic Tensor Optimization
- Scalable Influence and Fact Tracing for Large Language Model Pretraining
- Input Similarity from the Neural Network Perspective
- Optimizing ML Training with Metagradient Descent
- Understanding Black-box Predictions via Influence Functions
- Tulu 3: Pushing Frontiers in Open Language Model Post-Training
- Deep Ignorance: Filtering Pretraining Data Builds Tamper-Resistant Safeguards into Open-Weight LLMs
- Estimating Training Data Influence by Tracing Gradient Descent
- The WMDP Benchmark: Measuring and Reducing Malicious Use With Unlearning
- Procedural Knowledge in Pretraining Drives Reasoning in Large Language Models
- Scaling Up Influence Functions
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- SimpleFSDP: Simpler Fully Sharded Data Parallel with torch.compile
- Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks