Embedded Conditional Independence Tests for Large Language Model Generated Text with an Application to German Parliament Speeches
stat.ML, cs.AI, cs.LG, math.ST, stat.ME, stat.TH
Submitted: 2026-09-01
Updated: 2026-09-01
Terminology
Sources
- jina-embeddings-v5-text: Task-Targeted Embedding Distillation
- The Impossibility of Fair LLMs
- On the Robustness of Multilingual Text Embedding Rankings Across Learning Tasks, Languages, and Benchmark Datasets
- Intrinsic Bias Metrics Do Not Correlate with Application Bias
- A Survey on Fairness in Large Language Models
- NV-Retriever: Improving text embedding models with effective hard-negative mining
- New Encoders for German Trained from Scratch: Comparing ModernGBERT with Converted LLM2Vec Models
- Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
Related papers
- Behavior of prediction performance metrics with rare events
- Optimal Estimation of Generic Dynamics by Path-Dependent Neural Jump ODEs
- A Posterior-Dynamics Framework for Imaging Inverse Problems with Pretrained Diffusion Priors
- One Permutation Is All You Need: Fast, Deterministic Feature Importance and Model Stress-Testing
- Online Conformal Prediction for Non-Exchangeable Panel Data
- Deep Time-Series Forecasting in 10 Years: A Survey