Follow the Flow: On Information Flow Across Textual Tokens in Text-to-Image Models
cs.CL
Submitted: 2025-04-01
Updated: 2026-04-16
Code: https://github.com/tokeron/lens
Terminology
Sources
- Fine-grained Analysis of Sentence Embeddings Using Auxiliary Prediction Tasks
- On Identifiability in Transformers
- The Hidden Language of Diffusion Models
- Sparse Autoencoders Find Highly Interpretable Features in Language Models
- SAeUron: Interpretable Concept Unlearning in Diffusion Models with Sparse Autoencoders
- Training-Free Structured Diffusion Guidance for Compositional Text-to-Image Synthesis
- Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
- Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
- GenEval: An Object-Focused Framework for Evaluating Text-to-Image Alignment
- Does Liking Yellow Imply Driving a School Bus? Semantic Leakage in Language Models
- From Tokens to Words: On the Inner Lexicon of LLMs
- SuperBPE: Space Travel for Language Models
- GPT-4 Technical Report
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
- Adversarial Diffusion Distillation
- RAGDiffusion: Faithful Cloth Generation via External Knowledge Assimilation
- Gemini: A Family of Highly Capable Multimodal Models
- Gemma: Open Models Based on Gemini Research and Technology
- Padding Tone: A Mechanistic Analysis of Padding Tokens in T2I Models
- DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering