Shared Geometry As A Rosetta Stone: Cross-Modal Alignment Without Paired Data
cs.LG, cs.AI, cs.CV
Submitted: 2026-10-07
Updated: 2026-10-08
Code: https://github.com/bioptimus/releases
Terminology
Sources
- Granite Embedding Models
- Qwen3-VL Technical Report
- Microsoft COCO Captions: Data Collection and Evaluation Server
- mini-vec2vec: Scaling Universal Geometry Alignment with Linear Transformations
- Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale
- Brain-to-Text Decoding: A Non-invasive Approach via Typing
- Towards General Text Embeddings with Multi-stage Contrastive Learning
- Platonic Representations in the Human Brain: Unsupervised Recovery of Universal Geometry
- Exploiting Similarities among Languages for Machine Translation
- Vision Foundation Models for Computed Tomography
- Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- Text Embeddings by Weakly-Supervised Contrastive Pre-training
- Qwen3 Technical Report
- What Converges in the Platonic Representation Hypothesis? Structure over Geometry
- Jasper and Stella: distillation of SOTA embedding models
- Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks