Atomizer-IO: Beyond Pixels, Patches and Grids
cs.CV
Submitted: 2026-09-30
Updated: 2026-09-30
Terminology
Sources
- AnySat: One Earth Observation Model for Many Resolutions, Scales, and Modalities
- An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
- Lightweight Temporal Self-Attention for Classifying Satellite Image Time Series
- FRACTAL: An Ultra-Large-Scale Aerial Lidar Dataset for 3D Semantic Segmentation of Diverse Landscapes
- Deep Learning for 3D Point Clouds: A Survey
- xBD: A Dataset for Assessing Building Damage from Satellite Imagery
- ForestNet: Classifying Drivers of Deforestation in Indonesia using Deep Learning on Satellite Imagery
- Perceiver IO: A General Architecture for Structured Inputs & Outputs
- Foundation Models for Generalist Geospatial Artificial Intelligence
- FlexiMo: A Flexible Remote Sensing Foundation Model
- PANGAEA: A Global and Inclusive Benchmark for Geospatial Foundation Models
- Lightweight, Pre-trained Transformers for Remote Sensing Timeseries
- Galileo: Learning Global & Local Features of Many Remote Sensing Modalities
- Neural Plasticity-Inspired Multimodal Foundation Model for Earth Observation
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models