iSEE: Object Permanence Through Self-Supervision
cs.CV
Submitted: 2026-10-01
Updated: 2026-10-01
Terminology
Sources
- Invariant Slot Attention: Object Discovery with Slot-Centric Reference Frames
- DDLP: Unsupervised Object-Centric Video Prediction with Deep Dynamic Latent Particles
- Latent Particle World Models: Self-supervised Object-centric Stochastic Dynamics Modeling
- Slot State Space Models
- Reasoning-Enhanced Object-Centric Learning for Videos
- Selective Synergistic Learning for Video Object-Centric Learning
- TSA: Temporal Slot Activation for Persistent Object-Centric Video Representation
- What DINO saw: ALiBi positional encoding reduces positional bias in Vision Transformers
- OCK: Unsupervised Dynamic Video Prediction with Object-Centric Kinematics
- Learning Object Permanence from Videos via Latent Imaginations
- PlaySlot: Learning Inverse Latent Dynamics for Controllable Object-Centric Video Prediction and Planning
- Rethinking Object-Centric Representations for Video Dynamics Modeling
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models