VASC: Value-Aware Sparse Attention with Cross-Layer Memory for Efficient 3D Reconstruction
cs.CV
Submitted: 2026-10-01
Updated: 2026-10-01
Code: https://github.com/kosakayamahoo-design/VASC
Terminology
Sources
- IndexCache: Accelerating Sparse Attention via Cross-Layer Index Reuse
- Longformer: The Long-Document Transformer
- Token Merging: Your ViT But Faster
- GHOST: Geometry-Hierarchical Online Streaming Token Eviction for Efficient 3D Reconstruction
- Generating Long Sequences with Sparse Transformers
- Not All Patches are What You Need: Expediting Vision Transformers via Token Reorganizations
- TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?
- FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
- Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference
- InfiniteVGGT: Visual Geometry Grounded Transformer for Endless Streams
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models