ScenePilot: Grow-and-Repair Policy for Text-Driven 3D Indoor Scene Generation
cs.CV, cs.AI
Submitted: 2026-08-31
Updated: 2026-08-31
Project page: https://zjw-louie.github.io/ScenePilot
Terminology
Sources
- SAGE: Scalable Agentic 3D Scene Generation for Embodied AI
- InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
- LayoutGPT: Compositional Visual Planning and Generation with Large Language Models
- LayoutVLM: Differentiable Optimization of 3D Layout via Vision-Language Models
- SceneWeaver: All-in-One 3D Scene Synthesis with an Extensible and Self-Reflective Agent
- Text-to-Scene with Large Reasoning Models
- EditThinker: Unlocking Iterative Reasoning for Any Image Editor
- ReasonEdit: Towards Reasoning-Enhanced Image Editing Models
- RoomDreamer: Text-Driven 3D Indoor Scene Synthesis with Coherent Geometry and Texture
- MetaSpatial: Reinforcing 3D Spatial Reasoning in VLMs for the Metaverse
- Direct Numerical Layout Generation for 3D Indoor Scene Synthesis via Spatial Reasoning
- DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
- GPT-4o System Card
- OpenAI GPT-5 System Card
- Qwen3-VL Technical Report
- Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models
- Billion-scale similarity search with GPUs
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models