K-OPSD: Verifiable On-Policy Self-Distillation for Post-Training Vision-Language Models on AEC Drawings
cs.AI, cs.CV
Submitted: 2026-09-28
Updated: 2026-09-28
Terminology
Sources
- Qwen-BIM: developing large language model for BIM-based design with domain-specific benchmark and dataset
- Domain-Specific Fine-Tuning and Prompt-Based Learning: A Comparative Study for developing Natural Language-Based BIM Information Retrieval Systems
- Text2BIM: Generating Building Models Using a Large Language Model-based Multi-Agent Framework
- CEQuest: Benchmarking Large Language Models for Construction Estimation
- DrafterBench: Benchmarking Large Language Models for Tasks Automation in Civil Engineering
- Using Large Language Models for the Interpretation of Building Regulations
- A Generalized LLM-Augmented BIM Framework: Application to a Speech-to-BIM system
- Automatic Building Code Review: A Case Study
- A Large Language Model-Empowered Agent for Reliable and Robust Structural Analysis
- Large Language Model Agent for Structural Drawing Generation Using ReAct Prompt Engineering and Retrieval Augmented Generation
- SDIGLM: Leveraging Large Language Models and Multi-Modal Chain of Thought for Structural Damage Identification
- BIMgent: Towards Autonomous Building Modeling via Computer-use Agents
- Investigating the Potential of Large Language Model-Based Router Multi-Agent Architectures for Foundation Design Automation: A Task Classification and Expert Selection Study
- Large Language Models in Fire Engineering: An Examination of Technical Questions Against Domain Knowledge
- FloorplanVLM: A Vision-Language Model for Floorplan Vectorization
- STaR: Bootstrapping Reasoning With Reasoning
- Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
- Reinforced Self-Training (ReST) for Language Modeling
- Beyond Human Data: Scaling Self-Training for Problem-Solving with Language Models
- On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection