Omni2Web: Benchmarking Audiovisual Website Development
cs.SE, cs.CV
Submitted: 2026-09-20
Updated: 2026-09-20
Project page: https://omni2web-bench.github.io
Terminology
Sources
- ScreenAI: A Vision-Language Model for UI and Infographics Understanding
- MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction
- WebVR: Benchmarking Multimodal LLMs for WebPage Recreation from Videos via Human-Aligned Visual Rubrics
- Nemotron 3 Nano Omni: Efficient and Open Multimodal Intelligence
- MMOU: A Massive Multi-Task Omni Understanding and Reasoning Benchmark for Long and Complex Real-World Videos
- WebLINX: Real-World Website Navigation with Multi-Turn Dialogue
- Muse Spark Safety & Preparedness Report
- Qwen3-ASR Technical Report
- LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs
- Gemma 4 Technical Report
- Kimi K3: Open Frontier Intelligence
- Qwen3.5-Omni Technical Report
- Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
- FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties