SlidesGen-Bench: Evaluating Slides Generation via Computational and Quantitative Metrics
Yunqiao Yang, Wenbo Li, Houxing Ren, Zimu Lu, Ke Wang, Zhiyuan Huang, Zhuofan Zong, Mingjie Zhan, Hongsheng Li
cs.CL
Submitted: 2026-08-21
Updated: 2026-08-24
Code: https://github.com/YunqiaoYang/SlidesGen-Bench
Terminology
Sources
- PaddleOCR 3.0 Technical Report
- A Survey on Code Generation with LLM-based Agents
- Qwen Technical Report
- Enhancing Presentation Slide Generation by LLMs with a Multi-Staged End-to-End Approach
- The Llama 3 Herd of Models
- Mistral 7B
- Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation
- No Free Labels: Limitations of LLM-as-a-Judge Without Human Grounding
- Verbosity Bias in Preference Labeling by Large Language Models
- SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
- LLaMA: Open and Efficient Foundation Language Models
- Llama 2: Open Foundation and Fine-Tuned Chat Models
- OpenHands: An Open Platform for AI Software Developers as Generalist Agents
- Qwen2.5 Technical Report
- Auto-Slides: An Interactive Multi-Agent System for Creating and Customizing Research Presentations
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering