CypherTurn: A Multi-Turn Benchmark for Conversational Text-to-Cypher Evaluation and the Autonomy Divergence
cs.CL
Submitted: 2026-09-29
Updated: 2026-09-29
Code: https://github.com/BarryQ/CypherTurn
Terminology
Sources
- DeepSeek-V3 Technical Report
- GLM-5: from Vibe Coding to Agentic Engineering
- The Llama 3 Herd of Models
- BIRD-INTERACT: Re-imagining Text-to-SQL Evaluation for Large Language Models via Lens of Dynamic Interactions
- Multi-turn Natural Language to Graph Query Language Translation
- Text2GraphQuery-Bench: A Text to Graph Query Benchmark
- STRuCT-LLM: Unifying Tabular and Graph Reasoning with Reinforcement Learning for Semantic Parsing
- Think-on-Graph: Deep and Responsible Reasoning of Large Language Model on Knowledge Graph
- MiMo-V2-Flash Technical Report
- Gemma 2: Improving Open Language Models at a Practical Size
- Kimi K2.5: Visual Agentic Intelligence
- ERNIE 5.0 Technical Report
- Qwen3 Technical Report
- ReAct: Synergizing Reasoning and Acting in Language Models
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering