EEGAgentBench: Benchmarking LLM Agents on Short- and Long-Horizon EEG Analysis
cs.LG
Submitted: 2026-09-10
Updated: 2026-09-10
Terminology
Sources
- Deep learning-based electroencephalography analysis: a systematic review
- EEGAgent: A Unified Framework for Automated EEG Analysis Using Large Language Models
- BrainAgent: A Large Language Model-Driven Multi-Agent Framework for Autonomous Brain Signal Understanding
- EEG-Bench: A Benchmark for EEG Foundation Models in Clinical Applications
- NeuralBench: A Unifying Framework to Benchmark NeuroAI Models
- $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
- AstaBench: Rigorous Benchmarking of AI Agents with a Scientific Research Suite
- Holistic Evaluation of Language Models
- SzCORE: A Seizure Community Open-source Research Evaluation framework for the validation of EEG-based automated seizure detection algorithms
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks