BrainBench: New benchmark tests LLMs on comprehensive EEG understanding
From sleep scoring to neurocognition, this benchmark evaluates LLMs on 3,000+ real EEG tasks
Deep Dive
Key Points
- Zhejiang University's BrainBench covers 4 subsets, 17 datasets, and 3,000+ EEG tasks on real data
- LLMs evaluated across 100K+ executions using CodeAct and BrainAgent agentic paradigms
- Outputs validated via numerical, categorical, set, sequence, semantic, and artifact checks
Why It Matters
As LLMs enter neuroscience, BrainBench sets a standard for measuring real-world EEG analysis capability.