Introduces NeuroDoc and NeuroAudit to create a community-reviewed corpus of 53 EEG benchmark entries with 245 task definitions using a rulebook-guided task document and executable kernel.
URL https://proceedings.neurips.cc/paper_files/paper/2024/file/ 9547b09b722f2948ff3ddb5d86002bc0-Paper-Datasets_and_Benchmarks_Track.pdf
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
years
2026 2verdicts
UNVERDICTED 2representative citing papers
BUILD-AND-FIND is a protocol that measures how accurately and efficiently downstream agents can recover hidden repository specifications from AI-generated codebases using accuracy, repeatability, coverage, and effort metrics.
citing papers explorer
-
EEG Benchmarking Needs a Task Specification Layer: NeuroDoc for Rulebook-Guided, Executable Benchmark Construction
Introduces NeuroDoc and NeuroAudit to create a community-reviewed corpus of 53 EEG benchmark entries with 245 task definitions using a rulebook-guided task document and executable kernel.
-
BUILD-AND-FIND: An Effort-Aware Protocol for Evaluating Agent-Managed Codebases
BUILD-AND-FIND is a protocol that measures how accurately and efficiently downstream agents can recover hidden repository specifications from AI-generated codebases using accuracy, repeatability, coverage, and effort metrics.