The paper presents ChildAgentEval as the first psychometrically grounded benchmark comparing MLLM-based agents' reasoning performance to age-specific human cognitive stages.
The cognitive capabilities of generative ai: A comparative analysis with human benchmarks
3 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
roles
background 1polarities
background 1representative citing papers
A data-driven SU(3)-breaking analysis of B to PP decays yields QCD-factorization amplitudes that resemble dynamical predictions and require no enhanced annihilation terms.
LHCb measures f_L^d = 0.600 and f_L^s = 0.159 for B to K* Kbar* decays and reports a ratio L of 4.92 that confirms 4.4 sigma discrepancy with theory.
citing papers explorer
-
Evaluating Cognitive Age Alignment in Interactive AI Agents
The paper presents ChildAgentEval as the first psychometrically grounded benchmark comparing MLLM-based agents' reasoning performance to age-specific human cognitive stages.
-
QCD-factorization amplitudes from flavour symmetries: beyond the $SU(3)$ symmetric case
A data-driven SU(3)-breaking analysis of B to PP decays yields QCD-factorization amplitudes that resemble dynamical predictions and require no enhanced annihilation terms.
-
Measurement of the branching fractions and longitudinal polarisations of $B^0_{(s)} \to K^{*0} \kern 0.18em \overline{\kern -0.18em K}{}^{*0}$ decays
LHCb measures f_L^d = 0.600 and f_L^s = 0.159 for B to K* Kbar* decays and reports a ratio L of 4.92 that confirms 4.4 sigma discrepancy with theory.