A critique-and-routing controller cast as a finite-horizon MDP with policy-gradient optimization outperforms one-shot routing baselines on reasoning benchmarks while using the strongest agent for under 25% of calls.
Andrew Estornell et al
3 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
years
2026 3verdicts
UNVERDICTED 3roles
background 1polarities
background 1representative citing papers
CAF-Gen uses an iterative multi-agent creator-reviewer process to enrich shallow argument mining outputs into structurally richer CAF-compliant models with claimed improvements over single-pass generation.
A multi-agent LLM framework with schema enrichment and business rules achieves 78.1% semantic accuracy on the BIRD NL2SQL benchmark.
citing papers explorer
-
Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs
A critique-and-routing controller cast as a finite-horizon MDP with policy-gradient optimization outperforms one-shot routing baselines on reasoning benchmarks while using the strongest agent for under 25% of calls.
-
CAF-Gen: A Multi-Agent System for Enriching Argumentation Structures
CAF-Gen uses an iterative multi-agent creator-reviewer process to enrich shallow argument mining outputs into structurally richer CAF-compliant models with claimed improvements over single-pass generation.
-
AgentNLQ: A General-Purpose Agent for Natural Language to SQL
A multi-agent LLM framework with schema enrichment and business rules achieves 78.1% semantic accuracy on the BIRD NL2SQL benchmark.