Pith. sign in

Paper Citation Record · LEDGER

RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2503.10657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.10657 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:18:54.116775Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T03:46:44.819123Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c8c53eb2-396d-4395-b2ca-5008e9421336 · inbound

Universal Model Routing for Efficient LLM Inference cites this paper.

Universal Model Routing for Efficient LLM Inference RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T23:48:00.958174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:48:00.958174Z digest=sha256:486e57fad0b2fe65cc81799b39dc7336f1d23475ffaa01037d0c3d149cfd74e3

Observation f9b76be3-78ac-49d5-82ed-63d56d03b83a · inbound

LightRouter: Towards Efficient LLM Collaboration with Minimal Overhead cites this paper.

LightRouter: Towards Efficient LLM Collaboration with Minimal Overhead RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:11:17.431359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:11:17.431359Z digest=sha256:fbd2de06f842cea47b7d36ac044dd6581d3a7398ab51dea28c6f13b563fbfcc9

Observation 2deb8c0a-17a0-472d-bf21-777bba6b2774 · inbound

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants cites this paper.

The Avengers: A Simple Recipe for Uniting Smaller Language Models to Challenge Proprietary Giants RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:37.520825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:37.520825Z digest=sha256:1e53b6dc4089902eac3a63a3a495c274ffc974af9f57b83fa43e1ada29404b44

Observation e13a777d-61db-46c8-bd3e-3388bac034aa · inbound

Towards Efficient Multi-LLM Inference: Characterization and Analysis of LLM Routing and Hierarchical Techniques cites this paper.

Towards Efficient Multi-LLM Inference: Characterization and Analysis of LLM Routing and Hierarchical Techniques RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:56:56.742261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:56:56.742261Z digest=sha256:c29720c799115dda0d3f87c754fdf4edf40e9af6ee1b1526d336ca4b5342315f

Observation d102ab14-2f5e-4e1f-b484-6d775d8326ee · inbound

DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label Recognition cites this paper.

DART: Dual Adaptive Refinement Transfer for Open-Vocabulary Multi-Label Recognition RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T23:23:08.355033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:23:08.355033Z digest=sha256:b85815e780e48df35cac7e3cd6956e7482d1d34b08236b0769b6fdd8142d2e45

Observation 71724254-d1c4-43b2-b4fc-3e1f8ff3fd70 · inbound

A Greedy PDE Router for Blending Neural Operators and Classical Methods cites this paper.

A Greedy PDE Router for Blending Neural Operators and Classical Methods RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:32:36.156349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:32:13.577474Z digest=sha256:0b1fde008caf7d06c162eff4d9f97d9b79c94e5192c3f3ce480c717a7a0cfc66

Observation a8de10d9-19ca-4a24-bc9f-ac458b76624f · inbound

R2-Router: A New Paradigm for LLM Routing with Reasoning cites this paper.

R2-Router: A New Paradigm for LLM Routing with Reasoning RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T05:17:29.774492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:17:29.774492Z digest=sha256:90f196f0fe3d75f47981c985399327383646ed18e75f3bb7f24f13788c3f25b7

Observation e677c21e-75b1-4dde-a390-b216535a7791 · inbound

Impacts of Aggregation on Model Diversity and Consumer Utility cites this paper.

Impacts of Aggregation on Model Diversity and Consumer Utility RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T20:32:11.860234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:32:11.860234Z digest=sha256:038b6e7653356d93027517ae0fe9f3938f28bf0c3fe600543adc7ecfc2913f3b

Observation a13b1713-3b64-44fb-9594-e0881eaa6342 · inbound

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent cites this paper.

AgentOpt v0.1 Technical Report: Client-Side Optimization for LLM-Based Agent RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:25:52.580766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T19:51:39.564680Z digest=sha256:726a35718e1da15fd81153fcde8c9d3fa561559386415f5d282cc3934b2c84bf

Observation eb45bc2a-666a-4087-b0ea-097eaaaf8966 · inbound

ModelLens: Finding the Best for Your Task from Myriads of Models cites this paper.

ModelLens: Finding the Best for Your Task from Myriads of Models RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:59.512351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-11T01:55:57.223673Z digest=sha256:c42c5e5315d0094195f761c1ca77586ef69c5c4f37a7a39e0a1459e909bad194

Observation 73e15b42-552a-4048-92c9-241bb1b828d1 · inbound

SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning cites this paper.

SeqRoute: Global Budget-Aware Sequential LLM Routing via Offline Reinforcement Learning RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:14:01.051611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T23:13:41.456647Z digest=sha256:e94116ef84c3fca95d0995f4131509e23e8ebbda7bc0d34692fe9fcd0ed641d8

Observation f97faf56-7698-4646-a1b1-41303bc13c05 · inbound

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks cites this paper.

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-02T18:37:16.528637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-02T18:19:43.146102Z digest=sha256:0b8046eb304c46d45d0ea566dde30cd60ccd39c2207057886d0bc481b47b3f64

Observation 7e2939c7-f61e-4730-8b58-d7796b92fc78 · inbound

How Much of the Routing Gap Is Real? Decomposing the Router-to-Oracle Gap into Reproducible Specialist Advantage and Single-Draw Label Noise cites this paper.

How Much of the Routing Gap Is Real? Decomposing the Router-to-Oracle Gap into Reproducible Specialist Advantage and Single-Draw Label Noise RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T02:29:58.886380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T02:29:58.886380Z digest=sha256:82af576a3185f6926d8a996b7af42374d00a9e6e994e6c8088620b199b9decdf

Observation 111bb04e-7c64-47e8-88e0-5a68f44b1bb2 · inbound

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models cites this paper.

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-07-10T03:46:44.820435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-10T03:42:04.605214Z digest=sha256:2b9817792acc5b2e3fc039d9f441a8c0b62574f1546913be9c4ebd3473c9bc1a

Observation 11c8c7dd-b43e-4970-9362-f215dfb2c6e7 · inbound

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models cites this paper.

Resample or Reroute? Budget-Aware Test-Time Model Selection for Large Language Models RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T06:32:41.000897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T06:32:41.000897Z digest=sha256:ef7f77bfb73addc1734989b60ebc2cf21978b4b1a49802170e7c11f05516d352

Observation a3809a90-6dbd-4c46-a8ad-aa7e746c43cc · inbound

Adapting Embedding Models for Agent Capability Retrieval cites this paper.

Adapting Embedding Models for Agent Capability Retrieval RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T18:19:38.026339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:19:38.026339Z digest=sha256:c58999f0bf2f44c5a874533d31c873fd175c3fbab086e32791ffbc7dd4e92b9e

Observation 84fe4562-b79c-4a38-a339-a776b79ced88 · inbound

HACO: Hedged Agent Computing for Reliable LLM Systems cites this paper.

HACO: Hedged Agent Computing for Reliable LLM Systems RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:58.965763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:58.965763Z digest=sha256:b736a2fc238a1ba873ce91a1c6ba9e835cbb0eeb1ba62c5192d3011c79861b67

Observation acd73c12-76f2-472b-96cb-804c6265671b · inbound

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World cites this paper.

The Replay Gap: Static Evaluation of Model Switching in LLM Agents Scores the Wrong World RouterEval: A Comprehensive Benchmark for Routing LLMs to Explore Model-level Scaling Up in LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T00:18:54.116775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:18:54.116775Z digest=sha256:db004f8c5c9d1de03d8e0ba607b4346d2a40efb92ebc47557b6510156cd3234d