Pith. sign in

Paper Citation Record · LEDGER

Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 68 inbound Pith citation observations for arXiv:2404.14618.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.14618 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 68 of 68 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:29:18.681178Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6fe7d8a4-dcff-4e04-9d3b-c9b7afcb8f1e · inbound

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees cites this paper.

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:53:21.456777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-23T18:49:39.718108Z digest=sha256:8eaddc6fe795620ea6ef8ae8669a7abd1cef38eff530b8dc2bb895c806949b56

Observation b6318681-61bf-4aff-8b61-72ff304c87e8 · inbound

Real-time Adapting Routing (RAR): Improving Efficiency Through Continuous Learning in Software Powered by Layered Foundation Models cites this paper.

Real-time Adapting Routing (RAR): Improving Efficiency Through Continuous Learning in Software Powered by Layered Foundation Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T20:20:59.490706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:20:59.490706Z digest=sha256:efa7d0a21c8c79d5977d60decdf923da02df42f0b6a784d5187d7a31d321fb3c

Observation 771beb8d-25aa-4611-8297-56c34aa4a62f · inbound

Adaptive Routing of Text-to-Image Generation Requests Between Large Cloud Model and Light-Weight Edge Model cites this paper.

Adaptive Routing of Text-to-Image Generation Requests Between Large Cloud Model and Light-Weight Edge Model Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T16:16:26.203856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:16:26.203856Z digest=sha256:82ee3174cdb4a9e86a31b78e9ec17ee70d343b5b0d3b6293c3b998300fe4eb2c

Observation e5d89cde-ccfc-4b19-bac1-99bf2d7ce8c4 · inbound

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling cites this paper.

DiffServe: Efficiently Serving Text-to-Image Diffusion Models with Query-Aware Model Scaling Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T14:27:01.652242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:27:01.652242Z digest=sha256:06c9daaf6591a20e89dd818d789aa18916f80316c2af4a427af16c72aa2030d6

Observation 8d61f77b-aefb-4aea-8e61-98eba6932266 · inbound

DBRouting: Routing End User Queries to Databases for Answerability cites this paper.

DBRouting: Routing End User Queries to Databases for Answerability Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T13:39:45.298403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:39:45.298403Z digest=sha256:70106e8966e8529147fd7615500ff7cdbda38c6171ed2406cd30cfc7cd04e4ba

Observation f36df912-854e-40b9-b8d0-d525ae3e5d4c · inbound

What should an AI assessor optimise for? cites this paper.

What should an AI assessor optimise for? Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T19:24:07.302286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:24:07.302286Z digest=sha256:8b676a4390f87a2a6418037aac90a6dc58cd396e58e90056158f0e875872cc00

Observation eec7d799-9a97-44c5-bc9f-91d2f33b4537 · inbound

Fast Large Language Model Collaborative Decoding via Speculation cites this paper.

Fast Large Language Model Collaborative Decoding via Speculation Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T19:34:42.766560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:34:42.766560Z digest=sha256:e431fe696da77fcf370dacea4565f9ff4f8667a0777eee8497e5d1df80ea12b1

Observation 5a732764-f1be-4d51-913d-1c3427b92698 · inbound

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing cites this paper.

CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T13:58:45.307926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:58:45.307926Z digest=sha256:8839bbc855daa7154df4c0b677f0f1e0b2c572d2e4b5f62882cf0f9db1d3880f

Observation 68fd614a-c29e-46c9-92dd-bf214f6c1da0 · inbound

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing cites this paper.

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-09T11:22:28.858028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:22:28.858028Z digest=sha256:2a456503b5b37c1e547ad86b857ed6c39c84b2ec3262d501a8312970533423db

Observation 8884f231-6802-4c97-803d-b7e0fa6c4ec6 · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.192576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.192576Z digest=sha256:de17ea0428404ca2a4f7dbf6de26b84b2f309bb4dc1be14cd9f5b4dcce822b41

Observation 3dd1f6a2-1801-489f-9db0-1ee819644210 · inbound

MixLLM: Dynamic Routing in Mixed Large Language Models cites this paper.

MixLLM: Dynamic Routing in Mixed Large Language Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T18:11:31.690286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T18:11:31.690286Z digest=sha256:f240250dbcc337040e986c7c5108d633d49b587c5a8b7669536865bcf8ad8b2c

Observation 6501e6a1-0502-48f4-8399-773f84d5a93e · inbound

Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models cites this paper.

Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T10:29:18.681178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T10:29:18.681178Z digest=sha256:1f7a36959a0c19289162bf7dc698e9d8ba2191a7e2f4a5698239e945c4eb5bc2

Observation 366a8899-7733-4dba-b62c-f194f34143e2 · inbound

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers cites this paper.

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:14:57.444574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T15:13:28.927880Z digest=sha256:36e4cfdab49caa18857e538284dce58ba0b5a08f2ddb5f87d68f16940734493e

Observation 4bb8de0c-e21e-44ac-85a3-681a67684fd3 · inbound

IRT-Router: Effective and Interpretable Multi-LLM Routing via Item Response Theory cites this paper.

IRT-Router: Effective and Interpretable Multi-LLM Routing via Item Response Theory Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T12:00:09.976813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:00:09.976813Z digest=sha256:e5c64fe75bdc3c7b98a458a402e55f3d6fdf6be858c53aa731b4322a9e8d5a3d

Observation 369e619b-111c-4d37-9083-b7527010f245 · inbound

Towards Efficient Multi-LLM Inference: Characterization and Analysis of LLM Routing and Hierarchical Techniques cites this paper.

Towards Efficient Multi-LLM Inference: Characterization and Analysis of LLM Routing and Hierarchical Techniques Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T05:56:56.936401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:56:56.936401Z digest=sha256:d6a488da62236725ef0e5d58817d87b87b8ce71c7b20f2e6d2dbef22db6d7bba

Observation 8ee36174-92ea-4362-a257-d4b9ca289102 · inbound

FAA Framework: A Large Language Model-Based Approach for Credit Card Fraud Investigations cites this paper.

FAA Framework: A Large Language Model-Based Approach for Credit Card Fraud Investigations Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:08:12.871385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:08:12.871385Z digest=sha256:03abaa7d30e6b2113c307ebb7ac572ac633d644a2eee134f83afd6a4d5b7baa3

Observation 412ce91b-9870-4e91-aef4-efcfeedc9b4e · inbound

Semantic Scheduling for LLM Inference cites this paper.

Semantic Scheduling for LLM Inference Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T01:09:20.594776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:09:20.594776Z digest=sha256:dbd18e1f77afb9f35a08c9c05469dbbb13add800a5239a85805d8a6405611d0b

Observation 2802386e-945a-48bf-86ad-3d2abdc552f9 · inbound

Divide, Specialize, and Route: A New Approach to Efficient Ensemble Learning cites this paper.

Divide, Specialize, and Route: A New Approach to Efficient Ensemble Learning Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:44:04.913416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:44:04.913416Z digest=sha256:b363fe86a068c661e051240c729307d211cbb5b0fa16d1b65602a97bfaceb6ff

Observation 7c7ce47c-0123-4fdf-ab94-999afa9ad0fa · inbound

Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration cites this paper.

Toward Edge General Intelligence with Multiple-Large Language Model (Multi-LLM): Architecture, Trust, and Orchestration Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-06T21:16:14.892650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:16:14.892650Z digest=sha256:770670f5cf2eb2982d514a0fb6f697c52fc177ea2f86cdb1ecb743c8139de480

Observation 19b925c4-85bc-4eae-898d-2bc195724bd9 · inbound

Orchestration for Domain-specific Edge-Cloud Language Models cites this paper.

Orchestration for Domain-specific Edge-Cloud Language Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:10:13.650814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:10:13.650814Z digest=sha256:275a45bd8828718d9ee1cdab5f1a618287d42aaf89516ee6ac633fcdcf5cab2f

Observation cf501460-914e-4ed4-90cb-128d29a74534 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T17:53:45.603078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:53:45.603078Z digest=sha256:f3537b380600bd42624f136decbabe30bc50acefa97d724d78f17891e55e5e85

Observation 21b24a98-50b1-46c0-8bd2-5bd6c9f8990d · inbound

DSSD: Efficient Edge-Device LLM Deployment and Collaborative Inference via Distributed Split Speculative Decoding cites this paper.

DSSD: Efficient Edge-Device LLM Deployment and Collaborative Inference via Distributed Split Speculative Decoding Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:05:02.313837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:05:02.313837Z digest=sha256:d594e76b790496c913cfcc56610d61730e50f072c34eddbef390b091691fa9b8

Observation e5f72bfc-ae62-483f-9b75-0d0b2def476b · inbound

CoE-Ops: Collaboration of LLM-based Experts for AIOps Question-Answering cites this paper.

CoE-Ops: Collaboration of LLM-based Experts for AIOps Question-Answering Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T18:09:39.642136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:09:39.642136Z digest=sha256:a6a945b04f7f61e0b930bce1aa8654c3a74bbfef118f2cf33c10aa9ede4b3126

Observation 5d215f12-f43b-4d2c-be98-bc4d39c890e6 · inbound

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts cites this paper.

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T10:22:42.491345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:22:42.491345Z digest=sha256:50e5ece3cded47df01ba30aec13f04b11e5d86c78ba3173a3537a5c1c4df6777

Observation ccacbe4c-6a93-4c8f-8825-a28c2b94f68d · inbound

Balancing Information Accuracy and Response Timeliness in Networked LLMs cites this paper.

Balancing Information Accuracy and Response Timeliness in Networked LLMs Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T05:12:30.570192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:12:30.570192Z digest=sha256:c581d5dc41a85e1eb0a5a2085d2a0c51320616c3153c0a35aaaa5d8e3768a201

Observation d39c3808-a077-4ff5-b007-37eca2617bec · inbound

Adaptive LLM Routing under Budget Constraints cites this paper.

Adaptive LLM Routing under Budget Constraints Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T14:39:27.856593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T14:39:27.856593Z digest=sha256:2aa605713919cf03df49c1a04ab183d6d2fa0e2c8132dc840d82f1b74e4d91b0

Observation e048fe38-e460-433f-9515-8b4fc77ff1c0 · inbound

Towards Generalized Routing: Model and Agent Orchestration for Adaptive and Efficient Inference cites this paper.

Towards Generalized Routing: Model and Agent Orchestration for Adaptive and Efficient Inference Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T22:03:00.916850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:03:00.916850Z digest=sha256:ffb58221b822b8ee1e4587238e16953653340c4fd10c78ddeeaeb8b447704094

Observation ad557b72-5018-4611-9602-92a160538a6e · inbound

A Greedy PDE Router for Blending Neural Operators and Classical Methods cites this paper.

A Greedy PDE Router for Blending Neural Operators and Classical Methods Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:32:36.186197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T12:32:13.577474Z digest=sha256:be9735e8e8704122af45e3c48e79c18b492a08eced5f974629a359750454a6cb

Observation 44a14727-bea3-4317-881d-e25744decb6b · inbound

Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration cites this paper.

Optimal-Agent-Selection: State-Aware Routing Framework for Efficient Multi-Agent Collaboration Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T00:16:17.237087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:16:17.237087Z digest=sha256:314d1dc967502e64cb1418640d5dfb2ecc8339f2a1e4edcd47ef71a92161997b

Observation 7e7a120b-3b85-4949-ad5b-e459ff4360e2 · inbound

GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts cites this paper.

GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T15:58:04.079256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T15:53:07.496412Z digest=sha256:a51372b2c21c1fc23bf49cc472bfe05762bf7a8cdb47920bc6a618b040013b84

Observation 7e868d7e-b58e-4fd7-b42f-1972f7ed8be8 · inbound

The Workload-Router-Pool Architecture for LLM Inference Optimization: A Vision Paper from the vLLM Semantic Router Project cites this paper.

The Workload-Router-Pool Architecture for LLM Inference Optimization: A Vision Paper from the vLLM Semantic Router Project Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:45:12.103211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T06:40:27.945478Z digest=sha256:e5ae4586dc4227ed7dcb9f3dcc90530cfb0658e00692c26aabe37f13bf5f813e

Observation 0a38c288-737d-49b7-be35-e96195d62836 · inbound

Triage: Routing Software Engineering Tasks to Cost-Effective LLM Tiers via Code Quality Signals cites this paper.

Triage: Routing Software Engineering Tasks to Cost-Effective LLM Tiers via Code Quality Signals Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:15:59.510991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T17:15:39.613112Z digest=sha256:539d773f1b617f3901191a6d705ae3b7dc3bc22abf2857c8796ba262c9c5447e

Observation 640d6abd-6be1-4c70-918b-e97be19d732f · inbound

RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving cites this paper.

RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T21:25:52.318641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T16:36:25.868623Z digest=sha256:25b71638fd32c899df592fbe12a01c97741c9ebbd8db36993b142b56a51f132e

Observation f85d4c6c-da72-4428-bedd-6c8f196714c8 · inbound

Adaptive Test-Time Compute Allocation for Reasoning LLMs via Constrained Policy Optimization cites this paper.

Adaptive Test-Time Compute Allocation for Reasoning LLMs via Constrained Policy Optimization Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:00:22.103907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T11:57:44.680423Z digest=sha256:b6bdbb4a954dd2cd7a0f998c9ee159cecdad2a24457634dccbf76660a6041820

Observation 215646a6-d464-4d45-bcea-891b2adc732d · inbound

Privacy-Preserving LLMs Routing cites this paper.

Privacy-Preserving LLMs Routing Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T08:32:51.615035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T08:32:43.015370Z digest=sha256:60bf5c795a3ee4b164a80d75f1e175b9845a1228a6f7d985c3d8a623afbcb522

Observation f9c98f97-997b-4b6e-9584-7a9fdf3c970d · inbound

CADMAS-CTX: Contextual Capability Calibration for Multi-Agent Delegation cites this paper.

CADMAS-CTX: Contextual Capability Calibration for Multi-Agent Delegation Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:36:02.103883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T05:31:27.700576Z digest=sha256:75efc2cb743bf8a82e6afe83fa8e6c036a5f8b2c9511d6662348903a1693a2cb

Observation 67bab197-44c1-45c8-96d4-b204ed01084b · inbound

AgentFloor: How Far Up the tool use Ladder Can Small Open-Weight Models Go? cites this paper.

AgentFloor: How Far Up the tool use Ladder Can Small Open-Weight Models Go? Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:21:10.580723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T20:09:11.566825Z digest=sha256:a58392efe863a3f3e234ca9c82a8d25181b787b9eb984d00dcc85025da0870dd

Observation d695ccf2-cbaa-443f-a6e4-4ed6a6617e06 · inbound

When Less is Enough: Efficient Inference via Collaborative Reasoning cites this paper.

When Less is Enough: Efficient Inference via Collaborative Reasoning Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:42.435295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T19:27:04.267404Z digest=sha256:159615e295fa1a149d862eedc108a649ae1a1b7018d4b71312aa849b1735fdad

Observation 23200814-c51d-4eaf-8b8b-43cd74594acc · inbound

A Regime Theory of Controller Class Selection for LLM Action Decisions cites this paper.

A Regime Theory of Controller Class Selection for LLM Action Decisions Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:11:11.881896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T10:01:45.712230Z digest=sha256:81b6475754772eef5aea20b93b01dab8dc59cd0baedd56e1d418640c8b62c2ec

Observation 13de6052-996d-4b6d-afcb-5a5a0906fe41 · inbound

Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs cites this paper.

Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:26:24.443387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T01:10:03.459912Z digest=sha256:9affa165741519c403fb99f5872000794d9a51152fca6ea6c9cc2fe27c2fa8a8

Observation 34a1482e-0d75-4068-94ac-7316c4a03680 · inbound

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer? cites this paper.

LatentRouter: Can We Choose the Right Multimodal Model Before Seeing Its Answer? Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:47:04.430923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T01:42:54.802658Z digest=sha256:c2e1c14b81c354193c0a69d8959072fc23eb8209500836714e29e9b6f43a8e7b

Observation 0625b803-2811-4448-918e-4e50b86d9a65 · inbound

SOMA: Efficient Multi-turn LLM Serving via Small Language Model cites this paper.

SOMA: Efficient Multi-turn LLM Serving via Small Language Model Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:42:04.202613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T01:37:27.361504Z digest=sha256:a837aecbb30b1f8bbaaa588053172e7fd19ff6afaa55bac2662dd712a8c745f6

Observation e84027dd-5522-42fb-b1ca-e88c90a796cd · inbound

R2V Agent: Teaching SLMs When to Ask for Help cites this paper.

R2V Agent: Teaching SLMs When to Ask for Help Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:28:59.795014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T20:26:29.804427Z digest=sha256:220f05e4dadc5243d00a7fa57d07382da3c145e823cdf442072a86eb60da35c5

Observation ece29f62-8da1-4c0c-8045-1ae8155bbe94 · inbound

HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools cites this paper.

HyDRA: Hybrid Dynamic Routing Architecture for Heterogeneous LLM Pools Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T15:13:31.963292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T15:13:27.898371Z digest=sha256:316b9583678f0307fa3751a13ba60b73092e4da180194f87dc8e17247f1e361a

Observation f3cf284a-ea45-46f2-a226-f37bfd2022c1 · inbound

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing cites this paper.

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:23:50.775746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T23:23:45.561978Z digest=sha256:a3c5f366b12ad1a06fe52cc2f4f53d735fe7cc812764d0d369c74fd347d3f0b1

Observation 539a785e-176b-415a-8222-bf385bf10222 · inbound

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows cites this paper.

DecisionBench: A Benchmark for Emergent Delegation in Long-Horizon Agentic Workflows Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:18:11.676742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-20T10:16:38.920528Z digest=sha256:d95ace3b499d6d8749ef4760dc335aac81b9a116de5f62d99b8150491daf7732

Observation cd027529-07cf-4176-ad26-7c88dcfe1cd4 · inbound

Triaging Threats to Specialized Guardrails cites this paper.

Triaging Threats to Specialized Guardrails Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:32:44.295622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T22:27:44.703466Z digest=sha256:3fb207415540ce7eda59dea6c984a65af5c4297eb202e1a24acbf67fa064d2d9

Observation ce4b2b2f-585a-4198-abbe-6d24bf7ea9b6 · inbound

Trading Human Curation for Synthetic Augmentation in RLVR cites this paper.

Trading Human Curation for Synthetic Augmentation in RLVR Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:16:26.230031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T11:10:06.685591Z digest=sha256:88b94273e878cab3ec18d8731f4f68eaaa3c8393d2fc871c55bb546fbf2b95b5

Observation 6236a340-eb67-4ff2-934c-52bd733986ff · inbound

Trading Human Curation for Synthetic Augmentation in RLVR cites this paper.

Trading Human Curation for Synthetic Augmentation in RLVR Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T15:15:09.675778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T15:15:09.675778Z digest=sha256:0ebc538f70d1653810efef70ee7795b2cf04a9935ba0d54db2a6d3ef27fb6d5a

Observation 9827fbc2-9381-49aa-b420-8dcac91c7e5c · inbound

DLLG: Dynamic Logit-Level Gating of LLM Experts cites this paper.

DLLG: Dynamic Logit-Level Gating of LLM Experts Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:36:45.107785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-28T06:50:32.988192Z digest=sha256:2d9575d034094a4e854bf3d52c1563c6117a455fd3684606802188f21160a1b4

Observation 5a6e44bd-bbdd-4fe0-8f16-a3a86ab0e9b9 · inbound

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers cites this paper.

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T14:13:30.262342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T14:07:26.366172Z digest=sha256:4ac46f6e9b61283773e48dccd186b4e5493cd5f6757593ab663b6d0bfaa97f50

Observation bcf1baf2-c804-4f49-ae7b-3a4ed5bb9ae4 · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:09:36.573258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T16:15:22.543601Z digest=sha256:314f353bc75ee369c372681a0a0e2d113faa871d7f2300ef7af16485892cd700

Observation 3b1f1821-007f-48bc-98d3-e62ba70c0699 · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T10:48:59.226127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:48:59.226127Z digest=sha256:e1586098d75cac6ecb04180710c2041156e646c9969239b270cf2891a48f5186

Observation 9a08753f-9fbc-403d-92ee-fa7e0a6dfbba · inbound

Agent-as-a-Router: Agentic Model Routing for Coding Tasks cites this paper.

Agent-as-a-Router: Agentic Model Routing for Coding Tasks Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.010214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T08:40:26.716014Z digest=sha256:6cffe4b2b8eeb71afb2c6f53d95cb15604f3bb4d02a7547d44483bb647dfa332

Observation d780ba3d-d21c-4904-96ee-b89148a5809c · inbound

Agent-as-a-Router: Agentic Model Routing for Coding Tasks cites this paper.

Agent-as-a-Router: Agentic Model Routing for Coding Tasks Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:51.271659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T05:02:40.616741Z digest=sha256:a846012bbb6e106253210c5b0deead3b2960e5a97813a49ba058d5c9ded3dc85

Observation 09f1b869-8e22-44f9-a043-2fcf0c1f6633 · inbound

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries cites this paper.

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:55:35.898461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-01T06:50:19.566682Z digest=sha256:df006edb7ceb9c2fabd337a4a54697cb288767cfc20fb6d449cd43e2f84de568

Observation 1cede858-c685-4ad8-82c0-280dc5aa473f · inbound

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries cites this paper.

ComplianceGate: Classifier-Gated Multi-Tier LLM Routing for Inference in Regulated Industries Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:17:21.080725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-07-02T20:16:17.328340Z digest=sha256:db36d7488de9ace7562ff832561056323086cb4e4b033aaf59f8e83b77df58d7

Observation 0f57acb1-dc03-44b9-b456-ca01bfbf72fa · inbound

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks cites this paper.

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T18:37:16.526121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-02T18:19:43.146102Z digest=sha256:ea2d9485e09fc5b7018faf07a8a8994ea34d5407b20ad9d206689d832373421a

Observation d0ee6bea-867a-4ea6-b3cb-36eef49e6fd6 · inbound

A Workflow-Aware Serving Layer for Agentic Applications cites this paper.

A Workflow-Aware Serving Layer for Agentic Applications Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T05:56:18.797766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:56:18.797766Z digest=sha256:47ea936658eca9a0558718704850be3ad63e6bd4d9f3691f9d85e272c2807ee2

Observation 0abf5307-53e6-4f11-9896-255fbb611be2 · inbound

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning cites this paper.

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T05:38:27.357469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:38:27.357469Z digest=sha256:93f624fa79d06995cf98c177f0ea6214178119955daff9c352954f71dece3b7f

Observation d8f93125-a4b0-41b5-afd2-83c97395bbed · inbound

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning cites this paper.

TSRouter: Dynamic Modality-Model Selection for Time Series Reasoning Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T07:51:12.437343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:51:12.437343Z digest=sha256:4d54283d413194c1bc59f3519ede771b943bd56c6c7a25015c1961226a9c1490

Observation 6c35e317-8fe3-411e-a74a-e67c2c7d133e · inbound

HACO: Hedged Agent Computing for Reliable LLM Systems cites this paper.

HACO: Hedged Agent Computing for Reliable LLM Systems Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:57.574431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:57.574431Z digest=sha256:46590a40b06902b969e782e07c82aa8dc9c9c1f7e6e27670b2817b927b225321

Observation c7222c22-3a5e-4dda-9082-13cbf2705f88 · inbound

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference cites this paper.

PyroDash: Cost-Efficient Token-Level Small-Large Language Model Collaborative Inference Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T10:14:13.051550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:14:13.051550Z digest=sha256:229b6a98e911167cd1aa78c982dd7eaa436030b4f2393700e7bc043e2914362d

Observation 72386a27-6641-4b6c-ade4-b975beebfc64 · inbound

Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating cites this paper.

Routing Without Training: Controllable-Ratio LLM Offloading via Reliability Gating Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T12:43:45.819592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:43:45.819592Z digest=sha256:05e49d3545a935e169b698a2a5a6aadda8d5c94a3238b29bcb976ac55fa604fd

Observation 5835481e-a495-4a67-b0b1-29c78bc9202a · inbound

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs cites this paper.

LLM Serving in the Wild: An Empirical Study of Frameworks, Methods, and System Designs Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T14:55:56.408267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:55:56.408267Z digest=sha256:cf071f6385fd1457aebae475758ec35348d56242ef199b9dc805bd283ed0f4c8

Observation b1a11c41-d239-4cd2-9b00-07df7df0d872 · inbound

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First cites this paper.

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:21:11.021145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:21:11.021145Z digest=sha256:1522277e2f6e623d30bbf8f2623c71e8c0e97c2346d8e826e0e351b9188ef5bb

Observation d3bbfb66-1043-4fc6-8e86-d2d8313f11c0 · inbound

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents cites this paper.

Routing Is Least Learnable Where It Is Most Valuable: Bounds on Representation Routing for Web Agents Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T13:35:59.642970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:35:59.642970Z digest=sha256:833beaf9908dac69a4178ffebebbb65f5243c76c580369ebbad213a09a8104ac

Observation 0e68eff5-f8e7-4635-9969-f8f2c68ee57b · inbound

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing cites this paper.

Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing Hybrid LLM: Cost-Efficient and Quality-Aware Query Routing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:36:54.843239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:36:54.843239Z digest=sha256:91e5bad8d89bf0ed0b7ca39a5bbd10cd84049cec59d892bfc009a8317d0b8bae