Pith. sign in

Paper Citation Record · LEDGER

TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2408.12320.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.12320 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:09:39.632024Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T18:37:16.483139Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3a97c8e0-0473-4082-9ea9-96e08e198a90 · inbound

Bench-CoE: a Framework for Collaboration of Experts from Benchmark cites this paper.

Bench-CoE: a Framework for Collaboration of Experts from Benchmark TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T21:46:44.846975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:46:44.846975Z digest=sha256:dec2efb58a7e5a13f227d27fdac7b2a7c9cb1d28f7beb48ed4719e3933d91501

Observation ed4bc80d-a439-4b04-b0fc-dfde91b9f58f · inbound

Rerouting LLM Routers cites this paper.

Rerouting LLM Routers TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-10T22:24:45.610732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:24:45.610732Z digest=sha256:14300f2f0e3ff05571a57cac5654a7d287517971e464c96f3c71be89cd246c28

Observation 6bd6de6c-f446-4a20-876a-8e2dfb717e85 · inbound

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization cites this paper.

Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T22:49:51.798116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T22:49:51.798116Z digest=sha256:acefb0d7093459c2ecc749d45fb66cfee3aa5e5fe4f394b702521ce5a7510529

Observation c1ccbc90-a3eb-48bb-8206-e915844f58f4 · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.596740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.596740Z digest=sha256:e07afd8f20284832d1327159ca66e29d800de05ae2f0888e1445d7c014e286a1

Observation 8595c1b8-42d5-45b0-8531-555d397144ad · inbound

Cognify: Supercharging Gen-AI Workflows With Hierarchical Autotuning cites this paper.

Cognify: Supercharging Gen-AI Workflows With Hierarchical Autotuning TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T11:03:44.786357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:03:44.786357Z digest=sha256:a69791f1d7d2ef3ae56dad0c5e052ff7bbc27abe2179ad8bae54d677ad6630c1

Observation de3199c5-50de-4121-af18-8ded23b122ee · inbound

Universal Model Routing for Efficient LLM Inference cites this paper.

Universal Model Routing for Efficient LLM Inference TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T23:48:01.219001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:48:01.219001Z digest=sha256:941b50d47cb43feaddcf9bb95025e0852ea0047773f59be71004f0657369b96b

Observation b8195b0e-08f8-4ad3-9df8-93e946e2a568 · inbound

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble cites this paper.

Harnessing Multiple Large Language Models: A Survey on LLM Ensemble TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-23T02:25:19.639305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-23T02:22:28.649071Z digest=sha256:21c51302f358f7f93874c6cae6f55dca8af7f1a3d4cf11d02032249a091d56f1

Observation e3a61287-baf8-494d-9587-2858a4f224cd · inbound

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers cites this paper.

Rethinking Predictive Modeling for LLM Routing: When Simple kNN Beats Complex Learned Routers TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-22T15:14:57.437352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T15:13:28.927880Z digest=sha256:dd0fbdcc34c645df470751abc678072f62a42635b5a6f7ed0cc3e3bf04433bdc

Observation 2cc3dc16-44ef-4180-b152-3a1d93d18581 · inbound

CoE-Ops: Collaboration of LLM-based Experts for AIOps Question-Answering cites this paper.

CoE-Ops: Collaboration of LLM-based Experts for AIOps Question-Answering TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:09:39.632024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:09:39.632024Z digest=sha256:835eb29dae78494c72b3148a48e960306857c9b55d83eb5734676ff551b2c8e8

Observation d6fb5336-512d-48cd-8483-794d817d97f1 · inbound

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts cites this paper.

Quality-of-Service Aware LLM Routing for Edge Computing with Multiple Experts TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:22:42.484238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:22:42.484238Z digest=sha256:3ea448e1b5a49c9de48008e7ff0171ad187ed2de2daf97c5afb03dc0537fd96a

Observation 22942fe5-9c9a-4221-aacb-4b0c593af27c · inbound

One Head, Many Models: Cross-Attention Routing for Cost-Aware LLM Selection cites this paper.

One Head, Many Models: Cross-Attention Routing for Cost-Aware LLM Selection TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:09:31.562664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:09:31.562664Z digest=sha256:197209a62ce038c69d3fe9aa11a3be5172b29cd249a0ec662508473565b8cc14

Observation b6f7a51c-a0f0-4a2f-8ea9-f791d7c16382 · inbound

R2-Router: A New Paradigm for LLM Routing with Reasoning cites this paper.

R2-Router: A New Paradigm for LLM Routing with Reasoning TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T05:17:30.681496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:17:30.681496Z digest=sha256:e9cffae00f7946f6e12f6ef485fc2f10b59547b7f86f881a2968df71d33fbc42

Observation be57fbcd-1cff-44b5-b905-da0c9539a694 · inbound

Switchcraft: AI Model Router for Agentic Tool Calling cites this paper.

Switchcraft: AI Model Router for Agentic Tool Calling TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:50:56.562284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T01:02:55.275359Z digest=sha256:aae05ad0abbe9677a06fed066adda2f854ce3542314170036482f6aca65b2f0c

Observation 80108400-a299-49f9-9632-227e2828d298 · inbound

IR3DE: A Linear Router for Large Language Models cites this paper.

IR3DE: A Linear Router for Large Language Models TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:26:59.058078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:19:30.946221Z digest=sha256:33a6488d94b83550c26afd45cf15e76cf9bf6525f2aa1daf23c779728b445e3b

Observation 8129c5dd-b77f-4ebc-b55a-168d342940a8 · inbound

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval cites this paper.

CAMI: Cost-Aware Agent-Guided Multi-Indexing for Semantic Retrieval TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:24:38.783544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T11:07:31.861143Z digest=sha256:7fa742859e7eba445ac5e5130d00502e28028423c9717a2f7cb84511d9dc1e35

Observation 271ab246-196d-471c-9c93-9bf0357e303a · inbound

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks cites this paper.

SWE-Router: Routing in Multi-turn Agentic Software Engineering Tasks TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-02T18:37:16.484385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-02T18:19:43.146102Z digest=sha256:6f09191c2132bd5d9043a506ecde959456badeafb9b4b5c5611a3fbcb6cf80b6

Observation 22ee0efc-a464-4ee7-a2ea-698ede494ce3 · inbound

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes cites this paper.

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes TensorOpera Router: A Multi-Model Router for Efficient LLM Inference

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-04T07:49:39.967543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:49:39.967543Z digest=sha256:77b529e4b7a6d20197c3ce51e1a49d1f39cc13b4213ac6f6682fe29267191fbf