Pith. sign in

Paper Citation Record · LEDGER

Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2310.03094.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.03094 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T22:33:09.714455Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:09:44.136682Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 19bbf7df-5b0e-49c9-833e-5c1a8658f84b · inbound

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees cites this paper.

Optimal Query Allocation in Extractive QA with LLMs: A Learning-to-Defer Framework with Theoretical Guarantees Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:53:21.433797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-23T18:49:39.718108Z digest=sha256:804c1f0ae355a9b11b8052cc5e346433eefe2181b559e47b66e1e0dd591f2f8e

Observation 7c833d01-c69a-4345-87d9-481d769f9399 · inbound

When One LLM Drools, Multi-LLM Collaboration Rules cites this paper.

When One LLM Drools, Multi-LLM Collaboration Rules Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 129

Resolution
unresolved
no resolver link, observed 2026-08-08T22:33:09.714455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T22:33:09.714455Z digest=sha256:359b74be0f8d9cfa97646dcda5aadcc8c6b276dafd26ea5aaeb55105b66a77ca

Observation eae85b95-1cc1-44ad-964a-a3da894f18e1 · inbound

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length cites this paper.

AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T04:30:43.727123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:30:43.727123Z digest=sha256:f13732f0ab45b002c0a8cba9ce2ce243e88af597763955b4daa4400f41af7d54

Observation e511c90b-13a8-4ed9-984b-4cff45cb6ba5 · inbound

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute cites this paper.

BEST-Route: Adaptive LLM Routing with Test-Time Optimal Compute Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T22:05:29.245616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:05:29.245616Z digest=sha256:40a3eea793a1facb1e9ecf5e070ac687fc4b1a883d4d2be086decfdc412faeb4

Observation cb82fa50-a591-4d89-b43f-a0422c9afd68 · inbound

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation cites this paper.

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:53:14.009570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T22:52:58.524934Z digest=sha256:a018ffbef9026f01585a1e8ba720032af935d57b6c3dd8b0c7703d349a036d89

Observation b8eb085c-a640-47c5-aae6-f94e5d34b11b · inbound

Evaluating Small Language Models for Front-Door Routing: A Harmonized Benchmark and Synthetic-Traffic Experiment cites this paper.

Evaluating Small Language Models for Front-Door Routing: A Harmonized Benchmark and Synthetic-Traffic Experiment Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:23.073475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:38:17.474611Z digest=sha256:eb42a72f347ac77b31dcbc563aab4d46f3d13bcab377edab4d80976fda9f4fc7

Observation 9e2a85f5-e476-4d01-be07-0cd8f57165a7 · inbound

Privacy-Preserving LLMs Routing cites this paper.

Privacy-Preserving LLMs Routing Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:32:51.625602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:32:43.015370Z digest=sha256:ee45bd22f527409479b0cd69da0bda983e3127ff09a3f6915452795150b19e7e

Observation fedde61c-d6d3-490d-bcfa-a014536387a7 · inbound

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models cites this paper.

Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T14:25:29.196389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T14:24:55.735201Z digest=sha256:fb01a7640ecf4fe45fc660cd776779ebfd97fd3547f582607ed453c8688d6220

Observation 0ec3e192-c478-4372-8ad1-d81ee46e003d · inbound

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models cites this paper.

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:17.421068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T05:11:46.991452Z digest=sha256:0c0f2b8df03459f0add703088ddba2f956269d78c49ad36038ab7df216665a39

Observation 6e7fee11-b8dd-4822-8840-2a525bf18d90 · inbound

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models cites this paper.

SEMA-SQL: Beyond Traditional Relational Querying with Large Language Models Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:55:10.577413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T06:53:11.604043Z digest=sha256:17e91c85741f35f1c74155402013d3bc9f6b3e29976588ad2d18c7ded8b141ba

Observation 6b2207cc-b58c-4b5b-83e5-b26aefca08f3 · inbound

Online Learning-to-Defer with Varying Experts cites this paper.

Online Learning-to-Defer with Varying Experts Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 96

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T04:07:13.156453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T04:05:52.853863Z digest=sha256:bedccaf192cda9ea6d88811b5a1ac3dc2275eb1675b9d9d1a78acb60b2e870b8

Observation f2a63b40-57a6-4ff6-a951-03a2841e180d · inbound

Online Learning-to-Defer with Varying Experts cites this paper.

Online Learning-to-Defer with Varying Experts Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 96

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T08:09:51.328157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T08:07:35.959083Z digest=sha256:aef0e13783b35c34c7ba6dd789b83e194d034f0f630b30a139fef0696ac72887

Observation b78d259c-7345-4ac3-b67c-ed8d22174803 · inbound

From Sampled Outcomes to Capability Distributions: Rethinking Supervision for LLM Routing cites this paper.

From Sampled Outcomes to Capability Distributions: Rethinking Supervision for LLM Routing Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 151

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:17:08.804014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T22:54:28.452796Z digest=sha256:98418c26d77a04674bb7fe267a5c4ff9f17bcfc6cb965eaca2e9cba5a39507ea

Observation 1c332ccd-0c19-4907-aa8b-53723d7de47d · inbound

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing cites this paper.

RouteJudge: An Open Platform for Reproducible and Preference-Aware LLM Routing Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:19:13.488572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T21:17:29.543901Z digest=sha256:a58d75eeb6cd9ccbeb654cebf4251795c084bc3603f0ae1f73d2157bef6e0486

Observation 5dfd33e0-ec31-465d-bd4c-cccfddcbbd51 · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:09:36.813433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T16:15:22.543601Z digest=sha256:c94e16fc5cb257c3a0e5d16b3f6be484f64f4b55b285167838db3972cc3dc2c8

Observation 79bf9a32-a2f2-425a-9005-ddfeea326fb8 · inbound

Token-Operations-Oriented Inference Optimization Techniques for Large Models cites this paper.

Token-Operations-Oriented Inference Optimization Techniques for Large Models Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T10:49:00.720574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:49:00.720574Z digest=sha256:87ad92753865efca25a47d519168080b942f701a4b4ced8d5445adfd00651216

Observation d116d5c0-8419-4e0c-ac75-4b9ccc5ee714 · inbound

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving cites this paper.

RLM-Cascade: Response-Level Speculative Decoding for Cost-Efficient LLM API Serving Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:09:44.138484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T09:10:47.370434Z digest=sha256:c922c75473400f1531ad47d0de5b65e95f3c831f8cc0ab27296c6809311c17a8

Observation 7ef6c68d-2314-4cae-ab4f-1cefe6c15f45 · inbound

HACO: Hedged Agent Computing for Reliable LLM Systems cites this paper.

HACO: Hedged Agent Computing for Reliable LLM Systems Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T13:11:02.437028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:11:02.437028Z digest=sha256:8805eda1bdcc10a4aac48377eb4183b017db7c11be0788975cd6852972205127

Observation 09860c77-6b1f-4a43-bc81-491d1e7be16f · inbound

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference cites this paper.

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-31T03:38:02.101711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T03:38:02.101711Z digest=sha256:e51586ea4018a561ea9bea21cf6db09a0ce89a566b83d10953b0e5487e1f8983

Observation 7631a5a5-0a6e-492a-8c90-ce0746b4dee1 · inbound

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference cites this paper.

Conformal Cascade: Distribution-Free Accuracy Guarantees for Multi-Tier LLM Inference Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T01:49:19.070875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:49:19.070875Z digest=sha256:0d20ee7df45f93217756721bd29a467bc13dae2f5563e4ecda9855b7629edd2c

Observation a1ab2906-8665-43d2-8239-46f0fdb9c1d1 · inbound

How Often Should a Recommender Call an LLM? Value-Weighted Routing, Monitoring, and Seasonal Robustness cites this paper.

How Often Should a Recommender Call an LLM? Value-Weighted Routing, Monitoring, and Seasonal Robustness Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-31T02:16:44.828313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T02:16:44.828313Z digest=sha256:eb4e5d4ae2046ae1d1ca38fe1a6a2b4647b5254f926342d44f0c7556dae1d57c

Observation 7aa18767-100d-4500-9cd4-eaa65344a713 · inbound

Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents cites this paper.

Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-30T18:50:46.197482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T18:50:46.197482Z digest=sha256:c40f20e63ed528e9083628fc03152d6166fe98e5506e9e8eb684259e7d2560be

Observation c2151bb0-53f0-4594-a0fd-67999d38183f · inbound

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes cites this paper.

Training-Free versus Training-Based Intent Classification in LLMs: Accuracy, Robustness, and Failure Modes Large Language Model Cascades with Mixture of Thoughts Representations for Cost-efficient Reasoning

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-04T07:49:39.987618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:49:39.987618Z digest=sha256:b86cf31dff9f3db19d6abaad24b249865941e686a2ef2d10d92e439ad0f075c8