Pith. sign in

Paper Citation Record · LEDGER

Distilling Reasoning Capabilities into Smaller Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2212.00193.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.00193 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T17:37:40.763610Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c7785893-e39f-48a0-bc0d-a4e16d8f8e84 · inbound

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation cites this paper.

Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation Distilling Reasoning Capabilities into Smaller Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-09T17:37:40.763610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T17:37:40.763610Z digest=sha256:e7f62c2b3eaf3198aaeb0be7bb9fbd1a8dcffb7ea83bffc6afb484355f61a270

Observation b45c7894-61c3-4b03-b761-5fce8ea192e9 · inbound

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification cites this paper.

Don't Just Demo, Teach Me the Principles: A Principle-Based Multi-Agent Prompting Strategy for Text Classification Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T13:39:04.476069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:39:04.476069Z digest=sha256:21aac22214f4ec61b64f1d7113823a0875d934e6ddde5a5d5afb8ad0a921657e

Observation 257dc47b-fe8a-44c9-a2f9-2b9ff886e6f3 · inbound

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability cites this paper.

Rationales Are Not Silver Bullets: Measuring the Impact of Rationales on Model Performance and Reliability Distilling Reasoning Capabilities into Smaller Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:24.689688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:40:24.689688Z digest=sha256:3728e7f718838888cc5361a8545980d0d4d049448a1cdddc1de84d2feab524f2

Observation f26e25c4-142e-4fb3-bd9d-8a960bff3f6b · inbound

Learning to Insert [PAUSE] Tokens for Better Reasoning cites this paper.

Learning to Insert [PAUSE] Tokens for Better Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:06:19.933207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:06:19.933207Z digest=sha256:ba1e90b0f1974ab3a3e7401f6585cdfd008cfcc6add35eee1eb8a5a6705e3333

Observation 5e535acb-da1f-4c43-9b0c-23df43c80594 · inbound

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning cites this paper.

Revisiting Test-Time Scaling: A Survey and a Diversity-Aware Method for Efficient Reasoning Distilling Reasoning Capabilities into Smaller Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T10:42:41.023501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:42:41.023501Z digest=sha256:a7dc2b71ebb08513db6554518efc9b1c46c37ad0db05c4be2cede67fce40475e

Observation f5735be7-56cc-43b7-8cfa-c2bad23c8ae5 · inbound

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery cites this paper.

SciGPT: A Large Language Model for Scientific Literature Understanding and Knowledge Discovery Distilling Reasoning Capabilities into Smaller Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T21:37:22.374350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:37:22.374350Z digest=sha256:812bfcabe2f29371a8947be1cdafe84d1a9e613aaad7347aa517f2faecbea8d1

Observation 6de40924-e738-4283-918e-1dad23ad4ab0 · inbound

Deep sequence models tend to memorize geometrically; it is unclear why cites this paper.

Deep sequence models tend to memorize geometrically; it is unclear why Distilling Reasoning Capabilities into Smaller Language Models

Reference 166

Resolution
verified exact
arxiv_id, observed 2026-05-21T20:40:36.235761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T20:38:18.005002Z digest=sha256:c9c807c11d951a87a301b21fa1324ae53eb370c77d4e1c7cd4dda756e86c1a16

Observation 7d5052cc-e4bf-4c50-bf7b-4a2c0948d171 · inbound

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces cites this paper.

Characterize Then Distill: Mechanistic Reasoning in Large Output Spaces Distilling Reasoning Capabilities into Smaller Language Models

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-06-27T22:31:21.583158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T22:22:52.690010Z digest=sha256:cce55daaad5b80bb993ab06d199be09624946874242f40aaca93bd82e527761f

Observation 5d0924f7-af5f-4a83-9f8b-29861e27fa5b · inbound

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers cites this paper.

Contribution Weights: A Geometrical Analysis of Self-Attention Transformers Distilling Reasoning Capabilities into Smaller Language Models

Reference 109

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:32:46.691403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T23:29:02.457697Z digest=sha256:b2c3e241186bd3563fae03eb86883cd395d0093e7b025c1cf3b390b22f6d4a6f

Observation 191a9b54-8171-4add-b8d7-7cf65166ebc5 · inbound

Toward Calibrated, Fair, and accurate Deepfake Detection cites this paper.

Toward Calibrated, Fair, and accurate Deepfake Detection Distilling Reasoning Capabilities into Smaller Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-06-28T07:11:45.303811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T07:05:18.026601Z digest=sha256:34296fd65f434f01f4c870baf30cc542c9b14819cad10fd538890d2a85daea36