Pith. sign in

Paper Citation Record · LEDGER

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability

As of 17 August 2026, this Paper Citation Record lists 90 of 90 outbound references and 0 inbound Pith citation observations for arXiv:2504.16056.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16056 v1

Coverage vector

measured 90 of 90 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:14:33.817036Z

measured 90 of 90 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

90 of 90 outbound references displayed

  • verified exact12
  • verified fuzzy36
  • unresolved41
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 98b69021-97a2-44e2-9763-6abb1118bb58 · outbound

This paper cites Language models are few-shot learners,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Language models are few-shot learners,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.210698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.210698Z digest=sha256:7d0a81a3bdfd506f532d3806262a857aebd5d64fdad992cf8da4144c514b35e7

Observation b1812db4-a0bc-4667-ad44-fd6c7278b68a · outbound

This paper cites Language models are unsupervised multitask learners,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Language models are unsupervised multitask learners,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.219284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.219284Z digest=sha256:5319fd6d9ff840f5f741ef7f0e06bce09386057da1b68f9f14e81cd095ce42ae

Observation f3966a9c-5b91-4152-b30a-b478383c6509 · outbound

This paper cites How many data points is a prompt worth?.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability How many data points is a prompt worth?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.224801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.224801Z digest=sha256:1d6c47900ac81272a0dc47975b1eeb50e05320cb4f7520f6f58e925777cc0605

Observation 2f3a4864-4e04-447b-a013-b6613aa6a73b · outbound

This paper cites A Good Prompt Is Worth Millions of Parameters: Low-resource Prompt-based Learning for Vision-Language Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A Good Prompt Is Worth Millions of Parameters: Low-resource Prompt-based Learning for Vision-Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.230226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.230226Z digest=sha256:dcac9c4fcbd57d5af6fb97a907bd8615d4a942625b1a8ef91aa1b0472da5ea37

Observation 18db569c-24c1-45ec-bd85-31d81a8856fe · outbound

This paper cites Orca 2: Teaching small language models how to reason,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Orca 2: Teaching small language models how to reason,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.237630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.237630Z digest=sha256:2960f45e9873cb5db54350544d56e4d85f49dd3c355aac018be3324c72edaad9

Observation 88e77e2b-a18d-43a1-a0fc-3035525f1d79 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Chain-of-thought prompting elicits reasoning in large language models,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.243258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.243258Z digest=sha256:ad4e3d39cacd62348a1826c10961d1879e1b70af82e00e3e4f4676fda12359ba

Observation d769a000-8baa-4909-add5-02b14e90f632 · outbound

This paper cites Tree of thoughts: deliberate problem solving with large language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Tree of thoughts: deliberate problem solving with large language models,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.253840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.253840Z digest=sha256:163b669f10cbf512be98e179fce320c397701003b35c84dfa8f38b25c323f7bb

Observation 40c8ec38-e37d-48a2-884c-197e34eba25c · outbound

This paper cites Evaluating Consistency and Reasoning Capabilities of Large Language Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Evaluating Consistency and Reasoning Capabilities of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.262619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.262619Z digest=sha256:48910c5d931dbc020503462a5c6a917a90d85f9fefc68bb64a61c95a2e165061

Observation f3757aab-b2ac-4f54-8036-bfe20918ac78 · outbound

This paper cites LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.271377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.271377Z digest=sha256:ef1d3a1424201cc713365b6f37c7bc6f900e3fcfef87fc9c229e38bc14e4ddee

Observation e19bff66-49ed-4f99-af76-1264f20f4c32 · outbound

This paper cites Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Towards a Mechanistic Interpretation of Multi-Step Reasoning Capabilities of Language Models,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.278489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.278489Z digest=sha256:f505e2fb25cac6a29b2dbf139e0cbfcba97e0635f7e8e10b7b12224f980a9371

Observation 36bb8b42-14b4-4dbb-9604-457fedc2b329 · outbound

This paper cites Large Language Models: A Survey.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Large Language Models: A Survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.286742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.286742Z digest=sha256:a0d38448e4c955e2c7b94392d91b7d56d3d88da53d22b559c985173bf16ccb77

Observation d711e744-6740-4214-9a2a-d4fa81a14483 · outbound

This paper cites PaLM: Scaling Language Modeling with Pathways.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability PaLM: Scaling Language Modeling with Pathways

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.295471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.295471Z digest=sha256:99bd04d55210b12e853e99d7218e7b0e4cef1686c38466a7989d8488e77c246d

Observation 5f8d12ff-88b8-4e3b-ada2-ec0858bbe784 · outbound

This paper cites Distilling Task-Specific Knowledge from BERT into Simple Neural Networks.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Distilling Task-Specific Knowledge from BERT into Simple Neural Networks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.305643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.305643Z digest=sha256:4682649c940f4d416668dfd47bf836a37851bce8d9b7734292c3d51f711d6a9d

Observation 0c1b7307-c6b4-46b2-b806-67ef6513496a · outbound

This paper cites TinyBERT: Distilling BERT for Natural Language Understanding,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability TinyBERT: Distilling BERT for Natural Language Understanding,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.312409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.312409Z digest=sha256:e827951a6d5ca3ed72e18b47eecc7cad0ad56838dcb24ff48da639f04f19f0bd

Observation ae1240ab-bf61-4e4b-8bd5-08dac906f5d8 · outbound

This paper cites A survey on model compression for large language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A survey on model compression for large language models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.318425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.318425Z digest=sha256:4f0e07425a545db44834e3c5d8c72b06daec8e81fe835e6e62b012182d05e4af

Observation fc464acc-3672-4d20-8166-1b425735dad7 · outbound

This paper cites Robustness-Reinforced Knowledge Distillation With Correlation Distance and Network Pruning,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Robustness-Reinforced Knowledge Distillation With Correlation Distance and Network Pruning,

Reference 16

Resolution
verified exact
raw_fallback, observed 2026-08-16T11:14:35.465463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.330242Z digest=sha256:6a431709da2f52edbebfefe9aa8bc55fcaf347e887f9a896bf1418836f9aab85

Observation 205069e1-d13f-4f4a-bf81-09b72aac1138 · outbound

This paper cites Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Distilling Step-by-Step! Outperforming Larger Language Models with Less Training Data and Smaller Model Sizes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.335551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.335551Z digest=sha256:a70d51ccbbdee97148a74b9e457c16eac55b6acae67c714abdea099a3c2304eb

Observation dbdf1f77-978b-49a2-a739-23cf6e579ac4 · outbound

This paper cites PanDa: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability PanDa: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation,

Reference 18

Resolution
verified exact
raw_fallback, observed 2026-08-16T11:14:35.326248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.343242Z digest=sha256:6529398537ddc959932c95db75a4b4b4343b468a249905d88ef99f35a8c7cceb

Observation 7be016a7-3828-4fea-b47e-b9df5de62d82 · outbound

This paper cites Efficient Knowledge Distillation: Empowering Small Language Models with Teacher Model Insights.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Efficient Knowledge Distillation: Empowering Small Language Models with Teacher Model Insights

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:35.235959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.349173Z digest=sha256:ea3802822f6b86ed849b354c1ac9e5c3ecb288b2053131690560c28c905550f5

Observation 7f2bdad9-662d-4356-9eae-b177258ab2b4 · outbound

This paper cites Improve Student‘s Reasoning Generalizability through Cascading Decomposed CoTs Distillation,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Improve Student‘s Reasoning Generalizability through Cascading Decomposed CoTs Distillation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.701126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.355639Z digest=sha256:4e3171a68e073a4eb10449de4e17f067a0eda5e416f603fb0e27db4c7617a745

Observation 6f514958-2a04-47f7-8def-6d790ed36f17 · outbound

This paper cites Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.362273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.362273Z digest=sha256:24464323f97475ef2cd03e7a0d2f057d817a3cdc700875310c9d582e154521ae

Observation 68424ea8-48fd-4ce2-8154-e08f800a1ae8 · outbound

This paper cites One teacher is enough? pre-trained language model distillation from multiple teachers,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability One teacher is enough? pre-trained language model distillation from multiple teachers,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.682170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.376497Z digest=sha256:5403743bbcd2fec783e55e92872028d317b4182db5b25bb05968fa47bb9c98ad

Observation 32b23400-321c-4d2b-8556-6adc88e07744 · outbound

This paper cites Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Beyond Imitation: Learning Key Reasoning Steps from Dual Chain-of-Thoughts in Reasoning Distillation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.370369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.370369Z digest=sha256:5978fc195d4b6e9c9355d19db9828eb5ff4a1737ed25681f420c2689483114b8

Observation 9d81e9cb-a2f6-4b01-a567-0d722665be01 · outbound

This paper cites We’re getting a better idea of AI’s true carbon footprint,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability We’re getting a better idea of AI’s true carbon footprint,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.664620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.390193Z digest=sha256:201b00fbc75e5b9c5baf54030f490a301d1108a4fcfb24a2dbfe1ace6c38d71b

Observation 2baef9c0-9ad5-4038-971c-bc99aadde51f · outbound

This paper cites Distilling Large Vision-Language Model with Out-of-Distribution Generalizability ,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Distilling Large Vision-Language Model with Out-of-Distribution Generalizability ,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.383447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.383447Z digest=sha256:f1c21a9e4f4787308ea90bcaabc672fd97c7307d0543912eb52c15eab337509d

Observation a5d1de52-e38c-4275-b0d1-b30bed9c1900 · outbound

This paper cites How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability How Far Can Camels Go? Exploring the State of Instruction Tuning on Open Resources

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.401221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.401221Z digest=sha256:c373e76b72b6da918028ae915cfc967412f9398876132fc91dfafea8cc636292

Observation 7b777eb8-5b55-4c68-be52-747a82cf7873 · outbound

This paper cites Energy and policy considerations for modern deep learning research,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Energy and policy considerations for modern deep learning research,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.642068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.395348Z digest=sha256:631608a9c7ed676f7deb4218387766bce5f7f2571d427833d7d1b9ce38bad78f

Observation ef9a79be-8a3d-457c-90f4-63ee2fefb9ad · outbound

This paper cites Towards a rigorous science of interpretable machine learning,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Towards a rigorous science of interpretable machine learning,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.623378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.415619Z digest=sha256:24d098b3c70467448f7d01d87b1f5f4370b28c1cf2e9b5f8285096019b8a1c0a

Observation 2f72d52e-be27-4dd7-9f02-6c71301c8668 · outbound

This paper cites Code Llama: Open Foundation Models for Code.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Code Llama: Open Foundation Models for Code

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.409305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.409305Z digest=sha256:d8f885396c33faadbe7b12aa99ac1cd0822be71841bc08cca7f25200388e0130

Observation a0859352-8950-41cf-87d8-d7c3542e2479 · outbound

This paper cites Human Delegation Behavior in Human-AI Collaboration: The Effect of Contextual Information.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Human Delegation Behavior in Human-AI Collaboration: The Effect of Contextual Information

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.945203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.429650Z digest=sha256:8daeb83c03ff7f570cb2dbcc01e04c640377565b4ffd9f014cfc60b40bcddeae

Observation f388cb98-012c-41db-b296-e006bdcb0588 · outbound

This paper cites Explainability in AI Based Applications: A Framework for Comparing Different Techniques.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Explainability in AI Based Applications: A Framework for Comparing Different Techniques

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.987158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.422214Z digest=sha256:381428332574e9f3ca7580167997842e7dfa7a96d751990d1982439c44eae259

Observation 68e86581-1138-47e4-bf33-abba4f572450 · outbound

This paper cites Multimodal explanations: Justifying decisions and pointing to the evidence,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Multimodal explanations: Justifying decisions and pointing to the evidence,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.582192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.444498Z digest=sha256:7ec8ea3a1a8540a812d7f25f2576b1c7f23c280504a042743aa45ac81a3bb601

Observation 67acd858-bb00-4913-8335-0ff0bf6532d0 · outbound

This paper cites Textual explanations for self-driving vehicles,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Textual explanations for self-driving vehicles,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.603412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.436517Z digest=sha256:efce892be32e95e9ef7c99f667590b370f5567a2e89f4353da421a4bbb5d3c52

Observation 24a31e83-23f5-48b4-8c34-884a9706401d · outbound

This paper cites SCOTT: Self-consistent chain-of-thought distillation,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability SCOTT: Self-consistent chain-of-thought distillation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.557491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.457164Z digest=sha256:4d6259f7fa06fa70267cfea6a1c76d931e26bfb9e13e0f82e3ed7c1f44abbdb8

Observation 11020271-44dd-4c07-9a27-f2121b09e107 · outbound

This paper cites The Impact of Imperfect XAI on Human-AI Decision-Making.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability The Impact of Imperfect XAI on Human-AI Decision-Making

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.911322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.451101Z digest=sha256:503e856d517620766f3c068229cb9e77a8642e2adbe05d2d00606574d82f2d63

Observation e380decd-ef05-4d9c-bd54-89933584f5c1 · outbound

This paper cites GPT-NeoX-20B: An open-source autoregressive language model,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability GPT-NeoX-20B: An open-source autoregressive language model,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.534008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.470853Z digest=sha256:9eb0bbe83ec5bbad96389aa4ef19509eea72310de48a0f442d4d2ef5d3b8ef99

Observation d03cc4b7-2763-456b-9b92-07728ef5de20 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Constitutional AI: Harmlessness from AI Feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.464536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.464536Z digest=sha256:be1b4a3600bc1a893b789fc0bf5ad7e9829875ff84827364d01a1ffa7df47dca

Observation 8c5eeee1-4791-4d24-8e43-1c2e99c5ba13 · outbound

This paper cites MiniLLM: Knowledge distillation of large language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability MiniLLM: Knowledge distillation of large language models,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.514086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.484866Z digest=sha256:a9c78297e83036f3ad5b46c95089e92d988eff35f92c6b0e3fde168473c7968e

Observation 1dfe0a89-dace-4aa9-9165-a5fb0cfba560 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Distilling the Knowledge in a Neural Network

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.478311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.478311Z digest=sha256:d060a07bbfaa42cab8f87870db7bcaddb718f2d8f5efd8943faa64de11596d61

Observation ed1622c9-09f0-4769-b435-72e85aa870fc · outbound

This paper cites Collaborative Distillation for Ultra-Resolution Universal Style Transfer.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Collaborative Distillation for Ultra-Resolution Universal Style Transfer

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.794436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.499036Z digest=sha256:b76e9ba7fd479fa766e418ea4faf6dbe667a140cd844952d30a7beac4361e069

Observation e77b374e-b894-4c69-8e36-4e804050e515 · outbound

This paper cites ERDL: Efficient Retrieval Framework Based on Distillation from Large Language Models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability ERDL: Efficient Retrieval Framework Based on Distillation from Large Language Models,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.493565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.491419Z digest=sha256:2eef501ea1e48e8c8dc305a677fb142bea811393081ea557a2099a5025b5ce10

Observation 4df334a1-f535-470c-8972-dd470626dbe9 · outbound

This paper cites Learning Efficient Object Detection Models with Knowledge Distillation,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Learning Efficient Object Detection Models with Knowledge Distillation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.473530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.512642Z digest=sha256:46b81591c09dc8456bf0a55f61d69a96aa551dceab57da6500e7dec357ff9f43

Observation 7258449e-ce62-4ddb-b21a-61a3e922586d · outbound

This paper cites Triplet Distillation for Deep Face Recognition.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Triplet Distillation for Deep Face Recognition

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.760921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.505671Z digest=sha256:9b43f8c4282de275adab964e0f814e645ce1b96ea27359c1dfe952e99a946073

Observation ff7ee1f4-7ecc-407f-a6f4-453edda60d00 · outbound

This paper cites Training Compact Models for Low Resource Entity Tagging using Pre-trained Language Models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Training Compact Models for Low Resource Entity Tagging using Pre-trained Language Models,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.455034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.526719Z digest=sha256:1378c03ca8d3f0d81da47539dbc1fdca6c3fb295eae6710cacf9de44b95c8dfa

Observation 9b1ab665-5a3e-4740-aa82-e97f7594d065 · outbound

This paper cites A Survey on Knowledge Distillation of Large Language Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A Survey on Knowledge Distillation of Large Language Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.519953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.519953Z digest=sha256:55678f74d28e6935a6a47f31cc18dbbfb1c74c887b631daccf6001416e4e369b

Observation fae3d5ec-891f-44c9-8386-4323c85d1ab5 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Lora: Low-rank adaptation of large language models,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.541136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.541136Z digest=sha256:1ec6a8b25562dcb3597e383502e3c9767142ac3d1079f09aea7dfef2109f5b5c

Observation 076b0ddf-f35d-4b9b-abca-75a1a6637a35 · outbound

This paper cites On the effectiveness of adapter-based tuning for pretrained language model adaptation,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability On the effectiveness of adapter-based tuning for pretrained language model adaptation,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.533713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.533713Z digest=sha256:6dcc4b93a4a17be60a715fa38005f20efa5389bf001efc7e04e3f7bbf7dde728

Observation 5a332118-3ebb-4b38-aba3-8ad410930bb6 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability The Power of Scale for Parameter-Efficient Prompt Tuning,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.422585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.563083Z digest=sha256:bd1ddb847ef0a4ea72483bfbfbb0c1d3a70bbb8d4621442b66d88ca13d766132

Observation 829b6fe7-5bef-4449-b3d6-20d866f33253 · outbound

This paper cites Efficiency Optimization of Large-Scale Language Models Based on Deep Learning in Natural Language Processing Tasks,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Efficiency Optimization of Large-Scale Language Models Based on Deep Learning in Natural Language Processing Tasks,

Reference 49

Resolution
verified exact
raw_fallback, observed 2026-08-16T11:14:34.636249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.570311Z digest=sha256:df2d31ae48ad7ff5657f02e445dfda4e8b433b6320dce952e37ce131cdca3cd8

Observation 2726a0e4-0f88-476d-aebf-8984ffc4d260 · outbound

This paper cites LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability LoPT: Low-Rank Prompt Tuning for Parameter Efficient Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.555181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.555181Z digest=sha256:2578fdab35e943f875c0a0452d3f92956c2fdd0e342ebc0538219a659c1d8024

Observation 23af0cdf-774c-438c-8714-91bc58b841d1 · outbound

This paper cites A survey on the explainability of su- pervised machine learning,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A survey on the explainability of su- pervised machine learning,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.583008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.583008Z digest=sha256:18242517fb959ccb229a3d5f46fc550d93029cc7b8def0a546e57251e961487a

Observation e0b18b51-2d2f-4dc3-b554-1c944bd67083 · outbound

This paper cites Baize: An Open- Source Chat Model with Parameter-Efficient Tuning on Self-Chat Data,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Baize: An Open- Source Chat Model with Parameter-Efficient Tuning on Self-Chat Data,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.356019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.588971Z digest=sha256:3bcc2db0b027a545e93ea36b8119cd77f60f944412d2adc3b14e5c7dc53e7eb5

Observation 128b23c6-9508-4834-8853-a7a723bfff32 · outbound

This paper cites Causability and explainability of artificial intelligence in medicine,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Causability and explainability of artificial intelligence in medicine,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.398035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.576472Z digest=sha256:04c233c670fb17bbef9aadeb33903e6d952d4e47ad15be740666c4926d26c367

Observation 76cbda5e-7e7c-45f9-9ae2-008253d7765c · outbound

This paper cites Large Language Models Are Reasoning Teachers.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Large Language Models Are Reasoning Teachers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.602378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.602378Z digest=sha256:2c0923dc35a9562b936d9b965fecb8b52375bf32dc2d2ff24b6365839f55244e

Observation 457fff00-8fc6-4855-97f3-8a2b0e7d16e3 · outbound

This paper cites Specializing Smaller Language Models towards Multi-Step Reasoning.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Specializing Smaller Language Models towards Multi-Step Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.611181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.611181Z digest=sha256:d3d7ccb970a780be213b1a541e4033cb4781ad0b4c7a8c4c2f980de3abef2652

Observation a1e58e14-d50d-4ebb-9a7d-0f3a7962c376 · outbound

This paper cites Teaching small language models to reason,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Teaching small language models to reason,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.334869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.595210Z digest=sha256:c71cd84d5affa43da6942cb91832956c952521d2511e7e408d8324d37ba8e37a

Observation 8e6492d4-5e75-457b-a33e-4f4584da72e2 · outbound

This paper cites Orca: Progressive Learning from Complex Explanation Traces of GPT-4.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Orca: Progressive Learning from Complex Explanation Traces of GPT-4

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.623608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.623608Z digest=sha256:503d7e61208f122ef8065abdb9eea3f434d17652bc25844a89fe7c60106a1172

Observation 13966246-99c7-4232-ac64-bfcc5916e6f7 · outbound

This paper cites Explain Yourself! Leveraging Language Models for Commonsense Reasoning,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Explain Yourself! Leveraging Language Models for Commonsense Reasoning,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.317871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.630185Z digest=sha256:3f33967eaef9c4e1c5c32cbc1d49140f6d405eef4f460c5e54214b5cc4e589c7

Observation 0556d3fa-430b-4d1d-ba92-7e91b014fb34 · outbound

This paper cites Sci-CoT: Leveraging Large Language Models for Enhanced Knowledge Distillation in Small Models for Scientific QA,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Sci-CoT: Leveraging Large Language Models for Enhanced Knowledge Distillation in Small Models for Scientific QA,

Reference 59

Resolution
verified exact
raw_fallback, observed 2026-08-16T11:14:34.469818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.618376Z digest=sha256:6f5d05495f06f683ffa2af0b249e497bbc9d0f9b1fc05af7719719b965d6cb3b

Observation 2e8d94af-f530-4c5b-9e07-3c4d058a8757 · outbound

This paper cites Learning the Difference that Makes a Difference with Counterfactually-Augmented Data.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Learning the Difference that Makes a Difference with Counterfactually-Augmented Data

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.643474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.643474Z digest=sha256:2a22ace311f26a331773ce33dfd819d68b095f223ed3f03b4d7fbf4566318c69

Observation 4ebb9732-bf28-439a-a787-f4e1514b22ba · outbound

This paper cites Measuring Association Between Labels and Free-Text Rationales,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Measuring Association Between Labels and Free-Text Rationales,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.282778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.649662Z digest=sha256:d59f7fd0465dbc81fa109802ca1472245a957a77e2f53932c2cd965d945681f4

Observation 225e2463-0cf3-436e-beba-c3269bd93350 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Self-consistency improves chain of thought reasoning in language models,

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.636237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.636237Z digest=sha256:1a12236ea54cbb6eda3be60569bea663fd1a823469bea3b94d0433eec6f26a3e

Observation ac008d29-186c-4110-8306-28d05d17f829 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.660941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.660941Z digest=sha256:dae87fb9020eb01944336326ca50b4bc18edafedd5efa475b2bbb988800b6d32

Observation 8c144fd8-b7ff-4474-895d-920041793ea7 · outbound

This paper cites RLCD: Reinforcement learning from contrastive distillation for LM alignment,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability RLCD: Reinforcement learning from contrastive distillation for LM alignment,

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.240089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.666797Z digest=sha256:1928de488975bd660f56bfa5a54c82fbb3b3038535f4519722dd48949bb29fa6

Observation 00338ae6-a98f-43e8-b63b-8e589f616ace · outbound

This paper cites CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.262504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.655062Z digest=sha256:8b2e3cbfdcfaff2a4b0d79d032af255a30146f26ddeae344981d7faaca528ec1

Observation 9cca6e39-e9fc-45cb-94d1-cd1e82167214 · outbound

This paper cites Transformers: State- of-the-art natural language processing,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Transformers: State- of-the-art natural language processing,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.687221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.687221Z digest=sha256:e097ab525e282477e504b2f091474fc35c04294ec2db87654ca6ca923ce7d9c6

Observation c803238b-ad20-4141-b4cb-60ecf4a4824a · outbound

This paper cites Generating visual explanations,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Generating visual explanations,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.164806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.692926Z digest=sha256:40b264e972b834b2d998a21cabb43b813a38fc69a28664669e35cc46dacde5ab

Observation a63a36d2-7f36-4705-b61a-13668f557077 · outbound

This paper cites Available: https://openreview.net/forum?id=v3XXtxWK i6.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Available: https://openreview.net/forum?id=v3XXtxWK i6

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.220722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.675136Z digest=sha256:10af8b1c8539a24d1af3cc523ddeec71b7cc5a8b36188effaeb3bf39746155e3

Observation e399f061-52f8-41d4-8839-b75f29a21de7 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.201963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.681286Z digest=sha256:ef129c3cf317472d182e8c797bd6133c8192310ccd036e62f7062f5438128275

Observation 25b23c66-30be-43f4-888b-61102579c7a2 · outbound

This paper cites Exploring Evaluation Methods for Interpretable Machine Learning: A Survey,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Exploring Evaluation Methods for Interpretable Machine Learning: A Survey,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.095461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.719857Z digest=sha256:29d63ad4762a4d89ab660e1616247c51b29fed5f4a2025c2e324bebb47f1d775

Observation 38a404bc-869c-48f6-97b9-0020d1087550 · outbound

This paper cites Evaluating the Quality of Machine Learning Explanations: A Survey on Methods and Metrics,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Evaluating the Quality of Machine Learning Explanations: A Survey on Methods and Metrics,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.069454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.725446Z digest=sha256:5443c2bbe1341b248e0e4fe5ba3eed780843750454606a1923ec861350dd0ec7

Observation dc10041a-86fe-41c6-8535-0ecd00ef424d · outbound

This paper cites Prolific · Quickly find research participants you can trust.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Prolific · Quickly find research participants you can trust

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.141145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.699384Z digest=sha256:b6a6bb2c8ec4b689d90dc3a6c7f512276a15f240f0b272ef9ae225614be87664

Observation 600d2bdd-8c55-4ce4-bfab-a46d63be6a6b · outbound

This paper cites Oxford Learner’s Dictionaries | Find definitions, trans- lations, and grammar explanations at Oxford Learner’s Dictionaries,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Oxford Learner’s Dictionaries | Find definitions, trans- lations, and grammar explanations at Oxford Learner’s Dictionaries,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.019173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.737626Z digest=sha256:f7e0d2102672cee2661e7df14b93ad3a370a35fcc338265bd7918836fd4551a0

Observation 3d6068cb-7b34-457d-83f5-2b7b3bcd0778 · outbound

This paper cites Interpretation Quality Score for Measuring the Quality of interpretability methods.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Interpretation Quality Score for Measuring the Quality of interpretability methods

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:34.164010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.714226Z digest=sha256:6a757e9dfc5b5b4e3e11c9b95d01595be301675364366ca286152eb3538275bf

Observation 3fdea23a-9f9d-4a19-8025-a874fedfd1df · outbound

This paper cites Machine Learning Interpretability: A Survey on Methods and Metrics,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Machine Learning Interpretability: A Survey on Methods and Metrics,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:35.884649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.754508Z digest=sha256:b016da2b369f9d2578be588120f8ffd4d7f5e9496cc3ce831476ee08508efc6b

Observation 0805957b-667f-482a-9030-00eefd64f05d · outbound

This paper cites Leakage-adjusted simulatability: Can models generate non-trivial explanations of their behavior in natural language?.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Leakage-adjusted simulatability: Can models generate non-trivial explanations of their behavior in natural language?

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:35.827792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.761904Z digest=sha256:e0b1b655323d763e0b6e31273359728647f37b4ddb277e14087bbdba58b6b0f5

Observation 9f91952a-3cca-4886-b4d8-db24f7fb62e7 · outbound

This paper cites REV: Information-theoretic evaluation of free-text rationales,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability REV: Information-theoretic evaluation of free-text rationales,

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.046740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.730855Z digest=sha256:41fa5938f60bf079789c5ce8b781bf7e20e17bb5757d07aa01ada5ab92c53135

Observation d339e3a7-ec55-4a88-ad0a-73714273e4a9 · outbound

This paper cites A collection of principles for guiding and evaluating large language models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A collection of principles for guiding and evaluating large language models

Reference 78

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:14:33.976100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.774376Z digest=sha256:eb30d5aa0b300886b83e050ddf94e6f8f170326ec6460040f686d77405e21740

Observation 323866e7-6e58-48dd-97be-eb133c228396 · outbound

This paper cites Available: https://www.oxfordlearnersdictionaries.com/.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Available: https://www.oxfordlearnersdictionaries.com/

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:35.974511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.743200Z digest=sha256:cd165e37cdf078980da8b5c76fcb7de7ce18b2d012ba9ba7313988bd5b75db35

Observation 3ebabbc5-7f53-4593-b2dd-b92b6aead322 · outbound

This paper cites Measures for explainable AI: Explanation goodness, user satisfaction, mental models, curiosity, trust, and human-AI performance,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Measures for explainable AI: Explanation goodness, user satisfaction, mental models, curiosity, trust, and human-AI performance,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.748684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.748684Z digest=sha256:6467d39ead0d9cd78e503f6e08daf9291e47106b1593aa8ad5debdba29c9e47b

Observation 0d4e0bf6-5dc3-487e-a5b4-8c3402e7de39 · outbound

This paper cites Exploratory Data Analysis,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Exploratory Data Analysis,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.795484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.795484Z digest=sha256:7eb9413ab0c0ba56415cd84dd19f4948eb195ea2dc9c8234795d52fe1af0e710

Observation 3887b22d-771c-477c-8fa3-19ed49b8bfd4 · outbound

This paper cites Likert scales, levels of measurement and the “laws.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Likert scales, levels of measurement and the “laws

Reference 82

Resolution
malformed identifier
raw_fallback, observed 2026-08-16T11:14:35.726383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.802601Z digest=sha256:f4c6a60934d078fabf13aff66e4d549ed24c8faf923aeab1c4e17a7f40741f51

Observation abc4232b-48fb-4a5f-a94a-b64a6411855a · outbound

This paper cites Measuring the Quality of Explanations: The System Causability Scale (SCS): Comparing Human and Machine Explanations,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Measuring the Quality of Explanations: The System Causability Scale (SCS): Comparing Human and Machine Explanations,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.767813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.767813Z digest=sha256:e01c38d03012eb7010106b6cb847e37814829741ce6ddbab23e4e647f92eac40

Observation c8523882-fc5c-4cef-b152-c000da61d6f0 · outbound

This paper cites A Critique and Improvement of the.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability A Critique and Improvement of the

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:35.701353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.817036Z digest=sha256:8f1c98322631b01d8fb1ef2f0070aebcd350221a99f270304de89b448da30222

Observation 6c757055-83d5-4cc2-95ec-1101625928ac · outbound

This paper cites an unresolved cited work.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-16T11:14:35.796676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.781525Z digest=sha256:761395a05a52979ac2c5ab0dc85343abaf2742d843cc09bf62748bcec8d2cb39

Observation cf7422d9-6280-4b3d-bac9-58cef8b52da9 · outbound

This paper cites Bhattacherjee, Social Science Research: Principles, Methods and Practices.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Bhattacherjee, Social Science Research: Principles, Methods and Practices

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:35.760885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.788564Z digest=sha256:642d243e0185ba7651dae0c643844efe19602914fb13eaf8847ff0f0f9230d36

Observation a9d35239-64ad-4710-94e2-32d6b5fc3f7c · outbound

This paper cites Nonparametric Pairwise Multiple Comparisons in Independent Groups using Dunn’s Test,.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Nonparametric Pairwise Multiple Comparisons in Independent Groups using Dunn’s Test,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.809214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.809214Z digest=sha256:21a8aad221bee96be70871dfe12d1f73a8cc6ded00c9431489e96d76673a29bf

Observation 5e509fde-6c7c-4448-9ebc-7f118718c2d1 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability LoRA: Low-Rank Adaptation of Large Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-16T11:14:33.547063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:14:33.547063Z digest=sha256:af0259f40de3f1a7e93c1634c79da376f5f736c75cebe5ae7411e1d6985ba702

Observation 1bdaa55b-5841-4e5e-8125-86452fe0a676 · outbound

This paper cites Available: https://www.prolific.com/.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Available: https://www.prolific.com/

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.121111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.706941Z digest=sha256:421f0eb0e256dc56f93fb55d94d57cd2d7e464b508ab7f8dad7e3821055f884a

Observation 07ded7da-bf85-45ac-bbfc-2168717e6cdf · outbound

This paper cites Available: https://aclanthology.org/2024.tacl-1.85/.

Honey, I Shrunk the Language Model: Impact of Knowledge Distillation Methods on Performance and Explainability Available: https://aclanthology.org/2024.tacl-1.85/

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:14:36.719529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-16T11:14:33.324532Z digest=sha256:903fa3916f266d0bbc8d165d078f3e39e127dafc5b1bb8546c64e82e6107d00c

Pith citing papers

No inbound Pith citation observations are available.