Pith. sign in

Paper Citation Record · LEDGER

Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:1904.09482.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1904.09482 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T17:39:43.113333Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T12:05:43.645108Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 66fc3834-df64-4a0d-8379-927c900e7ceb · inbound

SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems cites this paper.

SuperGLUE: A Stickier Benchmark for General-Purpose Language Understanding Systems Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 119

Resolution
verified exact
local_arxiv, observed 2026-05-15T01:34:10.750216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-15T01:34:10.604864Z digest=sha256:e5388103fc1e8ca8cb12e73042d75810cf2fa4d1f0ef12805d66b98c0eaa14c4

Observation 09067d6e-ede4-4afc-a354-aa376cbfba05 · inbound

RoBERTa: A Robustly Optimized BERT Pretraining Approach cites this paper.

RoBERTa: A Robustly Optimized BERT Pretraining Approach Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-09T04:47:44.419233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-09T04:47:43.784327Z digest=sha256:68dd008b61d0dd14c43b4992deeab2c325a97d7a97a8e88b03ac8b823aed6e84

Observation 9ad833fc-1263-46d4-a4b4-351f49b19905 · inbound

Language Models are Few-Shot Learners cites this paper.

Language Models are Few-Shot Learners Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:05:38.287174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T12:05:38.045330Z digest=sha256:ef2379f5c54a2a1868b7c9c48d7b7983012b5bed2fba1bdcb036685a548ae5e2

Observation e433a410-ddc6-42cc-ade1-d9b807864690 · inbound

OPT: Open Pre-trained Transformer Language Models cites this paper.

OPT: Open Pre-trained Transformer Language Models Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T20:53:17.356706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-10T20:53:16.720145Z digest=sha256:49b057d44e85b78b2d84bdb4d06f9526286a0d27f184498b7ab3fd24039c7495

Observation 7710523c-049b-4ef7-af6f-651b572a44c7 · inbound

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization cites this paper.

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 75

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T12:05:43.646468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-07-01T02:32:19.425550Z digest=sha256:79c0522ce9f4724e06d92b03da6368d6fc55f9df1b8714bbbe135a4c6ae3dff6

Observation 4d60813d-d9fc-4c44-8f66-eea47d6b1818 · inbound

Learning in Deep Networks under Dale's Constraint cites this paper.

Learning in Deep Networks under Dale's Constraint Improving Multi-Task Deep Neural Networks via Knowledge Distillation for Natural Language Understanding

Reference 191

Resolution
unresolved
no resolver link, observed 2026-08-10T17:39:43.113333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T17:39:43.113333Z digest=sha256:528038030c862852e33c1622d9a77c5520926b08de5b0d02beb2100b576e400a