Pith. sign in

Paper Citation Record · LEDGER

What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2204.05832.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.05832 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T05:28:56.573142Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T09:14:16.427663Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fc8f302e-f3aa-45ac-a88a-6216586923dd · inbound

Flamingo: a Visual Language Model for Few-Shot Learning cites this paper.

Flamingo: a Visual Language Model for Few-Shot Learning What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:22:30.301662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T04:22:30.008355Z digest=sha256:81c17d9ee14c30cb3f9a48d25a6db778e86204b74bf7bb32523131967a410e81

Observation c2e2fcd7-b772-43f7-a518-6c368b7af8f0 · inbound

Emergent Abilities of Large Language Models cites this paper.

Emergent Abilities of Large Language Models What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:38:38.233526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-11T07:38:37.734402Z digest=sha256:7e553de355d8627a328adfa4fa987c50c9c8f840bb8aeaefa4fe6714d1c4ff48

Observation ac462a33-3b21-43aa-9cf4-c6e06d0879dc · inbound

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning cites this paper.

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-24T09:14:16.430799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-24T09:13:30.054153Z digest=sha256:b7b32fcf85374e0d2b0147aefe17085566d3d235c2eb4966e43a14ce1a6a683b

Observation 3228e870-568a-4312-94fc-2d83e6eb31b9 · inbound

The Falcon Series of Open Language Models cites this paper.

The Falcon Series of Open Language Models What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:46:10.099858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-16T09:46:09.701440Z digest=sha256:e80112869fd9f2c3402da019594e3b593b6733404e1cae728d86c8a01bb0f906

Observation 4a3f76f2-50fa-42fa-a61b-ba30395046ff · inbound

Benchmark Data Contamination of Large Language Models: A Survey cites this paper.

Benchmark Data Contamination of Large Language Models: A Survey What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 154

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:10:41.115358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T23:10:40.420241Z digest=sha256:e15047eecf866f6c6a1186b370ae4bc4576be164e7f8d76d1d89ccb6bfe00787

Observation 76e8086c-214f-41de-b50b-ecc3d91ab6f6 · inbound

Fine-Tuning Pre-trained Large Time Series Models for Prediction of Wind Turbine SCADA Data cites this paper.

Fine-Tuning Pre-trained Large Time Series Models for Prediction of Wind Turbine SCADA Data What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T05:28:56.573142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:28:56.573142Z digest=sha256:b9f086ca458d241d58d3ea54769d875a5ed19fdf051076e02648d7e217c053a4

Observation 57d26bf0-5097-4f8a-864a-88d0b2a76467 · inbound

Improving Language Transfer Capability of Decoder-only Architecture in Multilingual Neural Machine Translation cites this paper.

Improving Language Transfer Capability of Decoder-only Architecture in Multilingual Neural Machine Translation What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T23:54:06.547764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T23:54:06.547764Z digest=sha256:1e30a3b385ce580b685576a13f47c9d09bfda4bdc2bf276511a8396575656746

Observation 6d3fbd3f-d555-4de2-9fc1-11758331c79b · inbound

A Primer on Large Language Models and their Limitations cites this paper.

A Primer on Large Language Models and their Limitations What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T23:52:46.821787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:52:46.821787Z digest=sha256:9b14f86ed0b27a14261dea46c04c9954dd170bd50d36bcd054aa97286d990785

Observation dabb081f-ce5f-4125-99cb-800643504c8b · inbound

Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding cites this paper.

Top-Theta Attention: Sparsifying Transformers by Compensated Thresholding What Language Model Architecture and Pretraining Objective Work Best for Zero-Shot Generalization?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T05:29:16.641264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:29:16.641264Z digest=sha256:95f5e3211c46d6502c897b271fd52c6975208292b897068427eb781e89ff99a4