Pith. sign in

Paper Citation Record · LEDGER

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers

As of 21 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2502.03793.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03793 v2

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T00:49:12.272682Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:48:07.106359Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T17:48:07.604735Z

Reference resolution

56 of 56 outbound references displayed

  • verified exact3
  • verified fuzzy13
  • unresolved38
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cc187f64-08d4-4506-8f5c-9dfb66ab5f66 · outbound

This paper cites Devlin, M.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Devlin, M

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.089123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.089123Z digest=sha256:aca39f9346063f139114cbd85916b27e29db896071340f5b3269d3fbed4d4480

Observation 22cdd5a5-e432-466c-adbc-cbac4b4dd012 · outbound

This paper cites Vaswani, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Vaswani, N

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.777086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.092704Z digest=sha256:15a5e91568909994a68ffb5d8bcbbff4116944b70084a8f82ab2657f1f9bcfe0

Observation 49986245-a2c9-4247-bb85-3dec3673329e · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.095682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.095682Z digest=sha256:257b0e06c16f0158601322e088f4a15a84ac681a416f6c82b56ef85457f4ef9e

Observation 400316d9-f446-46eb-a295-5d494dcfdad4 · outbound

This paper cites The Llama 3 Herd of Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.098714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.098714Z digest=sha256:1f0e792b5ad4ec75eb9932b3232db8c7e8154c062eb32a198fb2f117a9f3a949

Observation 55f7d166-2e43-4751-9fcb-89d6573861f8 · outbound

This paper cites Qwen2.5 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Qwen2.5 Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.101823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.101823Z digest=sha256:5038aefb50472834b19deca83a37eea658bded940aea4611272714e54a164ec9

Observation ea9ebb7e-5310-4cad-ba3d-7e56cfcf2934 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.770286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.105079Z digest=sha256:f6598b63f595093f983514f8a7f8cf270c9dbfbca84a1445c19c4709daa77e3c

Observation 04652211-c8e1-4e13-96eb-02bfcbb15305 · outbound

This paper cites Radford, J.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Radford, J

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.763255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.108155Z digest=sha256:a8e097c40af592ccc41937ff6a49418f451e2f264890bfd7b92abe0b82b31282

Observation 00a61132-d8b6-42e9-b6f6-2ceac0fed9f1 · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Finetuned Language Models Are Zero-Shot Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.110889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.110889Z digest=sha256:c28d384e5843a0b21ca842a956b40d17a80ac8850b26a4ca8d0e7d00e1fdc15c

Observation d2d252ac-0ae0-4771-a493-3480d848cc2a · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.755261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.117045Z digest=sha256:5c72da8d568efcd65af330aaf63db62580fa6f58d2c551573987317d642db320

Observation 384a8075-f251-4be6-8579-106cb6463d61 · outbound

This paper cites Warner, A.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Warner, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.746731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.120107Z digest=sha256:6831e3ba40295cfa6f2055be4ca413f983231f7b91e13ffa730e4e98fadbd072

Observation 029cfc9e-e971-4cb0-afbc-c1fee4e52d77 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.124370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.124370Z digest=sha256:5add87c51c59fd33c68bfb5cbc504b0f705dc798c12ae0458a7dede6ced08ba2

Observation 42474086-e085-4513-bc04-7a813f4f7b3e · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.127371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.127371Z digest=sha256:0839bbc85602c03a06e6e28601be08b2d0d969be6e8ee9c34842fd5721074b7d

Observation 4f511a31-509a-484e-a37b-f6f478050150 · outbound

This paper cites Williams, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Williams, N

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.738655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.130541Z digest=sha256:7e05b327508b3a0a2fc1857430dd00558a3670ada41975fd342b21ddbf059faf

Observation 57a9767d-54d3-4d57-91c5-e4d6643aa27f · outbound

This paper cites Lewis, Y.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Lewis, Y

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.133110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.133110Z digest=sha256:dc6b6a63b0c852558538f46b1339e5b63263e0dc973a5ab976fbfb5360bdfcab

Observation 6c0f14fa-70c8-44af-b1f4-7ee5c8641108 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.730589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.135882Z digest=sha256:3779fd29d43d97cc3561c2341698a2d7d5bcb2b64a1d00638a4fb0ecde44780c

Observation 81724653-4351-41d6-96b8-b1da6a36bc38 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 17

Resolution
malformed identifier
no resolver link, observed 2026-08-09T00:49:12.138705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.138705Z digest=sha256:b6808f334e1bb597c9c2c0e04976372074961a8a5c29a04894e0aa65f54ee220

Observation 677a9b5e-a461-4546-8cea-3d2c65f71208 · outbound

This paper cites cloze procedure.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers cloze procedure

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.722993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.142256Z digest=sha256:d2826c305822c935ed6f2d2999e80a70dd93ba33ee08c3f08796fbe11398a319

Observation 1883359b-0aa2-4ad8-b81f-60b4cc9357a9 · outbound

This paper cites Schick, H.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Schick, H

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.147694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.147694Z digest=sha256:8913bf53a172341bb612e73869f2b1c0bbe7b14b820fd9a3b8958befe8a56bea

Observation 727558e6-826b-48f3-b14e-2ef42269f5aa · outbound

This paper cites BERTs are Generative In-Context Learners.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers BERTs are Generative In-Context Learners

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.150658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.150658Z digest=sha256:02c6d34ad4d6b5d4bb83c036b9c1944dd59cf4d48e96c7e6cac4755357b40a54

Observation 5eb73e16-027b-4e3e-ab2d-7b9b91c7bc34 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 22

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.333868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.153789Z digest=sha256:6fa64fa5911772c8f512533465365b5fd19abdc578f9187b351384a6c5a7dc82

Observation 9d14c827-305c-4ec8-827a-3e4a1f7c6089 · outbound

This paper cites DeepSeek-V3 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DeepSeek-V3 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.156833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.156833Z digest=sha256:e2dbf5080f1d50a79f13bae653eab99c1246a0bcf1c64b7226e47584bfdf8abe

Observation 2cc7376d-b3f4-4990-942b-f6a6e2989072 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.159808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.159808Z digest=sha256:0387ea6f43853d6467b35c3c8e4081b01ba91c5935b17f21a38385168a596671

Observation 49fb940e-5250-4f1d-a5cb-e951956d090b · outbound

This paper cites Qwen2 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Qwen2 Technical Report

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.162737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.162737Z digest=sha256:3f73536bf42e74b55fadbcb5164ea0b2cd6aa0ce545b9549f03c1b1e797c3db6

Observation 72e65d2c-17a6-4faf-a9f1-c353d719ebab · outbound

This paper cites Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.166031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.166031Z digest=sha256:ed3dc32cb7e60ea6b4418ca3caccd66c26ca17ae53956fd1132570e83573d342

Observation e8469c3f-7a12-4714-ba72-5335a943d876 · outbound

This paper cites DataComp-LM: In search of the next generation of training sets for language models.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DataComp-LM: In search of the next generation of training sets for language models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.169403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.169403Z digest=sha256:1b85c3e9814eeb73b479fd27e8ec7485dea48c3891e2bd60c7774263d1e742c5

Observation 48009407-6cc4-4b6b-84fa-d505198f9747 · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.172991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.172991Z digest=sha256:77b729f14452033854df1306f28d86fed37705c98b04fcb3b60915253aa3c383

Observation b9053eac-b2a3-4d70-a394-df095850ad93 · outbound

This paper cites Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.176185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.176185Z digest=sha256:aa2fe737fbf4584d4dc11a6f0320693ea07e1e373a1c074778736b6c5cec4f86

Observation 9c65427e-5506-45f0-a57f-493555cb3691 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.714808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.179707Z digest=sha256:fecac71c2a0fdfa8ae98724ee83d26dd2bd62dcafef444921d2e948c77f59529

Observation fd3c91fe-112f-4c6c-aae9-2a9de888dd46 · outbound

This paper cites Zhang, Y.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, Y

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.706544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.183270Z digest=sha256:cc9d117ae539c4ca51e9eca85acf994383b073fca4759b825180b97706315929

Observation 1575c2c5-855c-415a-8bec-64d3902e0bae · outbound

This paper cites Javaheripi, S.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Javaheripi, S

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.698090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.186606Z digest=sha256:efe2a3f1049f2ec60cbdacc08625a84e91d01547081e34a83fba1d55105e49cd

Observation b130bb73-1fce-494b-aaf2-bf70cd1024e0 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.189956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.189956Z digest=sha256:e16898d4e61f96892909521a3b315eae285333432c10ff5c7f863d81a498d242

Observation 49d80b63-c591-4c7c-978a-b3b79a642df8 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.690578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.193102Z digest=sha256:e9369d4266497a1a22355220605ee6b545a7d4eaafc3e549d2d2b7b36549add3

Observation c08affc9-d785-4f9a-968a-9606dcf46c3c · outbound

This paper cites DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding Sharing

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.195565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.195565Z digest=sha256:1c50b34920e6e68b0ea76f000617f96dbc7c1c56b3a66846a5f2e3c9f95d8852

Observation 67a90a28-3383-4623-b097-725b80457f00 · outbound

This paper cites Sutskever, O.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Sutskever, O

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.683142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.198723Z digest=sha256:c87c11c5a177d78c9cca830c8bb38bd808b34f1448f205c167529e6761495fdf

Observation 9c4216a1-3d61-45b9-a617-0cdd483ee7f9 · outbound

This paper cites Ra ffel, N.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Ra ffel, N

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.201569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.201569Z digest=sha256:b33e2764cc67310790523568aa01b80b2b4ce861cefc3b162b3ad1e1eb1a2d75

Observation eabbf38e-33b3-4748-91fc-a3fd99de6f55 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 38

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.312478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.204286Z digest=sha256:249b5d72ce462df847deb9fefb8269777391f8a25d606c10b54b40e695d28414

Observation 7c060403-72a8-45f4-97ea-5dd4431e0d1f · outbound

This paper cites Longpre, L.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Longpre, L

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.670374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.207043Z digest=sha256:21f9679b0836509956bbaeff4fefda1bf155bae99e3d77b234e64fed89ee7b24

Observation d1d0f590-36a1-4de5-ae50-26a263b8076d · outbound

This paper cites Hendrycks, C.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Hendrycks, C

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.662642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.210834Z digest=sha256:94f3838c183eeb72a5fd035ef120539f55c31c3bbe3b83080521f9c52c069e53

Observation cecd3773-fd12-44e2-90f6-cbc869d0af81 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.213813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.213813Z digest=sha256:37bb5b16244ad71d2d5c19fe0f453ee4d794548ad45bd74c99e267ae9ca32553

Observation 6fef0aae-67cd-4a5a-8bf2-3240be7a6197 · outbound

This paper cites Srivastava, G.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Srivastava, G

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.217095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.217095Z digest=sha256:bda34339596b583b8d5498469f8a4bb3d7758bf4316d2f8f0d7ae6c4df5e55a1

Observation 70ce846d-d0da-4fce-9eec-c1496881593b · outbound

This paper cites Khattab, M.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Khattab, M

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.650911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.220142Z digest=sha256:95908648a1c01bf809060b0915da092ef489d91d363c36f1e8160a35e271e8c9

Observation 4b016110-2dab-4397-bf2e-92efff9fd97a · outbound

This paper cites Text Embeddings by Weakly-Supervised Contrastive Pre-training.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Text Embeddings by Weakly-Supervised Contrastive Pre-training

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.223389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.223389Z digest=sha256:258faece15841c59676e82d44522f604d778e8a82175c77aa96a8dbaf10ba572

Observation 31aa18ef-6a75-4636-9511-47909dc375cb · outbound

This paper cites Zhang, et al., Jasper and stella: distillation of sota embedding models, arXiv e-prints (2024) arXiv–2412.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, et al., Jasper and stella: distillation of sota embedding models, arXiv e-prints (2024) arXiv–2412

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.641399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.227145Z digest=sha256:2162e6dba08dae7234ea2aafc8f9ac7b573c43377b1fed2981a27f56f3d54ed8

Observation 56b514b6-0342-4bc5-ac0d-acb81f14916a · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.234415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.234415Z digest=sha256:29ef11ada8f8b7d69c642d490bdc84cf8592577f3ebd4f7153027c146625131b

Observation ff5c76f0-f918-4ccc-93fa-93725ff048de · outbound

This paper cites RAFT: A Real-World Few-Shot Text Classification Benchmark.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers RAFT: A Real-World Few-Shot Text Classification Benchmark

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.237357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.237357Z digest=sha256:60ce0aba628719ef6e5023e54a8eb0b01839f66ad64718260c396927ec541196

Observation c26a35cf-1d59-43d3-aedd-873f115ce4ca · outbound

This paper cites Gurulingappa, A.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Gurulingappa, A

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.240445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.240445Z digest=sha256:7c189bc35f6b055f60f7998da31fc477fa822a756cbbedb0cbb934da1a520dfe

Observation 30f134f6-211e-41aa-bc9c-888d74ff69e0 · outbound

This paper cites Vajjala, I.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Vajjala, I

Reference 49

Resolution
verified exact
doi, observed 2026-08-09T00:49:12.300315Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.243751Z digest=sha256:6511c43ed7f4431ccc43cff79c954e67653c93c6dcea2e3ae90916ef3b416b90

Observation 2bf2b1d9-718d-44bf-b78a-68021688fab4 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-08-09T00:49:12.246217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.246217Z digest=sha256:efa4c1aa2590601eeb77d3da7560bb4a649f88f177f05fbe76f76f44017e7640

Observation d012e6c7-8181-4285-ae60-d561652ec04f · outbound

This paper cites Zhang, J.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Zhang, J

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T00:49:12.633224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.248745Z digest=sha256:e3f6433521e96cefbf73e70de15c72c9f9ddc5eb421dbb319614ba3313675110

Observation f94c6c45-71ce-4498-bce1-cc7bfe0b2175 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.625095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.252185Z digest=sha256:121f44a7b4ee8d928bec15d3847b20bc9f31341463bd620f364f3c1bd3947369

Observation 853fa846-d311-4f34-87a9-536cb65c7633 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.255533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.255533Z digest=sha256:dd81edeb6ed37d82dc9551f1a389cc9aee49715f36ce7712dd7e81fd7240de13

Observation 37139d8d-fe31-4959-8b98-5aefdd56cb93 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.259260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.259260Z digest=sha256:1cf61230fa6f542eb456986f28bd7249448b11a85c3207206b5710921776aae2

Observation eddd7f55-8fb1-4944-80ca-562e3ffc17f2 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.616581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.262838Z digest=sha256:21e8fc281623898c7678e438c1ddad4a375c3e140f9b70398b56ee395525adda

Observation 02a1e252-c7d2-49dd-9dbb-d72c34d7d4b6 · outbound

This paper cites Hermes 3 Technical Report.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Hermes 3 Technical Report

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.266170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.266170Z digest=sha256:3a69c2d35dff971a643e002849a284b1cf5d61e0cce0c03f174c0b63f8a27fe7

Observation 0a211ce9-50bb-4526-8191-9360ed615c75 · outbound

This paper cites A Survey on In-context Learning.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers A Survey on In-context Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T00:49:12.269561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T00:49:12.269561Z digest=sha256:974e9cfda5e7be2571e5f005314c3d18525d631420c86fb72d7c0c3acfbb34c9

Observation 75fdcb7c-54bd-4507-95c5-3b735ac1ead5 · outbound

This paper cites an unresolved cited work.

It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-09T00:49:12.609062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-09T00:49:12.272682Z digest=sha256:64d33152c2f33e2c8785ab2fa10d842496715267c3b2d9849bf0fbb3732d9d10

Pith citing papers

Observation f86633c8-6b4d-4265-b4bc-f7dbf11bcbe6 · inbound

Tiny Reward Models cites this paper.

Tiny Reward Models It's All in The [MASK]: Simple Instruction-Tuning Enables BERT-like Masked Language Models As Generative Classifiers

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:48:07.607922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-08-06T17:48:07.106359Z digest=sha256:7be76694c7025376f8085548c32142f82826f27bf2d828577edca747eaf86d80