Pith. sign in

Paper Citation Record · LEDGER

ByT5: Towards a token-free future with pre-trained byte-to-byte models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2105.13626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2105.13626 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:19:17.002827Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:06:58.467032Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b24f7341-1400-429e-a727-1d693db169ad · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 168

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:07.465115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:137c38466bfaf25a70354ed09efb247244b1c470937484fb98db14b59a8c20c4

Observation 9535f0fe-a3e9-452d-9c20-9fa2f60d2380 · inbound

SpeLLM: Character-Level Multi-Head Decoding cites this paper.

SpeLLM: Character-Level Multi-Head Decoding ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:19:17.002827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:19:17.002827Z digest=sha256:bcb927344cd8fd8b49882ee34d49335216fd884fbaeaccfa98ab0240b1b0924b

Observation 4d332bca-d614-425b-95a0-fe41d790c889 · inbound

chDzDT: Word-level morphology-aware language model for Algerian social media text cites this paper.

chDzDT: Word-level morphology-aware language model for Algerian social media text ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T12:17:30.454603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:17:30.454603Z digest=sha256:f8b6207070f9219dae95c29649a1ba827f4acf1f02186075142dda5f2585762a

Observation 4f98388b-ffa9-43f3-9e31-cae154c7295b · inbound

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty cites this paper.

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:14:40.890724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T13:08:33.172423Z digest=sha256:97d045b6ff4184714aa94057dd01c78b7d9b163d4615a2dc0a21c05eb3ea23ee

Observation 405de328-ea76-4f01-9d12-62617505fed6 · inbound

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments cites this paper.

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:26:35.412672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T09:13:26.501298Z digest=sha256:f14c7d23868fba23bcf61653e6740e58076fbe0ae688ec1f7c685c6d98501329

Observation a3f141fc-e872-4df1-a78e-7e876fff578b · inbound

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese cites this paper.

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:58.468663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-02T13:03:58.617626Z digest=sha256:89714f437f7b20256c573cd0727f85f948522c9b2f5b0c56622a5cfd4e72ddf3

Observation 7eb94844-1438-4a36-896f-bf3b19ecfd0a · inbound

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization cites this paper.

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T05:09:55.905970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:09:55.905970Z digest=sha256:0a84629fa940dc5dd7bf63f329771d45444ef54b26e40778b327a6385516b333