Pith. sign in

Paper Citation Record · LEDGER

ByT5: Towards a token-free future with pre-trained byte-to-byte models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2105.13626.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2105.13626 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:19:17.002827Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T13:06:58.467032Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b24f7341-1400-429e-a727-1d693db169ad · inbound

PaLM: Scaling Language Modeling with Pathways cites this paper.

PaLM: Scaling Language Modeling with Pathways ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 168

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:45:07.465115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T23:45:06.755839Z digest=sha256:678f3171b96e7bab2c96985c080a9903218f9a85e6085404904ab1e9ad231e6b

Observation 9535f0fe-a3e9-452d-9c20-9fa2f60d2380 · inbound

SpeLLM: Character-Level Multi-Head Decoding cites this paper.

SpeLLM: Character-Level Multi-Head Decoding ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:19:17.002827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:19:17.002827Z digest=sha256:bcb927344cd8fd8b49882ee34d49335216fd884fbaeaccfa98ab0240b1b0924b

Observation 4d332bca-d614-425b-95a0-fe41d790c889 · inbound

chDzDT: Word-level morphology-aware language model for Algerian social media text cites this paper.

chDzDT: Word-level morphology-aware language model for Algerian social media text ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T12:17:30.454603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:17:30.454603Z digest=sha256:f8b6207070f9219dae95c29649a1ba827f4acf1f02186075142dda5f2585762a

Observation 4f98388b-ffa9-43f3-9e31-cae154c7295b · inbound

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty cites this paper.

The Tokenizer Tax Across 25 European Languages: Domain Invariance, Cross-Lingual Few-Shot Effects, and the Ukrainian Penalty ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:14:40.890724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T13:08:33.172423Z digest=sha256:46216d3dfc25cba64edea2b617b9397c3238f285a53eecec4c6e8695930740c1

Observation 405de328-ea76-4f01-9d12-62617505fed6 · inbound

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments cites this paper.

MimeLens: Position-Agnostic Content-Type Detection for Binary Fragments ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T04:26:35.412672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:13:26.501298Z digest=sha256:9c4abae551f5af208fe54e2bcc7e9d061cea4a17913b57584d871778d07a772e

Observation a3f141fc-e872-4df1-a78e-7e876fff578b · inbound

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese cites this paper.

YOMI-Bench: A Benchmark for Evaluating Kanji Reading and Phonological Understanding of LLMs for Japanese ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:58.468663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-02T13:03:58.617626Z digest=sha256:e8114e8395cdeb3990b7abb9f92e7bcbb5870981130034f30f5e76c3459bf269

Observation 7eb94844-1438-4a36-896f-bf3b19ecfd0a · inbound

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization cites this paper.

Cross-Tokenizer On-Policy Distillation via Byte-Prefix Marginalization ByT5: Towards a token-free future with pre-trained byte-to-byte models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T05:09:55.905970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:09:55.905970Z digest=sha256:0a84629fa940dc5dd7bf63f329771d45444ef54b26e40778b327a6385516b333