Pith. sign in

Paper Citation Record · LEDGER

Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2310.03477.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.03477 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:41:13.911515Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T00:38:54.558626Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 092d28b1-d6e8-4de7-aee7-0cec5577cef0 · inbound

Bilingual BSARD: Extending Statutory Article Retrieval to Dutch cites this paper.

Bilingual BSARD: Extending Statutory Article Retrieval to Dutch Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T18:53:28.301573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T18:53:28.301573Z digest=sha256:404842f804c1b9d973a40818f66a90b3b9ed1fb206d2db3181086c871420f8c8

Observation 972bfbc2-b982-4496-85a2-f6865fc9acda · inbound

ChocoLlama: Lessons Learned From Teaching Llamas Dutch cites this paper.

ChocoLlama: Lessons Learned From Teaching Llamas Dutch Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T18:42:48.206508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:42:48.206508Z digest=sha256:0903f0f275994d80086d810e2221e11df5577940034bf47b0858603441cd5492

Observation 1b093bc9-d10c-4b2d-bebf-59dad12f6318 · inbound

Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages cites this paper.

Prompt, Translate, Fine-Tune, Re-Initialize, or Instruction-Tune? Adapting LLMs for In-Context Learning in Low-Resource Languages Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T18:41:13.911515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:41:13.911515Z digest=sha256:835bfb226ebf481e11785671196a12dcd89ea4ce44e7f4cf941ebcaac8d5a49f

Observation a8665cf7-5cfd-4b4a-929e-248b00747615 · inbound

Conditional Unigram Tokenization with Parallel Data cites this paper.

Conditional Unigram Tokenization with Parallel Data Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T18:35:44.523916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:35:44.523916Z digest=sha256:f0c43d600babf2a8efd8c294b7b3c6d0266ae660359e0fe1c69c25fe88b0d67a

Observation a27bf712-3fd6-4dd1-a375-8dfa8fd97caf · inbound

Writing-System-Level Tokenizer Adaptation for Byte-Level BPE cites this paper.

Writing-System-Level Tokenizer Adaptation for Byte-Level BPE Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 2019

Resolution
verified exact
local_arxiv, observed 2026-08-05T00:38:54.718186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-05T00:38:53.671265Z digest=sha256:739fc7d3905f63c609cf5c1ac2f28f8c3b19e60aeabd25b145edffe2a366f66c

Observation 81a440b6-2571-4595-9fb3-0738d6ccb310 · inbound

Disentangling Language Modeling and Boundaries cites this paper.

Disentangling Language Modeling and Boundaries Tik-to-Tok: Translating Language Models One Token at a Time: An Embedding Initialization Strategy for Efficient Language Adaptation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T16:21:01.287738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T16:21:01.287738Z digest=sha256:03889b4b5e3118814e4520b3aa060b58f9dd45829839902f6237cc391853911a