Pith. sign in

Paper Citation Record · LEDGER

TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

As of 25 July 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2410.06511.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.06511 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-24T06:31:00.690269+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T17:24:35.961717Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b62ac5ff-08e4-45d8-936e-b76c8dac6155 · inbound

MAGI-1: Autoregressive Video Generation at Scale cites this paper.

MAGI-1: Autoregressive Video Generation at Scale TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:31:15.813662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-13T20:31:15.700943Z digest=sha256:73dd90cb6f4f1901b21d86995a3648774ae3b8df70c74a5240d4fed673056b60

Observation e682afbd-f93a-4c04-9456-ff36c8ee2004 · inbound

FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space cites this paper.

FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T16:36:02.129439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T16:36:02.016696Z digest=sha256:9ea6d09da4928dbbccb00a24340dcb32a29384a982fdb8729d6f5114b6a7fab8

Observation 95ea3d80-ac3c-40c3-882e-ff5eea42491e · inbound

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models cites this paper.

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T23:21:21.580170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-16T23:19:02.358348Z digest=sha256:377554bb97dcf76623d8d6bcee94347432b5b5478a9c5f45eac1d7d5d581ef85

Observation 3f5ab610-3b51-4240-bd77-2165d03a2bb1 · inbound

Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML cites this paper.

Flint: Compiler Enabled Cluster-Free Design Space Exploration for Distributed ML TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:25:55.168463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T05:23:10.718195Z digest=sha256:30747790dca3ea32ced82fade20267cb674a97d8b03e257700de5fe444485363

Observation 1175d644-e51a-49f6-a538-8ff8ed301a29 · inbound

UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training cites this paper.

UniEP: Unified Expert-Parallel MoE MegaKernel for LLM Training TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T13:06:05.145883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-10T02:20:00.625923Z digest=sha256:9a0e1599bda6938965e1619478ad5e3f05b696393aecd930e18e7121800e1051

Observation ba7e42bd-b22a-4cd8-a2b6-13d5f377aa58 · inbound

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation cites this paper.

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T04:10:08.719495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-07T10:05:24.009460Z digest=sha256:64cc742c0f8cf36ee624d34fed96bd8cee4ed88277afe436bfbc5d089cad1be4

Observation e1ccaf99-0a12-47fe-a560-e2ea3f7c894a · inbound

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation cites this paper.

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T06:59:49.123448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T06:58:22.362262Z digest=sha256:e4af08cd7fe976a8880788ef7efda12645d994dea1d49d6a61d0a20e3c546ceb

Observation e2f07f89-81ed-4330-896b-ba542b4f3655 · inbound

A Rayleigh criterion for mechanical instability: inducing activity by chemo-mechanical coupling cites this paper.

A Rayleigh criterion for mechanical instability: inducing activity by chemo-mechanical coupling TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T17:24:35.961717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:24:35.961717Z digest=sha256:988eab2fb1b6942aca0537a962ba660c46d17b27dbcd08404bf3172bfad6c76b

Observation f512b045-6c78-4794-b02b-0263e3982fa9 · inbound

CCL-Bench 1.0: A Trace-Based Benchmark for LLM Infrastructure cites this paper.

CCL-Bench 1.0: A Trace-Based Benchmark for LLM Infrastructure TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:31:17.692290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-08T05:08:54.234138Z digest=sha256:f9a6522d53e3df9abd5b959fc068381d3866d810c2d8e538ce6024840535cced

Observation 5b103c06-654d-4126-83ce-0784668f198d · inbound

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale cites this paper.

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:06:27.867024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-12T04:33:41.411292Z digest=sha256:51cf11857352d33301f7bb96423b452497db26bfd7d34cb0369185166e1c2f28

Observation e412fe96-077e-42eb-821a-b7431523de41 · inbound

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale cites this paper.

LoKA: Low-precision Kernel Applications for Recommendation Models At Scale TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:59:46.082489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-15T04:55:01.973832Z digest=sha256:5cab6c9bf69aba820926ac4f18156a0767c5265925b5e498fb89489e718815d2

Observation b9dba34c-b3e3-4fe3-bbb7-ddd8db792513 · inbound

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs cites this paper.

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:23:07.108084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-20T07:19:22.751655Z digest=sha256:f0b7e6b449a88e422e07a18b2c67280242402706b3241b75d17c4a4cb04005b5

Observation bf0ea9e9-a2ed-4cb4-b892-b03a12f6522a · inbound

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs cites this paper.

CODA: Rewriting Transformer Blocks as GEMM-Epilogue Programs TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:24:02.563243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-21T07:20:56.839167Z digest=sha256:48c18344717d0b542f8a59b415463eee9cd5e3dc74b6bbcc39c65b1fb723a9ae

Observation d2a7d010-f8b8-4022-b246-1fccc711969f · inbound

DynaFlow: Transparent and Flexible Intra-Device Parallelism via Programmable Operator Scheduling cites this paper.

DynaFlow: Transparent and Flexible Intra-Device Parallelism via Programmable Operator Scheduling TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:36:17.026922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-05-22T08:34:47.923752Z digest=sha256:3c766d574ca55e28b0407eaa9aeeadbbc5ae75fa8e74062eb86940e9fc05fdf6

Observation 73ec02c8-d106-4cfc-9ca3-bda0d954d474 · inbound

DMuon: Efficient Distributed Muon Training with Near-Adam Overhead cites this paper.

DMuon: Efficient Distributed Muon Training with Near-Adam Overhead TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:39:58.570696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-06-26T02:41:02.917064Z digest=sha256:cddd98afa1086bf0a2e8c7896322f4f84deaa21cd37f50431742dc968cec60e8

Observation 25f3fb68-a996-4130-879c-0b1f9a03315a · inbound

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models cites this paper.

Mixture-of-Parallelisms: Towards Memory-Efficient Training Stack for Mixture-of-Experts Models TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T06:17:41.158704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-07-03T06:16:56.851399Z digest=sha256:77adb424c2a8b8ad5b4c083dcf6dc519dddf5dd7b834ccff49652c60638b24f6

Observation 286fa92a-4180-4ff2-8b61-e59afc8690a6 · inbound

MatrixFSDP: communication-free matrix optimizers under ZeRO-3 parameter sharding cites this paper.

MatrixFSDP: communication-free matrix optimizers under ZeRO-3 parameter sharding TorchTitan: One-stop PyTorch native solution for production ready LLM pre-training

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T21:35:37.641393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-24T06:31:00.690269+00:00.

source=pdf_text observed=2026-07-08T21:33:20.705837Z digest=sha256:4c45a885d0c9d3e376697b3a8820d8972a765c64c305d75748aa58e1ebe119a9