Pith. sign in

Paper Citation Record · LEDGER

Seesaw: High-throughput LLM Inference via Model Re-sharding

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2503.06433.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.06433 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:32.805859Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T19:25:31.176024Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 97826bf4-382f-47e4-a23f-1c48ba6253a6 · inbound

Hardware-Efficient Attention for Fast Decoding cites this paper.

Hardware-Efficient Attention for Fast Decoding Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:32.805859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:32:32.805859Z digest=sha256:3f610703aa07d18c0530f10feb6e0a93b97cbfc71f6c061e4f1f67fe7ce649ff

Observation a2843cd6-6086-4013-a65e-50a35eefc970 · inbound

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure cites this paper.

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:40.300419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:27:40.300419Z digest=sha256:a5409a5ab9484e5feaffcd3ef3d6ab573536b94fbcb66c31900ddd3048aa8259

Observation f90f5743-a83d-4467-9673-c90090343477 · inbound

Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference cites this paper.

Learning to Shard: RL for Co-optimizing the Parallelism Degrees and Per-operator Sharding Dimensions in Distributed LLM Inference Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T13:52:30.186309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:52:30.186309Z digest=sha256:33592e73fbebfd9cd977926094b38424a01711c0d51b1345ac2f3d383c0cdb4c

Observation aa051f31-3a6f-4b06-b44a-918fcd2191a2 · inbound

Amoeba: Runtime Tensor Parallel Transformation for LLM Inference Services cites this paper.

Amoeba: Runtime Tensor Parallel Transformation for LLM Inference Services Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:01:31.471942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T14:59:38.194894Z digest=sha256:f9ba7a6f388548e08a154b8d3564a091572d1e78a9e42a3113d38404d57a32e2

Observation 3032c9cd-70e1-4ccb-ad37-ff338f069620 · inbound

Understanding and Improving Communication Performance in Multi-node LLM Inference cites this paper.

Understanding and Improving Communication Performance in Multi-node LLM Inference Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:25:31.178045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T19:24:34.459446Z digest=sha256:c7e0a7808329ff408854f39dc527217469f7c5f1a4568f223ca12af74e1d2c0c

Observation 477b74ee-b1b7-40a1-81df-105cc8c4381c · inbound

PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers cites this paper.

PipeMax: Enhancing Offline LLM Inference on Commodity GPU Servers Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:10:41.559447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-08T18:49:56.357400Z digest=sha256:e2e0667bba3421c721628be3a4339df12e4a3b8e3b5aa3e7f76c8f6938820c25

Observation d03c1a45-79c5-4cc0-aa89-63acdc06e022 · inbound

Requests of a Feather Must Flock Together: Batch Size vs. Prefix Homogeneity in LLM Inference cites this paper.

Requests of a Feather Must Flock Together: Batch Size vs. Prefix Homogeneity in LLM Inference Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:46:09.132527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T14:06:01.934357Z digest=sha256:3a10a3f458bf0a7645bbf0590b7bb4f57856a10cc53e4a8c06b9d1f0006abcbf

Observation d5ce4034-caa2-4db7-b85a-033e718f27c3 · inbound

Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations cites this paper.

Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T04:56:20.002453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T04:56:20.002453Z digest=sha256:3f3b2d2bfeec8a9ba4d9547307f7e26a62ab37ac7d1de0ad2094f7250cc8d39e

Observation c21354b9-9887-4bd3-b3e7-198b286803d2 · inbound

Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations cites this paper.

Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations Seesaw: High-throughput LLM Inference via Model Re-sharding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:18.447867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:18.447867Z digest=sha256:aa466db40c423134b248510b996413a1a876613a2bc6800da964a48d26e32b61