Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:54:46.812096Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 1 inbound Pith citation observation for arXiv:2505.02533.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T00:54:46.812096Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-08T09:45:57.201837Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T20:16:10.526836Z
17 of 17 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6a272bea-1e7c-4b8e-bc43-2cb9a607dc3e · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge EdgeShard: Efficient LLM Inference via Collaborative Edge Computing
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf602b55-b084-4bb5-8deb-4cd5c7136544 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge SplitLLM: Collaborative Inference of LLMs for Model Placement and Throughput Optimization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a82c498-8154-4fba-85de-ccc1754349e3 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Galaxy: A resource-efficient collaborative edge ai system for in-situ transformer inference,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac62089e-ffc0-4aed-a02c-a83caae85127 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Language mod- els are few-shot learners,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e0edf7a8-d28a-45ea-93bb-41b95800e1f1 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge 6g technology overview,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5256ce74-4766-410f-b1c0-838730c2dac3 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Dnn partitioning and inference task offloading in 6g resource-constrained networks,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f2541d85-f014-479b-a75f-e832e283a9e8 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Splitplace: Ai augmented splitting and placement of large-scale neural networks in mobile edge environments,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 769c876a-f3ff-4c81-92c2-2d9ea2300583 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Joint optimization of the partition and scheduling of dnn tasks in computing and network convergence,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3f934288-8d4d-40db-a7c1-b8df6506c1fe · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Joint optimization of dnn partition and continuous task scheduling for digital twin-aided mec network with deep reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 64a78c78-37e5-43fc-9dde-006eb77ebf0c · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Huang, Y
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3391209d-cded-4304-821b-27f1cd320ebd · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Pipedream: generalized pipeline parallelism for dnn training,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c0667170-2416-4641-b0ab-8d515958be94 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Autopipe: A fast pipeline parallelism approach with balanced partitioning and micro- batch slicing,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d2f0ab22-8510-4a26-92a6-ec7f821d95c9 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Memory-efficient pipeline-parallel dnn training,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1d614240-fd22-4cae-a086-59489a1db3b5 · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge 3d parallelism for transformers via integer programming,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 457e8af3-24fa-402b-b136-fa2a9b3e526d · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Distributed Transformer Inference Simulator
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f32dd22c-79ea-43f2-8aaf-30b4c766e4cc · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Google cluster-usage traces: format+ schema,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4c9939d6-f98b-4cfa-831d-019d7571081a · outbound
Large Language Model Partitioning for Low-Latency Inference at the Edge Attention is all you need,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d738a7-7593-408d-b3bb-463e39467121 · inbound
Network Edge Inference for Large Language Models: Principles, Techniques, and Opportunities Large Language Model Partitioning for Low-Latency Inference at the Edge
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.