Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:23.582268Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2505.14085.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:23.582268Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c86273b6-479b-4f88-b177-5d17cecd19d0 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration 6g wireless communication systems: Applications, requirements, technolo- gies, challenges, and research directions,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a715e3e-690f-4f32-997b-9cf3b6109eb0 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration White Paper on Broadband Connectivity in 6G
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11bedb4e-9721-4d3d-ab78-bd7c0e58505f · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Ai-native network slicing for 6g networks,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8ae12871-c2b9-44fc-a741-3ef16b254497 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Towards 6g wireless communication networks: Vision, enabling technologies, and new paradigm shifts,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3aaa0cb3-ba2d-4600-bde1-edcc9012da77 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Optimizing llm inference clusters for enhanced performance and energy efficiency,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation baca5998-54b7-4d72-ad0a-dd51c56179a4 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Performance modeling and workload analysis of dis- tributed large language model training and inference,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 171f39d0-57f8-4d81-b5ed-c4e50d2e374b · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3cffd5d7-2ef3-4502-84bf-5bfca4a043d6 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Reducing activation recomputation in large transformer models,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5aed8f0-a8e7-435c-8072-a37046d6fe4c · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Efficient large-scale language model training on gpu clusters using megatron-lm,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2252d4ab-9b9f-4f5d-ab01-f409ad64ac1e · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Efficient memory management for large language model serving with pagedattention,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83d66fd5-4371-4bd0-b743-0b6af6ac3726 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration RelayAttention for Efficient Large Language Model Serving with Long System Prompts
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4db954af-3db0-403d-8890-0f273a857ee4 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Prompt cache: Modular attention reuse for low-latency inference,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ece9a834-b1c9-4f7d-a5cb-c5cd559267ec · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Model pruning enables efficient federated learning on edge devices,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94330cc7-8d05-49a8-8d41-dbe980d5b3e6 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration EDGE-LLM: Enabling Efficient Large Language Model Adaptation on Edge Devices via Layerwise Unified Compression and Adaptive Layer Tuning and Voting
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5614e127-983e-4639-b28f-1dc4facfd617 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Optimizing mobile-edge ai-generated everything (aigx) services by prompt engineering: Fundamental, framework, and case study,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d0856ff-9b02-4b2d-80a0-05f524d75af0 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration CE-CoLLM: Efficient and Adaptive Large Language Models Through Cloud-Edge Collaboration
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e991f7-3307-4c1a-9155-e0b38881e6c1 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Infinite-LLM: Efficient LLM Service for Long Context with DistAttention and Distributed KVCache
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e009fbe8-b573-4466-92ef-600fb12c5026 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Zero: Memory optimizations toward training trillion parameter models,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b6e9b1d-0818-485e-b782-db74ee69e692 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration LLM-Pruner: On the Structural Pruning of Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c59d05b9-46f2-4c8f-9be7-de32966817b3 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5d6b63f-1400-4643-a2f9-df03bc2f9948 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Cachegen: Kv cache compression and streaming for fast large language model serving,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0bdeff-ff8b-439d-9382-c80aa88a7e44 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f5775f5-be36-4c5c-9531-eba11776302d · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Ada-KV: Optimizing KV Cache Eviction by Adaptive Budget Allocation for Efficient LLM Inference
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfcd4484-3cb2-4f0a-8378-cbbd363a2109 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Model tells you what to discard: Adaptive kv cache compression for llms,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d872b2c-9fbb-48cc-aec1-fa9723b252f4 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Netgpt: An ai-native network architecture for provisioning beyond personalized generative services,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a974abcd-187d-46b9-9315-854d348d651f · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Compute Or Load KV Cache? Why Not Both?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d64e7a03-6e3a-4b0d-90ca-73cfd009e03b · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration {OSCA}: An{Online-Model}based cache allocation scheme in cloud block storage systems,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9278466e-630c-4ef2-beb6-d14d133d7670 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Similarity of neural network models: A survey of functional and representational measures,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2972a50a-33da-4f1c-a58f-be2f92ce2e4c · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration ThinK: Thinner Key Cache by Query-Driven Pruning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f91f44f-fe55-4ee4-8225-5ed1061ea45e · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration A Survey of Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2d31dae-3178-4f20-bd3b-9b124471e737 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e72d29dd-9468-4454-bdb9-2060b6317fcc · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Aligning ai with shared human values,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e423a7f1-e62a-48d2-961a-bf5cf2914e95 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c155ec4-fc16-4007-a30a-75c7b896324c · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration After applying an orthogonal trans- formation, it becomes: Oc =O eQ(26) whereQ T Q=I
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d10a037-0e94-40d0-9f0b-61708886f2f2 · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 496173f2-1f8a-4559-8dc3-3e891f1321ef · outbound
CE-LSLM: Efficient Large-Small Language Model Inference and Communication via Cloud-Edge Collaboration Similarity of Neural Network Representations Revisited
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.