Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:19:49.664592Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2501.14205.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T15:19:49.664592Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T18:53:34.664986Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:58:42.766816Z
40 of 40 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d84a032f-a3d7-4da2-bffa-b46f63043aba · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading When Large Language Model Agents Meet 6G Networks: Perception, Grounding, and Alignment
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10c0479d-77f9-42de-b2dd-70ba141c59fe · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Language mod- els are few-shot learners,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba4d54c5-4232-4607-8a4d-db3262dcc5e0 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Sparks of Artificial General Intelligence: Early experiments with GPT-4
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07bcb665-a1fa-4fa8-bcd8-b820ba246c89 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Large Language Model (LLM) for Telecommunications: A Comprehensive Survey on Principles, Key Techniques, and Opportunities
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f4ee1f1-789c-46ed-873a-0ec185a40574 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading The CAP Principle for LLM Serving: A Survey of Long-Context Large Language Model Serving
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 324835ab-36ac-41f0-a772-95bd158246ae · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Galaxy: A Resource-Efficient Collaborative Edge AI System for In-situ Transformer Inference
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c30e738c-4fea-4cca-8947-4a9a28a11baa · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Retention-aware container caching for serverless edge computing,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f5daf523-da96-4009-b6ea-729c2d61fe2e · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Cache- enabled federated learning systems,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be5ebe3d-bac9-4f38-9898-d668c14f884f · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Edgeadaptor: Online configuration adaption, model selection and re- source provisioning for edge dnn inference serving at scale,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bafd9c87-ef0f-48d7-a912-e7242e487e42 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Cooperative service caching and workload scheduling in mobile edge computing,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4f9d75c6-b1d1-42e9-8a1b-d87f9a6764be · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading MemServe: Context Caching for Disaggregated LLM Serving with Elastic Memory Pool
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b926fd4d-4dd1-4781-bf88-329a2c568c07 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading LLM-dCache: Improving Tool-Augmented LLMs with GPT-Driven Localized Data Caching
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c5236eb-bcb1-4540-9a57-1cc26038a826 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Chain-of-thought prompting elicits reasoning in large language models,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 705d8997-317d-498a-9e27-59e210c9deb9 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Self-Consistency Improves Chain of Thought Reasoning in Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c18bf34f-319d-459a-b120-e013e31636f6 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Proximal Policy Optimization Algorithms
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d234bdd-393a-4d3f-971b-85005cbf12ef · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Learning to (Learn at Test Time): RNNs with Expressive Hidden States
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfe2d729-d1ce-43e7-80a6-73929e45fe0a · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Pushing Large Language Models to the 6G Edge: Vision, Challenges, and Opportunities
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d87c17d-e2a4-4fdd-8215-d8abd3b885b2 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading PerLLM: Personalized Inference Scheduling with Edge-Cloud Collaboration for Diverse LLM Services
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 558a7db0-9503-4bee-8958-d1bb52c45b96 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Titanic: Towards production federated learning with large language models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d75add3e-0cf4-4cfb-8a7d-7cfc7c88d351 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Generative inference of large language models in edge computing: An energy efficient approach,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ddf08fb1-3e2e-4b54-96ff-e77a2abb545e · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Two time-scale joint service caching and task offloading for uav-assisted mobile edge computing,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 20ac9efd-3b09-4f68-b21c-f5e6764c2d81 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading TrimCaching: Parameter-sharing Edge Caching for AI Model Downloading
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c27acdd5-8e20-49ec-aa36-8258fa3e4b5e · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading A3c-based computation offloading and service caching in cloud-edge computing networks,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e482c0ba-9f64-42f9-ad97-5883a39ea08f · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Deepcache: A deep learning based framework for content caching,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73400278-8708-4782-9b07-d901433e05c5 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Deep reinforcement learning-based computation offloading and distributed edge service caching for mobile edge computing,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d1b9ca53-b7dd-4cb3-ab6e-296d6691b3ad · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Neighboring- aware caching in heterogeneous edge networks by actor-attention-critic learning,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation febaaaee-ece9-4f37-a2ca-6b25a529ef91 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Cooperative task offloading and service caching for digital twin edge networks: A graph attention multi- agent reinforcement learning approach,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 14ba85ae-987e-4d82-926f-cf3bb749626d · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Large language models (llms) inference offloading and resource allocation in cloud-edge com- puting: An active inference approach,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4b35c3fb-0782-46ce-a0cc-d39fc95bd21c · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Are transformers universal approximators of sequence-to-sequence func- tions?
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 39190684-d285-4a76-9702-86ce715496c3 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Why Can Large Language Models Generate Correct Chain-of-Thoughts?
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1138a40f-56d7-45f9-aaa3-3df80e5f49cb · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading A Latent Space Theory for Emergent Abilities in Large Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7abc262e-0c74-49b5-b4b3-2dbab7d53f38 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Cached Model-as-a-Resource: Provisioning Large Language Model Agents for Edge Intelligence in Space-air-ground Integrated Networks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bff87d8-a937-46a8-afd5-3eac5f51219f · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Imagebind: One embedding space to bind them all,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb3af874-46ee-4c30-8ae5-3dc7885dc6b0 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Solving General Arithmetic Word Problems
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9625d1c3-274b-4f21-b68a-0c4480f8d821 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77168914-eb45-40cd-8636-c0a7241f0c22 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Did aristotle use a laptop? a question answering benchmark with implicit reasoning strategies,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b0c23f5-f14f-4c7a-9454-88602f16b0ba · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading CommonsenseQA: A Question Answering Challenge Targeting Commonsense Knowledge
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81301fb8-5a9e-45f0-b363-1aa1c5868a03 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Adversarial GLUE: A Multi-Task Benchmark for Robustness Evaluation of Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 305919c9-70f4-428d-8177-f427c898aea5 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Improve diverse text generation by self labeling conditional variational auto encoder,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7234443e-7d09-4bb8-b0cb-9494d4685cf6 · outbound
Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading Training Verifiers to Solve Math Word Problems
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0444af8-1be3-4fcc-8942-70a4602b0f22 · inbound
The Price of Anarchy in Disaggregated Inference Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation aae32842-d1be-4de0-846a-af0918afdbb6 · inbound
LMEdge: QoS-Aware LLM Inference Orchestration on Edge Clusters Serving Long-Context LLMs at the Mobile Edge: Test-Time Reinforcement Learning-based Model Caching and Inference Offloading
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.