Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T10:33:00.445749Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 1 inbound Pith citation observation for arXiv:2604.16400.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T10:33:00.445749Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T12:16:10.904456Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-12T18:15:00.917874Z
46 of 46 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 775ea295-3671-4c55-a06b-644f12684d2b · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters (2023) Github copilot
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9192094e-7af7-40e9-a642-b62d659bdfcd · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ed653408-5948-4088-9d6f-0c33e71ca280 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters (2022) Chatgpt: Optimizing language models for dialogue
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 65a5ec62-2864-4bc9-b10d-8c05786e360f · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters A review on edge large language models: Design, execution, and applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f2113891-e3dd-47f7-8a76-d74edbc4d43e · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Mobile edge intelligence for large language models: A contemporary survey
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f6bbf74a-4cbd-4060-bf8d-106b9bad304c · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Exploring parameter-efficient fine-tuning to enable foundation models in feder- ated learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 96d713be-7495-44c1-8ddb-92f6f4c5a614 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Heterogeneous LoRA for federated fine-tuning of on-device foundation models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e1195fcc-cfbb-4e19-969c-390963878770 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Federated fine-tuning for pre-trained foundation models over wireless networks
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2ddfaebd-12b1-497d-8eda-e1f240aff3bf · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters LoRA: Low-rank adaptation of large language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c3cbe103-acca-4c20-a093-1928c8fbbe31 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Beyond scale: the diversity coefficient as a data quality metric demonstrates llms are pre-trained on formally diverse data
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 49bcd51f-07f6-4fe6-8c07-ffcd311d5c5a · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters DeepBoot: Dynamic Scheduling System for Training and Inference Deep Learning Tasks in GPU Cluster
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3ddc8af0-704b-4d46-997c-109aed6be3f6 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Multiplexing dynamic deep learning workloads with slo-awareness in gpu clusters
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4b8318f0-0267-4140-b2ee-3a1106ac983d · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Scaling Laws for Neural Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation db49b4a8-191e-4a31-b0ac-7b94630a49fe · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Lyra: Elastic scheduling for deep learning clusters
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2c5c4f4c-e305-4ee6-9ee7-b5077943bad7 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Serving hetero- geneous machine learning models on Multi-GPU servers with Spatio- Temporal sharing
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2a03124e-a738-473a-b4ff-8092765c21fd · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters (2023) Multi-process service (mps)
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5e589751-e464-4d9e-87e2-0527db0ee71a · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Shepherd : Serving DNNs in the Wild
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 77f0ca7c-c4e4-40c6-bed2-4fa9fa9740ae · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Federated Learning while Pro- viding Model as a Service: Joint Training and Inference Optimization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cba83386-c1de-4b08-9c7d-0c460e8cc593 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Communication-efficient learning of deep networks from decentralized data
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5b5e2263-c77a-47ff-91f4-4d3f828cf7f8 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Fedadapt: Adaptive offloading for iot devices in federated learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 074ad771-90d2-4000-8bd0-d8228d83cb25 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Qlora: Efficient finetuning of quantized llms
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e3827efd-b301-4c5a-b170-b681ef926f0d · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters FedPara: Low-Rank Hadamard Product for Communication-Efficient Federated Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 993e34c6-12b1-4303-a6ee-487bc8fb2f64 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Few-Shot Parameter-Efficient Fine-Tuning is Better and Cheaper than In-Context Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4e16538c-c1d6-4348-9893-df8789664171 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Don’t decay the learning rate, increase the batch size
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 856899ea-8e0c-4e58-9267-f87205099c00 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Efficient Coordination of Federated Learning and Inference Offloading at the Edge: A Proactive Optimization Paradigm
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1c6d0ce5-703b-49f2-bc55-8b14ca1d3a10 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Human-in-the-loop machine learning: a state of the art
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6392a0aa-4de5-4cb6-8a33-36c724ff6297 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Illustrating reinforcement learning from human feedback (rlhf)
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5c17ec63-b5f9-4c24-931e-c6dacb33ca8d · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters manim code
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ff13a1f7-48aa-4196-b4d1-a180755a37da · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Codealpaca-20k
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation ffbc9b68-049d-469a-849c-4b136ccf2b18 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters code instructions 120k alpaca
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2ea815c5-59d5-4377-9774-f45126575679 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 376dbce9-94fa-438d-a184-177b33a9420d · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Gpteacher-general-instruct
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cd0fcd64-4f94-4791-859e-c6fc646776e7 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters open-instruct-v1
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7ba89d86-4fc0-48c2-a7bd-96e8b5223ab0 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters URL https://doi.org/10.1109/HPCA61900.2025.00102
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 691d4f20-924d-4d59-b161-6d5e23210a1a · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters dLoRA: Dynamically Orchestrating Requests and Adapters for LoRA LLM Serving
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6bb1b16b-5f8c-4308-aec2-d01ccc6cf215 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Peft: State-of-the-art parameter-efficient fine-tuning methods
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9deabcba-3c8a-4423-9614-3d4b59ffd3a6 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Accelerating End-Cloud Collaborative Inference via Near Bubble-free Pipeline Opti- mization
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b9b026c3-cf21-40d7-9e7c-92d9739bbaf2 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Adaptive parameter-efficient federated fine-tuning on heterogeneous devices
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0ba0af59-037b-42e1-a304-f5b9b825924d · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Haflq: Heterogeneous adaptive federated lora fine-tuned llm with quantization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7e3457b9-2896-4ac6-94b7-9f907d69f673 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters HAFLQ: Heterogeneous Adaptive Federated LoRA Fine-tuned LLM with Quantization
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dce4c978-3b52-49dd-ad44-1d2bd507a384 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters FwdLLM: Efficient Feder- ated Finetuning of Large Language Models with Perturbed Inferences
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eda66662-e4db-46cd-b83d-02ad304d67c3 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Partitioned collaborative inference for on-device models via evolution- ary reinforcement learning
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a0263514-3ee5-4ba5-8939-f3e0a02b3bae · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 91d9592c-0a08-4e0b-bf58-a671bfb206fe · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Online Resource Allocation for Edge Intelligence with Colocated Model Retraining and Inference
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e09f437-8138-45aa-9a9b-5e8af549157f · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters LLMStation: Resource Multiplexing in Tuning and Serving Large Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5fd616b3-b39c-44b0-af8c-264fbae54aa0 · outbound
CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters Flexllm: Token-level co-serving of llm inference and finetuning with slo guarantees
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 20c7d213-03f7-42e7-b8d9-1be3412ad707 · inbound
Not Every Sync Is Safe: Calibrated DiLoCo Scheduling for Shared AI Infrastructure CoLLM: Continuous Adaptation for SLO-Aware LLM Serving on Shared GPU Clusters
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.