Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:31:42.126884Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 6 inbound Pith citation observations for arXiv:2502.07222.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:31:42.126884Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:06:15.845262Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T14:52:41.305916Z
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c43fe9cd-617f-4505-ac56-f4d5740570ed · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Language Models are Few-Shot Learners
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0be1fcc-0bdd-44f7-b901-b882f8008b00 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models LoRA: Low-Rank Adaptation of Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e1eff25-4cd1-4051-8c2d-52f462d793c4 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Chain of LoRA: Efficient Fine-tuning of Language Models via Residual Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 723f37ca-d918-4659-b148-8b7389db98c5 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Subspace Optimization for Large Language Models with Convergence Guarantees
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1dd313c-bc00-4608-967b-bcdeee29589f · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Flora: Low-Rank Adapters Are Secretly Gradient Compressors
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3416b57f-a747-44e1-aa4f-53cd5a854338 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Enhancing Zeroth-order Fine-tuning for Language Models with Low-rank Structures
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ba0fd91-1273-4f46-8c25-048b37646563 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models LoRA Learns Less and Forgets Less
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 962db0d4-9782-4af4-8935-6fcb5d559388 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Memory-Efficient LLM Training with Online Subspace Descent
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7257a49f-7d7f-4d3f-bb85-442a9660787d · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Fira: Can we achieve full-rank training of llms under low-rank constraint? arXiv preprint arXiv:2410.01623, 2024b
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7221865-1152-4da2-a76c-69f946720f23 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models GWT: Scalable Optimizer State Compression for Large Language Model Training
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5512ac83-5cbc-40c7-8f4b-cfb6d32aa5ec · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Second-Order Fine-Tuning without Pain for LLMs:A Hessian Informed Zeroth-Order Optimizer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51eb4ee5-6276-4bf5-93c6-39a0e6e05ef5 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Training Deep Nets with Sublinear Memory Cost
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8999f02c-d660-4eb4-96f3-afa40bc85f4f · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models {Zero-offload}: Democratizing {billion-scale} model training
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b2700321-27a8-44be-bb10-b64b6f316864 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Unified Convergence Analysis for Adaptive Optimization with Moving Average Estimator
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0e2b3cf8-96b4-49c7-9de4-517f0a01e3aa · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc7a6963-77c3-4ea5-9329-d7b9aa275ed5 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models We also report the memory overhead and total training time for each method
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7d117e1a-e632-4350-889c-e33cbc2f503e · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Decoupled Weight Decay Regularization
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dbf2f18-468c-4ae3-9814-34798345d6b9 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models QLoRA: Efficient Finetuning of Quantized LLMs
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81b546d6-cd09-41b1-b43e-8fe64df9ee27 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a964815e-6c11-4934-8371-7c97acad7c5c · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3347d7e8-4a3d-48c4-92ac-90dd0a3a50dd · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0cd6cf5-8f8b-4b11-99ce-5024066cf95e · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models GPT-4 Technical Report
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfea90d6-d746-4ff9-877f-9862585cee59 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models G10: Enabling an efficient unified gpu memory and storage architecture with smart tensor migrations
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe92f3d9-0eb2-4e29-bbda-735c56a10869 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models The Llama 3 Herd of Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c68f13ae-ebba-426c-98f6-6fdfa3823743 · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Adam: A Method for Stochastic Optimization
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 762db2e0-86aa-46b6-9993-bec36b85da3a · outbound
A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models Variance-reduced Zeroth-Order Methods for Fine-Tuning Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 975a7038-cd34-481d-aa55-a59dd83bf1a5 · inbound
FZOO: Fast Zeroth-Order Optimizer for Fine-Tuning Large Language Models towards Adam-Scale Speed A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d48eac71-acc1-4a57-a145-b1334f1674a9 · inbound
CR-Net: Scaling Parameter-Efficient Training with Cross-Layer Low-Rank Structure A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b9b5e486-9ed5-43bd-9601-4a81e8f559c6 · inbound
BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fe18d896-397a-46a8-b4d8-f4cd285c224a · inbound
AdaMeZO: Adam-style Zeroth-Order Optimizer for LLM Fine-tuning Without Maintaining the Moments A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ec5d3761-5ce8-44bf-84cd-c91704fedc2a · inbound
BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 15be2455-b720-4779-9d82-bceece556a22 · inbound
BROS: Bias-Corrected Randomized Subspaces for Memory-Efficient Single-Loop Bilevel Optimization A Memory Efficient Randomized Subspace Optimization Method for Training Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.