Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T13:39:53.206470Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2605.28573.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-29T13:39:53.206470Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 803ca0db-dfff-419d-a183-a4fb9dd8a4d4 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Princeton University Press, 2008
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb999300-a5e4-4e24-8c8d-c2217c04322f · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers GQA: Training generalized multi-query transformer models from multi-head checkpoints
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b47e68-0876-4c50-be0d-b28ea703478d · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unitary evolution recurrent neural networks, 2016
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc58cfc9-086b-4e95-a96e-02da620e13d7 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Can we gain more from orthogonality regularizations in training deep cnns?, 2018
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33d73fa0-0af7-48be-a9c3-3c6ae314c3eb · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Cambridge University Press, 2023
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba6164c2-f849-4d16-aa4f-a427d06134b3 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6ca715-2088-4ab3-a9d6-8df46e5a36c8 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 611fa5f3-172d-4acb-a452-19547ff7e3f0 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Reducing overfitting in deep networks by decorrelating representations, 2016
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d66f8d9-5628-406e-a03b-55ab48ecf332 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers IOS Press, October 2024
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3cefa28-6997-4de1-979c-4220b5cc62b4 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Documenting large webtext corpora: A case study on the colossal clean crawled corpus, 2021
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2551fa4-5e2b-4dd5-8c8a-9dd6b84f3f37 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers The Llama 3 Herd of Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71be59f2-fb3a-44fb-8de5-47c194ac368b · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers The approximation of one matrix by another of lower rank
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4888d7ca-8ad8-431e-938a-937d7a7f5710 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Arias, and Steven T
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcd53eef-4ceb-4e08-9487-bd379e41b4e2 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Full-rank no more: Low-rank weight training for modern speech recognition models, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813dba24-c7c6-4b33-8aeb-6b60d7b52e69 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Deep feedforward networks.Deep learning, 1:161–217, 2016
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46786ee7-0e73-4c67-aad8-4297ab996ab7 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers From gpt to llama: Tracing the growth of large language models.Theoretical and Natural Science, 142:144–155, 11 2025
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9514802-bbac-4a32-9529-7845407417e0 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Sltrain: a sparse plus low-rank approach for parameter and memory efficient pretraining, 2024
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8de2ee8c-5454-41af-a9e5-0f048087185b · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Rae, Oriol Vinyals, and Laurent Sifre
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0da1ecbc-3590-4db9-b3ba-f30a79bdd00a · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 880813a3-cbb5-4af9-a3a7-62de6314e3d9 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71870e05-89b7-4237-a8a1-79d0fc10ab72 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9df9d7a1-5758-4e84-9db4-2f7221b41631 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Initialization and regular- ization of factorized neural layers
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c958fc24-c032-4584-bc87-a468cc2d1af2 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Initialization and regular- ization of factorized neural layers, 2022
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56d8b64a-c1f2-4d7a-8ec5-512369ef81fc · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Lost: Low-rank and sparse pre-training for large language models, 2025
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea4f1aee-409b-404d-b0c6-989b44cf5229 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Relora: High- rank training through low-rank updates, 2023
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb52c38-215f-4810-af53-aa8909c67714 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Cola: Compute-efficient pre-training of llms via low-rank activation, 2025
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 583e94d8-c514-4a0b-8f1d-d813c34878f7 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Large language models: A survey, 2025
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22064a8e-f43f-43f4-a44b-3d622b4420e8 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Parameter and memory efficient pretraining via low-rank riemannian optimization
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73748c87-332c-4023-8d3f-228568de0478 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Olmo 3
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 05b2e6af-98d9-4ff0-9121-53322f2be3f7 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 151ff332-23a2-4de5-96af-322218e385bc · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Robust low-rank training via approximate orthonormal constraints, 2023
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b38f402-96f7-4b3e-96f6-1da574865883 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Low-rank lottery tickets: finding efficient low-rank neural networks via matrix differential equations, 2022
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7d54399-7dba-4ffe-b6c2-0192317e9479 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Compact: Compressed activations for memory-efficient llm training, 2025
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63391c11-29bf-44e0-a151-124d2792ff23 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Cuttlefish: Low-Rank Model Training without All the Tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 60b0ab53-a955-47f1-85d7-e9ed03adf229 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Dynamic rank adjustment for accurate and efficient neural network training.arXiv preprint arXiv:2508.08625, 2025
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2abeaca3-03fd-43da-b0ce-a732d94e8a36 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Elrt: Efficient low-rank training for compact convolutional neural networks, 2024
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 790e8e70-9ee6-484f-bc15-3ead2327b3be · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Llama: Open and efficient foundation language models, 2023
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23a9d67c-b4c8-44ae-af57-8144981f721b · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f155ceb-2721-4cd2-a698-aceedf5e7b84 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c4922ffc-7840-41ad-bfa1-9a3c4f178240 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Investigating low-rank training in transformer language models: Efficiency and scaling analysis, 2024
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8ba27e5-d6b7-4ddb-8ddd-7ee8a7e031ca · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Transformers: State-of-the-art natural language processing
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5342ad7d-e7ae-4b99-be1f-0251853b6497 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Learning low-rank deep neural networks via singular vector orthogonality regularization and singular value sparsification, 2020
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea414b5f-f6cc-46e0-9436-fbf34691dc6f · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers InRank: Incremental Low-Rank Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e1e5c323-c1c1-42e6-8c8d-901080d194ed · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Galore: Memory-efficient llm training by gradient low-rank projection, 2024
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56fa5527-1306-4c1f-a86c-3bbd767502ba · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers A survey of large language models, 2026
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2649f54a-9539-4b9a-bdf8-7d6532208ee3 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Hence the forward operator norm is controlled exactly byσ max
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fac4b32b-a6ed-4524-b5c2-c69b443f8280 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Hence the backward operator norm is controlled by the same quantity
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80cd6486-57d4-4de3-aeb7-5e3cb76b4c21 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Likewise, ifδlies in the represented output subspacespan(U), then σmin√r ∥δ∥2 ≤ ∥W ⊤δ∥2 ≤ σmax√r ∥δ∥2
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55bc3c9b-b5f1-4372-b960-4aad0885c27a · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers If ui and vi denote the i-th columns of UandV, then W= 1√r rX i=1 σiuiv⊤ i , ∂L ∂σi = 1√r u⊤ i Gvi
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eea9d66c-12e0-4a25-8437-458ef6f9f0e5 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bd55ed0-fcae-4ba4-8119-e0d6361c1273 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 205b66b8-5fff-461d-8e1e-577004d1e6f5 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Orthonormality does not remove this fundamental low-rank bottleneck, but it does prevent additional instability caused by badly scaled basis factors
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e642d319-2108-4ce2-bbd2-e2cbca30cb20 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dab857e3-bf4c-43e2-a22c-c7b94da8d6c4 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbc9a5c2-7d97-4b38-b1c8-f6dd55d208a8 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a62f5f2-1ac0-42e2-8e5b-16376d31b118 · outbound
Efficient Pre-Training of LLMs through Truncated SVD Layers Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.