Pith. sign in

Paper Citation Record · LEDGER

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization

As of 8 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2603.00910.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.00910 v3

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T02:39:03.519913Z

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1ff4ae0-1b69-4642-866a-75b8bb08bf62 · outbound

This paper cites Table 3: Hyperparameter configurations for allocation and pruning.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Table 3: Hyperparameter configurations for allocation and pruning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:03.519913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:03.519913Z digest=sha256:a18c2bc9db1db01a711d11186576f905355cabb2a76958c93a167126f549b1de

Observation 53ca9af1-1210-48f7-a8fe-8abdf552cd19 · outbound

This paper cites doi: 10.18653/v1/N19-1300.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.18653/v1/N19-1300

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.651246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.651246Z digest=sha256:37ac9003e6c3fd90aaa9a5c378c29395925e34ae6d97ed9b2e9a5fe30e3b022f

Observation 5edf0fa6-4ec4-49ff-8a22-4a1ef6cb491d · outbound

This paper cites doi: 10.18653/v1/2023.emnlp-main.907.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.18653/v1/2023.emnlp-main.907

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.139671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.139671Z digest=sha256:bb8b92f3b431d58439adb29e1c8fbfdb9f3b07de5f5ce0614759687b335cc268

Observation b5c9ff05-d477-4c0f-9434-fa3f8c06b6bb · outbound

This paper cites Mistral 7B.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Mistral 7B

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.270504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.270504Z digest=sha256:fc9606ae0d5c1c3e9487802d3214187dcc9e3b581b30349d0639834104d52f70

Observation aa0312ea-3f23-4095-a576-29c69ce08c65 · outbound

This paper cites Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Bridging Predictive Coding and MDL: A Two-Part Code Framework for Deep Learning

Reference 14

Resolution
malformed identifier
no resolver link, observed 2026-08-03T02:39:01.840773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.840773Z digest=sha256:dbf365463144b3dee2528a1721bfac7333894363e428d66930e4b6af4f4efb9f

Observation 7cfc5459-c9d0-4868-a2fd-766ee5d4818d · outbound

This paper cites AlphaLoRA: Assigning LoRA Experts Based on Layer Training Quality.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization AlphaLoRA: Assigning LoRA Experts Based on Layer Training Quality

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.949150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.949150Z digest=sha256:c015257e02777e6fcac8e7620033286e7230e66c7c2b63a86ca6128e0304e3b4

Observation bdb869fc-5010-4523-a8d0-051550d02818 · outbound

This paper cites Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.050833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.050833Z digest=sha256:13fd043609f4d7ecf225887592120cac8a069cb2e778e3cb2721f1ee3f88a4d9

Observation 548ad9e8-4bd1-409e-843b-c249518b5ee7 · outbound

This paper cites doi: 10.1145/3474381.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.1145/3474381

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.200285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.200285Z digest=sha256:f26a041c4abba65aa130729d388e13914b976b53b26d255f31acbb2b153ce36e

Observation 006238de-8937-4c5b-ba18-17ce5d0b8b74 · outbound

This paper cites Layer Importance and Hallucination Analysis in Large Language Models via Enhanced Activation Variance-Sparsity.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Layer Importance and Hallucination Analysis in Large Language Models via Enhanced Activation Variance-Sparsity

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.457506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.457506Z digest=sha256:a5c62302b1f776650b46872e4ee29ada69fdb0da8f4731c910507ff8d1eee691

Observation 04ecd127-62fe-46cc-bd4e-67a02e1a84b7 · outbound

This paper cites CommonsenseQA: A question answer- ing challenge targeting commonsense knowledge.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization CommonsenseQA: A question answer- ing challenge targeting commonsense knowledge

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.615387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.615387Z digest=sha256:930543aa703f2fa3b3681fc016cc6e34da0c1f47383aca504179d20b8c26a6c3

Observation 61b64283-719d-4d6d-a15e-845866b36ade · outbound

This paper cites doi: 10.18653/v1/N19-1421.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.18653/v1/N19-1421

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.750852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.750852Z digest=sha256:b415f20c562e0200c961b9c6cab1769fbf383df0263a81473bc4f8995b240ef6

Observation b7cbcc05-75b8-4ca4-b866-736564a0d9c5 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Gemma: Open Models Based on Gemini Research and Technology

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.915391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.915391Z digest=sha256:541d4bca5dfbc35c1d2d2206d43f236ebfaf50d5e048bfb251b5abea959e1e32

Observation 5cbd3c02-d0fa-4a3c-af6c-9a7a6779838a · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:03.042786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:03.042786Z digest=sha256:f49a3de6f572deea73d4f5d4714397dc7dc612b97241a92273e3ee56715b17b0

Observation 018bff00-7817-4319-a0aa-350dfc2d866c · outbound

This paper cites doi: 10.1162/tacl_a_00290.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.1162/tacl_a_00290

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:03.120885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:03.120885Z digest=sha256:13e047903fc2f979d5439f970464742eef1e1ab5afd15e214d5c6d442532fdfe

Observation 8cfaee59-c966-4d99-91d6-f5d29d31f941 · outbound

This paper cites More recent work scales these ideas to modern architectures using diagonal Fisher approximations Martens and Grosse [2015], Kronecker-factored curvature Botev et al.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization More recent work scales these ideas to modern architectures using diagonal Fisher approximations Martens and Grosse [2015], Kronecker-factored curvature Botev et al

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:03.286836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:03.286836Z digest=sha256:01795e0c181e7b3b4e78fb6c8d7fdf6c7bb5b1c78e45dd633637d7fa77e535bc

Observation caf4ed65-5e1b-49a7-a0e2-dba18a90f07d · outbound

This paper cites an unresolved cited work.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:03.390339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:03.390339Z digest=sha256:8c45a64d1fd90d2ca9d9d2cb1583751c9ec5f40a53df92757854ab2bed071cbe

Observation e7287afc-f3cf-405f-9dd6-d01d4d3df87e · outbound

This paper cites cc/paper_files/paper/1989/file/6c9882bbac1c7093bd25041881277658-Paper.pdf.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization cc/paper_files/paper/1989/file/6c9882bbac1c7093bd25041881277658-Paper.pdf

Reference 1989

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.439077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.439077Z digest=sha256:09b7d97bb77bae9dbd83c38b932f637e2e7bf5796f6b89cadfac27a3b9dd95e3

Observation 6113e4ce-45a5-4d2f-9d74-437b9389c6a0 · outbound

This paper cites Shwai He, Run-Ze Fan, Liang Ding, Li Shen, Tianyi Zhou, and Dacheng Tao.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Shwai He, Run-Ze Fan, Liang Ding, Li Shen, Tianyi Zhou, and Dacheng Tao

Reference 1993

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.057577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.057577Z digest=sha256:6630aa258268abce70252840c399051e28a73a10aa2eb39a840d8c83bb01decb

Observation 443bc2a7-5b09-4bb9-baa4-7a94750210b0 · outbound

This paper cites doi: https://doi.org/10.1016/S0893-6080(96)00127-X.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: https://doi.org/10.1016/S0893-6080(96)00127-X

Reference 1997

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:02.294054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:02.294054Z digest=sha256:c6a4e44b0ff574e3d48fc063361a8b06a6627e4f39d450d651e66a76142dbed3

Observation f4975cf5-b060-440b-8fe2-92ceb3130fe7 · outbound

This paper cites Christopher Clark, Kenton Lee, Ming-Wei Chang, Tom Kwiatkowski, Michael Collins, and Kristina Toutanova.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Christopher Clark, Kenton Lee, Ming-Wei Chang, Tom Kwiatkowski, Michael Collins, and Kristina Toutanova

Reference 2004

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.490098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.490098Z digest=sha256:c568b7eeb1442ef43ae192b28bc00216ea4aa996b7d303ef38ec7a9bc040c004

Observation a6d9dc6e-4a6f-4915-ba90-51031b37d6a7 · outbound

This paper cites Can a suit of armor conduct electricity? a new dataset for open book question answering.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Can a suit of armor conduct electricity? a new dataset for open book question answering

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.710424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.710424Z digest=sha256:0f0af9e11abd6c32994c04a60205ba6255db1e2b651a4462fefcc1ab8d31077a

Observation 3e1e642b-b84f-4050-8b14-43b97d8e860e · outbound

This paper cites DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.340481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.340481Z digest=sha256:2de246a942898e12161f44b6c10b700d34e2800ea3ffea252a969b09403565ef

Observation 4a89ba97-e1e9-4bf2-8c53-4cd99afa3262 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.762257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.762257Z digest=sha256:736f7b5f4111b53e13ca9d76942903375730d28ab930cee18d764b42e95a1636

Observation 42b50bf9-f184-4f8b-af8d-6aca4e3ae54d · outbound

This paper cites doi: 10.1016/j.neunet.2019.04.009.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization doi: 10.1016/j.neunet.2019.04.009

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.406059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.406059Z digest=sha256:f3dc6627da5fa43c103f204b995b596ab29e7049ef8d1d541a4e4a0c2ca2514c

Observation 89a30109-6188-4d5c-a112-7b43b945e5b6 · outbound

This paper cites Sanae Lotfi, Marc Finzi, Yilun Kuang, Tim G.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Sanae Lotfi, Marc Finzi, Yilun Kuang, Tim G

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:01.580653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:01.580653Z digest=sha256:a14abfc26891fd845189153c2985c6684c34c6b6284985936880dd7451437037

Observation 36a94c27-d56c-428d-8eec-757579577cc1 · outbound

This paper cites SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.860624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.860624Z digest=sha256:d7aef382b32e707c2103e5e588fe00589ae3e7850d63988009cbec206f26b61d

Observation 36cc4a16-0943-46c5-8820-9362d59c7c8a · outbound

This paper cites Higher Layers Need More LoRA Experts.

Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization Higher Layers Need More LoRA Experts

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T02:39:00.957520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:39:00.957520Z digest=sha256:d4d581638f0474957c346bbb314f908580b01ddd652dab9345d66d674affad92

Pith citing papers

No inbound Pith citation observations are available.