Pith. sign in

Paper Citation Record · LEDGER

Super Weights in LLMs and the Failure of Selective Training

As of 22 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2607.08733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.08733 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T02:26:06.457776Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0a8abeb-2053-4019-9092-1b42944753ed · outbound

This paper cites Intrinsic dimensionality explains the effectiveness of language model fine-tuning.

Super Weights in LLMs and the Failure of Selective Training Intrinsic dimensionality explains the effectiveness of language model fine-tuning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.143871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ff2e4006adf93ee4deefc8437f65d763acb92c9a370f57000db3dd5760ee1756

Observation a3f4058f-d427-4686-a6f3-52ab0e7e9cbe · outbound

This paper cites Systematic Outliers in Large Language Models.

Super Weights in LLMs and the Failure of Selective Training Systematic Outliers in Large Language Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-10T02:26:42.779156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:3bd481bfddd6990fde03c42eebb399041b7a0b8e1c1c3db1ac060c817e33e718

Observation f2b77621-530d-4b1a-b52f-664713cad9fe · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Super Weights in LLMs and the Failure of Selective Training Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T02:26:42.774803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:0972ae403a4d5bae4217a573b57b8208249f672250db523cd782410e5f0eec91

Observation c6e8e3b4-f2e3-4777-8640-cfe2a4a889b0 · outbound

This paper cites GPT3.int8() : 8-bit matrix multiplication for transformers at scale.

Super Weights in LLMs and the Failure of Selective Training GPT3.int8() : 8-bit matrix multiplication for transformers at scale

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.133504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:88726c89a80dfe8045b251325c547e1879bf797268f15b1c41c6e743ca9b8b88

Observation ff173f1a-fbb0-4a46-922a-5cdd91c0f1cb · outbound

This paper cites [Fou23] Nicolas Fournier.

Super Weights in LLMs and the Failure of Selective Training [Fou23] Nicolas Fournier

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T02:26:42.770708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:1ffe1034d7670723e8c55943c8b12eb6666b436765cb6b43456a02d4290ad2ad

Observation 8226ece8-143b-4959-92ff-4e618bbf3f60 · outbound

This paper cites OLMo : Accelerating the science of language models.

Super Weights in LLMs and the Failure of Selective Training OLMo : Accelerating the science of language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.127977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:41abfb101061219fad6de354a9ffae8080842cbe79dbe9aeb81fb24cc0cdc0a1

Observation 18b1d65e-1866-409d-a29b-39e246318568 · outbound

This paper cites LoRA : Low-rank adaptation of large language models.

Super Weights in LLMs and the Failure of Selective Training LoRA : Low-rank adaptation of large language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.126021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:56006aedf474b8bf50ebe84f0e3e50a732dd5ff1cc3a36da6176fa11cdf24e75

Observation 613a7e7d-62cb-44de-9543-beaa55a1613a · outbound

This paper cites The emergence of essential sparsity in large pre-trained models: The weights that matter.

Super Weights in LLMs and the Failure of Selective Training The emergence of essential sparsity in large pre-trained models: The weights that matter

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.140719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:6c6de484178774dafc19fbbef93a85769e8565695a7dadcbda409743f6cf96cb

Observation 1fa34088-dd81-4eb1-8876-f794b7979979 · outbound

This paper cites On relation-specific neurons in large language models.

Super Weights in LLMs and the Failure of Selective Training On relation-specific neurons in large language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.129809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:990f655e8464cc25ddcb57c6aa5e947768c5b53303037afbce896f945ac15795

Observation 918d72e4-cf13-45be-bd16-c3e2f2cd2c61 · outbound

This paper cites Pointer sentinel mixture models.

Super Weights in LLMs and the Failure of Selective Training Pointer sentinel mixture models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.142202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:821b0827a471779056491227e5fe717ecf4788a4991f8adffa1e090793e16e6b

Observation 0a54edfb-eee5-4f3f-a28c-4de355db50d7 · outbound

This paper cites Morris, Niloofar Mireshghallah, Mark Ibrahim, and Saeed Mahloujifar.

Super Weights in LLMs and the Failure of Selective Training Morris, Niloofar Mireshghallah, Mark Ibrahim, and Saeed Mahloujifar

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T02:26:42.777802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ea9c7ab714e43e7a0eccd5b3205dc29fbaeef28053b8be25545b24f9f2320605

Observation 0793b112-447a-467f-b3c2-81f2030b99a0 · outbound

This paper cites WinoGrande : An adversarial winograd schema challenge at scale.

Super Weights in LLMs and the Failure of Selective Training WinoGrande : An adversarial winograd schema challenge at scale

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.146078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:272a3a4fa10db7812477a3f0f70386f55edcebdede244d3f6a03292ec46f6203

Observation 8270d540-b70e-4ff5-9455-ebf0af5ea9cc · outbound

This paper cites Maximum-margin matrix factorization.

Super Weights in LLMs and the Failure of Selective Training Maximum-margin matrix factorization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.142406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:5d4d93c3453e693833b7726146a50b77e40b0e7f03bdd1d4e72bcce9cdd886da

Observation 21992097-4b64-4a36-960d-b6419e7797d9 · outbound

This paper cites Massive activations in large language models.

Super Weights in LLMs and the Failure of Selective Training Massive activations in large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:a6f6a040cf5a50042d540eca2cb461fe783f7dd3494b138ee7359aeceb447dad

Observation b345d8bb-1f66-427e-8488-41ef38298df2 · outbound

This paper cites The Super Weight in Large Language Models.

Super Weights in LLMs and the Failure of Selective Training The Super Weight in Large Language Models

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T02:26:42.780422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ddd76be3f51051ef7eace345141b38e4638151d90e06ee5ebf8dfe2b188adf39

Observation f8bcc10c-e703-4665-a259-2ce32a4d52c3 · outbound

This paper cites BitFit : Simple parameter-efficient fine-tuning for transformer-based masked language-models.

Super Weights in LLMs and the Failure of Selective Training BitFit : Simple parameter-efficient fine-tuning for transformer-based masked language-models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.137283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:b7d0dfa0663f8b2a4a13f837e329df1f8b09c4fa287be2276d48977c17859188

Observation a812e5ef-4241-49a7-b0d7-6c225253bcd9 · outbound

This paper cites Adaptive budget allocation for parameter-efficient fine-tuning.

Super Weights in LLMs and the Failure of Selective Training Adaptive budget allocation for parameter-efficient fine-tuning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.138983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:2d86f93b26e158a9a76d1d82d084729c14a29870c355e1fa8d8ad94197735331

Pith citing papers

No inbound Pith citation observations are available.