Pith. sign in

Paper Citation Record · LEDGER

Super Weights in LLMs and the Failure of Selective Training

As of 9 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2607.08733.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.08733 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T02:26:06.457776Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact2
  • verified fuzzy12
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a0a8abeb-2053-4019-9092-1b42944753ed · outbound

This paper cites Intrinsic dimensionality explains the effectiveness of language model fine-tuning.

Super Weights in LLMs and the Failure of Selective Training Intrinsic dimensionality explains the effectiveness of language model fine-tuning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.143871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:3df2a07fbebf36b896b6f53858791db609924260708b499e0933277f9b2ea203

Observation a3f4058f-d427-4686-a6f3-52ab0e7e9cbe · outbound

This paper cites Systematic Outliers in Large Language Models.

Super Weights in LLMs and the Failure of Selective Training Systematic Outliers in Large Language Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-10T02:26:42.779156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:f61adc0e2a8ad5f26873dbbf9665aee632c979082be9f210af1d24eb3e99c919

Observation f2b77621-530d-4b1a-b52f-664713cad9fe · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

Super Weights in LLMs and the Failure of Selective Training Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T02:26:42.774803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:57d4f247435ed3d874191512324f432ba76c3307b11dcafa604e7f2b7c5c470c

Observation c6e8e3b4-f2e3-4777-8640-cfe2a4a889b0 · outbound

This paper cites GPT3.int8() : 8-bit matrix multiplication for transformers at scale.

Super Weights in LLMs and the Failure of Selective Training GPT3.int8() : 8-bit matrix multiplication for transformers at scale

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.133504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:71eb14195c2ccc18ca92facbfbcae39f315717c7f1cdd0ab19689fd2735ab230

Observation ff173f1a-fbb0-4a46-922a-5cdd91c0f1cb · outbound

This paper cites [Fou23] Nicolas Fournier.

Super Weights in LLMs and the Failure of Selective Training [Fou23] Nicolas Fournier

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T02:26:42.770708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:232b91dc2b84ec1a4f5c42b990ba88a9513003d79ade0be1ab48be44b2bda472

Observation 8226ece8-143b-4959-92ff-4e618bbf3f60 · outbound

This paper cites OLMo : Accelerating the science of language models.

Super Weights in LLMs and the Failure of Selective Training OLMo : Accelerating the science of language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.127977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:5069102e25e6a5dd9e20b4a504530fd01797d7a946f050effb2175650823dd1b

Observation 18b1d65e-1866-409d-a29b-39e246318568 · outbound

This paper cites LoRA : Low-rank adaptation of large language models.

Super Weights in LLMs and the Failure of Selective Training LoRA : Low-rank adaptation of large language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.126021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ebad283c92421e0a70cd8348f7aa21b5310612405b473f60e78f3ffaf63acc05

Observation 613a7e7d-62cb-44de-9543-beaa55a1613a · outbound

This paper cites The emergence of essential sparsity in large pre-trained models: The weights that matter.

Super Weights in LLMs and the Failure of Selective Training The emergence of essential sparsity in large pre-trained models: The weights that matter

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.140719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ef9e6adb889de89bff1f49d5023870918ff0812aacdac022981428040b350ae2

Observation 1fa34088-dd81-4eb1-8876-f794b7979979 · outbound

This paper cites On relation-specific neurons in large language models.

Super Weights in LLMs and the Failure of Selective Training On relation-specific neurons in large language models

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.129809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:0988546e52b892bad3e9e01cf5b58f21bcb9a2745d9c34b92a7e04e2853ae407

Observation 918d72e4-cf13-45be-bd16-c3e2f2cd2c61 · outbound

This paper cites Pointer sentinel mixture models.

Super Weights in LLMs and the Failure of Selective Training Pointer sentinel mixture models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.142202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:529208edf68498af5ae9998c7a55aa0ca9fef8f6b7a9e63a8d2cba2ff823ee45

Observation 0a54edfb-eee5-4f3f-a28c-4de355db50d7 · outbound

This paper cites Morris, Niloofar Mireshghallah, Mark Ibrahim, and Saeed Mahloujifar.

Super Weights in LLMs and the Failure of Selective Training Morris, Niloofar Mireshghallah, Mark Ibrahim, and Saeed Mahloujifar

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-10T02:26:42.777802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:833f233026d341d4c971360402f6584c2a509bfe667e33a9a750a01f383b5e88

Observation 0793b112-447a-467f-b3c2-81f2030b99a0 · outbound

This paper cites WinoGrande : An adversarial winograd schema challenge at scale.

Super Weights in LLMs and the Failure of Selective Training WinoGrande : An adversarial winograd schema challenge at scale

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.146078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:6f346e9ddef1e547b15553f142678ece95d6927989739894011a7305fee3c54e

Observation 8270d540-b70e-4ff5-9455-ebf0af5ea9cc · outbound

This paper cites Maximum-margin matrix factorization.

Super Weights in LLMs and the Failure of Selective Training Maximum-margin matrix factorization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.142406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:d9a44fc41d58e17e46172cda12970a27e495bbbbcd4c04d47bf96b729f8df76a

Observation 21992097-4b64-4a36-960d-b6419e7797d9 · outbound

This paper cites Massive activations in large language models.

Super Weights in LLMs and the Failure of Selective Training Massive activations in large language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:c00896b9006b2282fa3a138e94d00afe31f63706972d9e5e9ef48abde44ddec3

Observation b345d8bb-1f66-427e-8488-41ef38298df2 · outbound

This paper cites The Super Weight in Large Language Models.

Super Weights in LLMs and the Failure of Selective Training The Super Weight in Large Language Models

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T02:26:42.780422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:f81e7d818b816003a8470c80db9567df062b494832693947fdd839f693165467

Observation f8bcc10c-e703-4665-a259-2ce32a4d52c3 · outbound

This paper cites BitFit : Simple parameter-efficient fine-tuning for transformer-based masked language-models.

Super Weights in LLMs and the Failure of Selective Training BitFit : Simple parameter-efficient fine-tuning for transformer-based masked language-models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.137283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:ba12c453c7d5128a985feed4fd05251e4056e584e3d9550a9d3f60651a93e06f

Observation a812e5ef-4241-49a7-b0d7-6c225253bcd9 · outbound

This paper cites Adaptive budget allocation for parameter-efficient fine-tuning.

Super Weights in LLMs and the Failure of Selective Training Adaptive budget allocation for parameter-efficient fine-tuning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T02:26:43.138983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-10T02:26:06.457776Z digest=sha256:6daed004eb7113cebe9ce47e264a2fd73bf48d4e489a11b418efae5b758ee6f0

Pith citing papers

No inbound Pith citation observations are available.