Pith. sign in

Paper Citation Record · LEDGER

Inverse Scaling: When Bigger Isn't Better

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2306.09479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.09479 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:55:42.806627Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T18:51:18.126997Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8ed30a97-a648-4d9a-b493-5215ca33d997 · inbound

AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents cites this paper.

AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents Inverse Scaling: When Bigger Isn't Better

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:35:13.477819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T06:35:13.331872Z digest=sha256:0c8c4d2ea6c103c068b26b9a3dda37255563423f2e3ff3ebc89cd572b27d2775

Observation 7a7409d8-78fb-465c-a08a-030eafdd4c16 · inbound

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation cites this paper.

Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation Inverse Scaling: When Bigger Isn't Better

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:24:13.032084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T07:24:12.845841Z digest=sha256:dbc3e2cbff2c7112c41d9f878c4646f8530ee71c5fefba17c3bef61f9024dc0a

Observation 7c1d1965-1218-410f-ab29-39d92aea046d · inbound

Towards Efficient and Effective Alignment of Large Language Models cites this paper.

Towards Efficient and Effective Alignment of Large Language Models Inverse Scaling: When Bigger Isn't Better

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:42.806627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:42.806627Z digest=sha256:aa5940ec5cc6a77d6edc349d73efa1e3e5fbe6dc20413bd82f0c152b00d37991

Observation b084aec8-6fee-4087-a879-fb1e266d68de · inbound

Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench cites this paper.

Which LLMs Get the Joke? Probing Non-STEM Reasoning Abilities with HumorBench Inverse Scaling: When Bigger Isn't Better

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:48:08.300856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:48:08.300856Z digest=sha256:4750684b1d92d12eb97eac97ffe5d7d530259419e693a1bfeca5cf45e787acf8

Observation c2d957b2-d75c-4872-b1ad-3a6ae4c1b2ad · inbound

A Survey on Data Security in Large Language Models cites this paper.

A Survey on Data Security in Large Language Models Inverse Scaling: When Bigger Isn't Better

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T05:05:02.822438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:05:02.822438Z digest=sha256:b10418afe04941a6d5814662376c43b3c77efb088fab03a54609623dc9c92267

Observation 7b3d9530-2f85-41c7-a555-20bda3ccf042 · inbound

On the Fitness Landscape in the $NK$ Model cites this paper.

On the Fitness Landscape in the $NK$ Model Inverse Scaling: When Bigger Isn't Better

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-05T19:28:59.286934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:28:59.286934Z digest=sha256:08308ac2e29e7b45eac7947d621ea72bd9ffc40447f6d5096a09a1040ac389b6

Observation e118af3c-f90f-4c3e-8708-0e9d9d41a799 · inbound

MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes cites this paper.

MoReBench: Evaluating Procedural and Pluralistic Moral Reasoning in Language Models, More than Outcomes Inverse Scaling: When Bigger Isn't Better

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T09:16:40.238706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:16:40.238706Z digest=sha256:bccc4c4639d3600b88256c4f2ca3ea831c5cfcd0f6ed42c456df56dbbd1886ef

Observation 3cdbfb9a-6550-4619-ab3c-e460597eccd6 · inbound

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings cites this paper.

Evaluating Code Reasoning Abilities of Large Language Models Under Real-World Settings Inverse Scaling: When Bigger Isn't Better

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:28:34.185656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T21:23:44.762007Z digest=sha256:f71bf1f01a4ab1c01860402650d3353ea5c45219fefa3e0baaa327af30910360

Observation d9887284-e5bd-42de-8a56-1dc3992d6388 · inbound

Complexity Horizons of Compressed Models in Analog Circuit Analysis cites this paper.

Complexity Horizons of Compressed Models in Analog Circuit Analysis Inverse Scaling: When Bigger Isn't Better

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:26:10.413487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T16:37:05.972866Z digest=sha256:c9fdba391929fd6517390bd422a1b9bc3d62c4a7055c46e7d635970d8de8b8b5

Observation 81f35394-32d5-44e0-9558-e39a8e346078 · inbound

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities cites this paper.

Absurd World: A Simple Yet Powerful Method to Absurdify the Real-world for Probing LLM Reasoning Capabilities Inverse Scaling: When Bigger Isn't Better

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:46:26.123113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T02:02:06.788056Z digest=sha256:c6fe06af3d792e01ba4e0de2421e9ec9ae4fef9921b1f908cd4da211bcf96cee

Observation cc9d31cb-c5e8-4bc1-b320-521a5899b68d · inbound

History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions cites this paper.

History Anchors: How Prior Behavior Steers LLM Decisions Toward Unsafe Actions Inverse Scaling: When Bigger Isn't Better

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:57:33.653876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-14T17:53:08.128110Z digest=sha256:6d3785e01eab584e5fabbd8f0582133342fbd93861d9e3ae751ac3d726370c7e

Observation 722c0ed2-3983-484c-8307-d99a164d6d25 · inbound

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning cites this paper.

GraphARC: A Comprehensive Benchmark for Graph-Based Abstract Reasoning Inverse Scaling: When Bigger Isn't Better

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-28T22:42:46.635573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T22:39:04.557903Z digest=sha256:0a4c8d65c8eecd2e1feda7fa91538bd41c0b274939c2bf5fb8465bd35a143348

Observation 285602ae-c361-4927-8a19-b4be547df4b2 · inbound

Resonant Context Anchoring: Decoupling Attention Routing and Signal Gain at Inference Time cites this paper.

Resonant Context Anchoring: Decoupling Attention Routing and Signal Gain at Inference Time Inverse Scaling: When Bigger Isn't Better

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:26:21.635105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T14:27:00.523125Z digest=sha256:fb1d7574a69c20c111b379545af50a8dfcd6c8ab4a4f4f4ab520907bbac63a6e

Observation 4ea3e0dd-15ba-4319-8d8e-fb0524bc0f33 · inbound

The Variance Brain Foundation Models Forgot: Third-Order Statistics Predict Cognition Where Billion-Parameter Models Fail cites this paper.

The Variance Brain Foundation Models Forgot: Third-Order Statistics Predict Cognition Where Billion-Parameter Models Fail Inverse Scaling: When Bigger Isn't Better

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:12:34.916457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T19:07:34.240009Z digest=sha256:31e9c29d5d854c872240f2137e2eeb183ff7af8c1b94a5b551b4e355fac51eb1

Observation 9e846436-773d-4e02-8d6e-4d1f41169cdc · inbound

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers cites this paper.

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers Inverse Scaling: When Bigger Isn't Better

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:56:46.962340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T06:36:29.222547Z digest=sha256:db67d98aaec3bd686bfcc0ee9154f42d22039b81363352b17294e5bae3d94029

Observation e7789daa-1c20-42c6-973f-e77661ecd46c · inbound

The Granularity Gap: A Multi-Dimensional Longitudinal Audit of Sycophancy in Gemini Models cites this paper.

The Granularity Gap: A Multi-Dimensional Longitudinal Audit of Sycophancy in Gemini Models Inverse Scaling: When Bigger Isn't Better

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-05T18:51:18.129206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-05T18:43:10.949515Z digest=sha256:05337dfb5fe4713ce13f3800860a6b9538f94324e5345d8e5749504a106079fa

Observation 5f684e49-b9bc-4789-b72e-862839dbc61a · inbound

LEVANTE-bench: Multi-Scale Comparison of VLMs to Children Using Cognitive Tasks (or, "Is Your VLM Smarter Than a 5th Grader?") cites this paper.

LEVANTE-bench: Multi-Scale Comparison of VLMs to Children Using Cognitive Tasks (or, "Is Your VLM Smarter Than a 5th Grader?") Inverse Scaling: When Bigger Isn't Better

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:46.455279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-28T06:39:08.246174Z digest=sha256:917d29aaf410d7fd1dba7f891058480c2fcb893db12fa0fea606ccdc36ee401d

Observation 73950eb4-b306-43a1-829c-e24b58f8457f · inbound

From Verdict to Process: Agentic Reinforcement Learning for Multi-Stage Fact Verification cites this paper.

From Verdict to Process: Agentic Reinforcement Learning for Multi-Stage Fact Verification Inverse Scaling: When Bigger Isn't Better

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:58:33.400385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T06:47:09.681690Z digest=sha256:8623b15245785706fd1b4dc34293570db2b3345fe8d5d8e01b50feaa3c10ed49

Observation 951592dc-74d9-4a91-8109-72cfe45d5af7 · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Inverse Scaling: When Bigger Isn't Better

Reference 77

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T19:50:11.027188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-25T20:58:53.119386Z digest=sha256:3b203c80399383c3c8901f3688ed336ab63fd08315223d28e0c7c4f381b9a05d

Observation 4f328d89-fadd-4e69-9177-df7dd0ca40be · inbound

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation cites this paper.

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation Inverse Scaling: When Bigger Isn't Better

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T10:16:44.591107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:16:44.591107Z digest=sha256:fbe0ce2dda2b2a5c169d77c68e6a0f91ab59a52f3e581b42b54726d1a9d01b28

Observation c1efe193-7930-406e-886c-22de33ccdf41 · inbound

EntroRouter: Learning Efficient Model Routing via Entropy Regulation cites this paper.

EntroRouter: Learning Efficient Model Routing via Entropy Regulation Inverse Scaling: When Bigger Isn't Better

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:44:21.602491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-30T07:42:35.962864Z digest=sha256:aed88157c6a0fddd655f2d73a93d27926c25f054bd1e8302939e33653e6ffb8d

Observation ab677952-96c7-4650-9ac1-aa5ff5369776 · inbound

Spaghetti Architect: A Contamination-Resistant, By-Construction-Labelled, Multi-Language Code Dataset Generator cites this paper.

Spaghetti Architect: A Contamination-Resistant, By-Construction-Labelled, Multi-Language Code Dataset Generator Inverse Scaling: When Bigger Isn't Better

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T14:50:59.137848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T14:50:59.137848Z digest=sha256:e511f5875d5e3d1615b0fbbc54717e786e08b3a1179bd9f70cf6c509adb622f1

Observation 7bb0ed5f-2ac5-4089-be19-0124e4760598 · inbound

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation cites this paper.

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation Inverse Scaling: When Bigger Isn't Better

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T04:28:47.303456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T04:28:47.303456Z digest=sha256:30954439ee12eff11e4d737bcdc216c2368693ce78e8d07404a0a1cabd87e5f8