Pith. sign in

Paper Citation Record · LEDGER

SLOT: Sample-specific Language Model Optimization at Test-time

As of 16 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 6 inbound Pith citation observations for arXiv:2505.12392.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.12392 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:38:17.037405Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T16:07:30.821966Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved35
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 986b9a51-ba5a-442c-b056-75302ff2b13e · outbound

This paper cites The Surprising Effectiveness of Test-Time Training for Few-Shot Learning.

SLOT: Sample-specific Language Model Optimization at Test-time The Surprising Effectiveness of Test-Time Training for Few-Shot Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.862667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.862667Z digest=sha256:6dac02a0e90439ab1be2e5f55544884825ce875b5c174a4fdb3908089cac7491

Observation 7573b81f-ef2f-46f0-b8d6-928fc77b82b8 · outbound

This paper cites Qwen Technical Report.

SLOT: Sample-specific Language Model Optimization at Test-time Qwen Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.868055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.868055Z digest=sha256:d3cecec784c1bde84dd34ec7bb8b5d0f15e6ad8dc3e7594b75bce941987740dd

Observation 38bd0204-7d45-41d6-a50e-f718c2400647 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

SLOT: Sample-specific Language Model Optimization at Test-time Evaluating Large Language Models Trained on Code

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.872974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.872974Z digest=sha256:9a010e3b8808ea5f25992e2553332b446eeb583366eba15f4cb422fda6a66f3d

Observation c6cc15dd-cc75-419f-8a0a-0ca998016cba · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

SLOT: Sample-specific Language Model Optimization at Test-time Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.877288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.877288Z digest=sha256:6db4f7471548e13c8c98867cfd9d0641f5f25feba9389f6cb5dbaacd52e6a0e1

Observation 73273f8b-7451-4c14-bdc0-6fc91be149bf · outbound

This paper cites Learning How Hard to Think: Input-Adaptive Allocation of LM Computation.

SLOT: Sample-specific Language Model Optimization at Test-time Learning How Hard to Think: Input-Adaptive Allocation of LM Computation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.881661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.881661Z digest=sha256:1dbff984d8bd7b16306335234a7fcb830d020025f267ba2f66689d0a9c81d0da

Observation c3eef4c6-626b-4496-bec7-3d797728c6bc · outbound

This paper cites Test-time training can close the natural distribution shift performance gap in deep learning based compressed sensing.

SLOT: Sample-specific Language Model Optimization at Test-time Test-time training can close the natural distribution shift performance gap in deep learning based compressed sensing

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.510693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:16.885975Z digest=sha256:41f7ae94b86cd448b7c10e7cf3a42986af9f6060bcb8504a6815a3131ff5ddae

Observation c13331ff-6287-4d59-8acc-3f7f42a7a7ec · outbound

This paper cites Test-time training with masked autoencoders.Advances in Neural Information Processing Systems, 35:29374–29385, 2022.

SLOT: Sample-specific Language Model Optimization at Test-time Test-time training with masked autoencoders.Advances in Neural Information Processing Systems, 35:29374–29385, 2022

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.890188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.890188Z digest=sha256:04ec588afac43b19fea7d6ac83d62096680cd0e6925f9717a70750eef10244c5

Observation de5708aa-45a5-40a7-87d8-b6d6520f7581 · outbound

This paper cites The Llama 3 Herd of Models.

SLOT: Sample-specific Language Model Optimization at Test-time The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.894534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.894534Z digest=sha256:eb667512bcc0844b5a5462afa5399ab7b22b5d1de52c8951d1c1f38b567bcff2

Observation d08dbc8e-943e-4908-a970-5bbd07da5d4c · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SLOT: Sample-specific Language Model Optimization at Test-time DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.898650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.898650Z digest=sha256:045fb4202352cb0e306d2da49ef418e9ebdf3c37d9ad46ba9b33f0605936c568

Observation 04b713be-57c2-414f-bd5c-c2d525023b2a · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

SLOT: Sample-specific Language Model Optimization at Test-time Measuring Mathematical Problem Solving With the MATH Dataset

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.903127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.903127Z digest=sha256:2a9855ebbc506ae2118825125436817c1b9e8451d8acc37313578453f8f18bf6

Observation 050b98c3-b7a7-4a8b-bada-11457ac3ce3e · outbound

This paper cites Parameter-efficient transfer learning for nlp.

SLOT: Sample-specific Language Model Optimization at Test-time Parameter-efficient transfer learning for nlp

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.907495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.907495Z digest=sha256:71934d159fc7be10324716c5353a4c5267f516526dcd12b55bd88ce23aba4590

Observation 341ac7c5-d098-4b84-b1aa-7a438fe5725d · outbound

This paper cites Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022.

SLOT: Sample-specific Language Model Optimization at Test-time Lora: Low-rank adaptation of large language models.ICLR, 1(2):3, 2022

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.911973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.911973Z digest=sha256:ba24502adf37dce9eb7057e9c02f4e5558dd980ad787eb1953f2c157736b9d66

Observation d3980604-bdb2-4681-9c7d-0dc3fdbac35e · outbound

This paper cites C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models.Advances in Neural Information Processing Systems, 36:62991–63010, 2023.

SLOT: Sample-specific Language Model Optimization at Test-time C-eval: A multi-level multi-discipline chinese evaluation suite for foundation models.Advances in Neural Information Processing Systems, 36:62991–63010, 2023

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.916224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.916224Z digest=sha256:ed19475023d3d4ab2153a4ae725fdd87112fcec8ba58af8eb58830219cda55bb

Observation e7d67827-d582-48df-b1f9-7d6f4ef54b2c · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

SLOT: Sample-specific Language Model Optimization at Test-time A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.920727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.920727Z digest=sha256:8d9ea91fc12995cd3d8bd04908faa83bcdd92341020ce557249f99853b1fb5a1

Observation fe6deb32-b237-471c-a812-1e017f76f086 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

SLOT: Sample-specific Language Model Optimization at Test-time Gonzalez, Hao Zhang, and Ion Stoica

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.466594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:16.925003Z digest=sha256:02dac70991730ecb82c5a723454966cc7f07c16349abbe9f37844b7b62dd76fc

Observation 52d4ba75-e817-4553-ad86-78f8dd9a3b45 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

SLOT: Sample-specific Language Model Optimization at Test-time The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.929184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.929184Z digest=sha256:687d378830f3719dc040297f48ba67080ccae1b435cd6f9b2fcda3a0ad7bd9d6

Observation 8a7e0f14-ebde-48ff-be69-7d0b1c2def29 · outbound

This paper cites Prefix-Tuning: Optimizing Continuous Prompts for Generation.

SLOT: Sample-specific Language Model Optimization at Test-time Prefix-Tuning: Optimizing Continuous Prompts for Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.933344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.933344Z digest=sha256:71158e4fd70ee0fd74eef32789f267509656dd79c20fbe9c501f15a80ba84014

Observation 02772e5a-6f2f-4b64-95a9-e3ef28d7b969 · outbound

This paper cites A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks.

SLOT: Sample-specific Language Model Optimization at Test-time A Survey on LLM Test-Time Compute via Search: Tasks, LLM Profiling, Search Algorithms, and Relevant Frameworks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.937355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.937355Z digest=sha256:c6cabd34a2c8423d637155c33fdf1e9139b1d0562deab60941cdc32cb84b1819

Observation 7ca01e10-82d6-4c5b-bc2c-171388286475 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

SLOT: Sample-specific Language Model Optimization at Test-time Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.941819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.941819Z digest=sha256:eb9ff00573db9e4b8863e5f710be6c65a2d66843c667cf041801e98bed944aab

Observation e84261a0-1bfe-46af-906c-ab656e6e042d · outbound

This paper cites P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks.

SLOT: Sample-specific Language Model Optimization at Test-time P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.946128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.946128Z digest=sha256:35f080a9a46d55c1de5c1b11bd45575581adeb2102113606286f7f1cf0a0094f

Observation feedad1f-a017-45ea-b62b-ffe4e02c4fe2 · outbound

This paper cites Ttt++: When does self-supervised test-time training fail or thrive?Advances in Neural Information Processing Systems, 34:21808–21820, 2021.

SLOT: Sample-specific Language Model Optimization at Test-time Ttt++: When does self-supervised test-time training fail or thrive?Advances in Neural Information Processing Systems, 34:21808–21820, 2021

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.950336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.950336Z digest=sha256:142cf5b9bdcce53942a68c64cd4cf25461ce2c0b2d63c4206bdcfce7f90f72d4

Observation 3bdb711e-f310-4e16-8474-8d909e164ffe · outbound

This paper cites Decoupled Weight Decay Regularization.

SLOT: Sample-specific Language Model Optimization at Test-time Decoupled Weight Decay Regularization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.954211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.954211Z digest=sha256:a4e1f43e7990957ff914019122e9591d7d47938412c9e91d720643f0fc6256fa

Observation 33c4cd93-3ea7-403c-8d5c-fc70557544d2 · outbound

This paper cites Language Models are Few-Shot Learners.

SLOT: Sample-specific Language Model Optimization at Test-time Language Models are Few-Shot Learners

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.958199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.958199Z digest=sha256:b8acec623dc2a0e6790675a5770b0d8be34d869b5ccae44aa62b9cf0e3a1e77f

Observation e752f4b3-7b39-4de7-b8c4-3264b7c11ae7 · outbound

This paper cites Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation.

SLOT: Sample-specific Language Model Optimization at Test-time Adaptive Inference-Time Compute: LLMs Can Predict if They Can Do Better, Even Mid-Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.962058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.962058Z digest=sha256:35bdcd9495b846aea943900a55b1e37e35db4b2356bb96dee3027af68f5137f4

Observation c6303c76-ab54-481f-a432-3ae39a21d958 · outbound

This paper cites Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?.

SLOT: Sample-specific Language Model Optimization at Test-time Rethinking the Role of Demonstrations: What Makes In-Context Learning Work?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.965914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.965914Z digest=sha256:70f93454a9ceafd74b6b8059ab4a7122598195a667df14679f8bcfe8c805881f

Observation 09d81cfd-dc9a-43bd-baa1-17059074609d · outbound

This paper cites Improving Black-box Robustness with In-Context Rewriting.

SLOT: Sample-specific Language Model Optimization at Test-time Improving Black-box Robustness with In-Context Rewriting

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.970305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.970305Z digest=sha256:c6ea0daa3c543ea17983f5ec2102ac1c08566cad17b027bde5328c0497420659

Observation 088ed9d4-eccb-4cf5-b1f0-49a5e60867f7 · outbound

This paper cites Tttflow: Unsupervised test-time training with normalizing flow.

SLOT: Sample-specific Language Model Optimization at Test-time Tttflow: Unsupervised test-time training with normalizing flow

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.447277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:16.974689Z digest=sha256:04167092d6936e40431b0054393083edf1ea6ee123a3081ef51a4a8102c594db

Observation 4653fa01-6f0f-4a58-89f0-457f52d14f08 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

SLOT: Sample-specific Language Model Optimization at Test-time Gpqa: A graduate-level google-proof q&a benchmark

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.978638Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.978638Z digest=sha256:873d3d654a2a7d7b070236e2e025fe32db872452a7551ff9fa1ec3582d9e69ae

Observation 60f8ee74-291c-4823-9140-bfc16300c2c4 · outbound

This paper cites Towards real-world test-time adaptation: Tri-net self-training with balanced normalization.

SLOT: Sample-specific Language Model Optimization at Test-time Towards real-world test-time adaptation: Tri-net self-training with balanced normalization

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.427328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:16.982354Z digest=sha256:4d268cdb43c56d252cf6b121ce1b2756695438673ce49ed8b043c88d2d8f67e8

Observation 434269cb-6a58-4f86-ae91-c49652ab7c27 · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

SLOT: Sample-specific Language Model Optimization at Test-time Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.986205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.986205Z digest=sha256:30dbb7554b8c87054ca13817e02a2be96cce349a552fdf0547b49a1602f580de

Observation 5dd43cf7-598d-4a59-8e54-d04d7a48a18f · outbound

This paper cites Learning to (Learn at Test Time).

SLOT: Sample-specific Language Model Optimization at Test-time Learning to (Learn at Test Time)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.990116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.990116Z digest=sha256:67299c20a8ac59ca96ce1ab742f3b47a2b67db59453c06a0cb4b912135beec36

Observation 2985d54d-e242-4666-bfa3-b984451a513b · outbound

This paper cites Test- time training with self-supervision for generalization under distribution shifts.

SLOT: Sample-specific Language Model Optimization at Test-time Test- time training with self-supervision for generalization under distribution shifts

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.413863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:16.994116Z digest=sha256:7858401feb946a3f4a00bb21952845a77b9454ae30120b500d3c1070d748a65c

Observation 4045d2b4-f3dc-4d6e-a4cf-40acf894e090 · outbound

This paper cites Tent: Fully Test-time Adaptation by Entropy Minimization.

SLOT: Sample-specific Language Model Optimization at Test-time Tent: Fully Test-time Adaptation by Entropy Minimization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:16.998243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:16.998243Z digest=sha256:4f3b269b74d5c0d577d5bccd7e741150d9085121d8f18a5599d1c43ab07814ca

Observation fbf0a411-64df-460b-acae-f9d75d19d280 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

SLOT: Sample-specific Language Model Optimization at Test-time Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.002426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.002426Z digest=sha256:4ba5b7d5e9b3e21d0f6e235e1e4f92feeee79609a27d24de0a24238a1deba539

Observation 4407432e-bcf3-4586-93c4-dda52f9fc7bd · outbound

This paper cites Beyond Model Adaptation at Test Time: A Survey.

SLOT: Sample-specific Language Model Optimization at Test-time Beyond Model Adaptation at Test Time: A Survey

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.006489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.006489Z digest=sha256:42a8905b1ba67d04d264a34d1a93759acbb0f4ae252ec36a7e4eb7c048c67949

Observation f7738cbe-275b-4833-83e5-26a53ddc0209 · outbound

This paper cites Stta: enhanced text classification via selective test-time augmentation.PeerJ Computer Science, 9:e1757, 2023.

SLOT: Sample-specific Language Model Optimization at Test-time Stta: enhanced text classification via selective test-time augmentation.PeerJ Computer Science, 9:e1757, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:38:17.399769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-15T20:38:17.011778Z digest=sha256:50e4f41fb27176948817ad4cfe4349f3d05bc993938e9abc5142b2afd03eba19

Observation fc4348f3-88b0-4e4d-886f-8fc7fc2a73ca · outbound

This paper cites Qwen2.5 Technical Report.

SLOT: Sample-specific Language Model Optimization at Test-time Qwen2.5 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.016069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.016069Z digest=sha256:31f34723aaa997ad106aeb3239c2349bd105eb1ee34b57a080003e3f77adc956

Observation e3b0f746-0e29-4d0a-9b41-948f4c0e9bd4 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

SLOT: Sample-specific Language Model Optimization at Test-time Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.020163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.020163Z digest=sha256:ca791279bd49503b9b52a690722a1c0867b145e87b5bb43ecd4e91ae41b701a8

Observation a45991d6-62c8-4263-b08a-3f53a23d45be · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.Ad- vances in neural information processing systems, 36:11809–11822, 2023.

SLOT: Sample-specific Language Model Optimization at Test-time Tree of thoughts: Deliberate problem solving with large language models.Ad- vances in neural information processing systems, 36:11809–11822, 2023

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.024323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.024323Z digest=sha256:b13b163602e4199db342a4b681efe8a2692a3fd96a96751e3a61b248b012cdb1

Observation 0c2265a9-7a01-4b92-a256-86071c0e2f5f · outbound

This paper cites Benchmarking Reasoning Robustness in Large Language Models.

SLOT: Sample-specific Language Model Optimization at Test-time Benchmarking Reasoning Robustness in Large Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.028835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.028835Z digest=sha256:7fe55ebb3fb7c8fd9b5af5a2220d8c3a577d77a1abefef372653dddc9c529b4b

Observation d4d6e143-130f-4034-b844-af8d5a89411e · outbound

This paper cites On Pitfalls of Test-Time Adaptation.

SLOT: Sample-specific Language Model Optimization at Test-time On Pitfalls of Test-Time Adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:38:17.033229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.033229Z digest=sha256:9becb827ee90e9aa90a9f293d872b5a13f6a7e379dce3e400253a85d7c550bd6

Observation 3bad336c-8c0b-465e-88e6-291adcbcccc7 · outbound

This paper cites TTRL: Test-Time Reinforcement Learning.

SLOT: Sample-specific Language Model Optimization at Test-time TTRL: Test-Time Reinforcement Learning

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-08-15T20:38:17.037405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:38:17.037405Z digest=sha256:bfba372ff3c232788ee557d2580769323f3d7a7041806a4f9cd0c84ea6e07aba

Pith citing papers

Observation 8fb8a73f-287e-4ea6-a7d3-064619ab26af · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle SLOT: Sample-specific Language Model Optimization at Test-time

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:30.821966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:30.821966Z digest=sha256:a770c954464de88f18cff184d78a0b699d5212ca6012aa75f37f9c1a3a96c976

Observation eac35b3f-19ea-4b2f-b0b4-0324713cec3c · inbound

Self-Reflective Generation at Test Time cites this paper.

Self-Reflective Generation at Test Time SLOT: Sample-specific Language Model Optimization at Test-time

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T12:41:42.525028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:41:42.525028Z digest=sha256:beb2a12727a01fb19d00e64c1eb5621d7bdf836f66aa2b9d5c0029a0ba969ef0

Observation b3c2337d-b6b2-40a8-8e46-7a20d768bf08 · inbound

From Exposure to Internalization: Dual-Stream Calibration for In-context Clinical Reasoning cites this paper.

From Exposure to Internalization: Dual-Stream Calibration for In-context Clinical Reasoning SLOT: Sample-specific Language Model Optimization at Test-time

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:10:50.521245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T19:19:15.053725Z digest=sha256:a8fbce3f2d6fc2a13c631cf700f4399cf866157128364dc0e9bfad6f8d3c2265

Observation d68e740f-b228-4c32-bb1b-63e1ccb1f9e1 · inbound

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models cites this paper.

Rethinking Entropy Minimization in Test-Time Adaptation for Autoregressive Models SLOT: Sample-specific Language Model Optimization at Test-time

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:51:14.022947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-12T00:48:47.213681Z digest=sha256:c0ce198cb5e21504bba7dfa73426e928331cbae366af2b092d9ea43f30fcf4d7

Observation dc5440c8-50dc-43f8-9f09-d48e2557360d · inbound

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation cites this paper.

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation SLOT: Sample-specific Language Model Optimization at Test-time

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T11:24:08.699295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-21T11:21:30.867480Z digest=sha256:da67bc0f898c5cfac14cf50718a564308e1fea80b1e4300811b45af3a99a16f4

Observation e802226d-8ba5-45de-8ab2-7db931ba651c · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning SLOT: Sample-specific Language Model Optimization at Test-time

Reference 78

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:38:55.966851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:f8d4ac36451de33be1ba43b980bbb0748225029431df6a6e25eed159debaa283