Pith. sign in

Paper Citation Record · LEDGER

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2608.03545.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03545 v2

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:52:37.405393Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact0
  • verified fuzzy42
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ed278b8-71eb-4459-9c40-e8e4f961a4ae · outbound

This paper cites an unresolved cited work.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:52:38.363913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.125672Z digest=sha256:09ebc2e7da3e2e211801de03e862c92b14087f94b97cc7c03f7e750fe08a3709

Observation 35873716-33c1-4ee7-a5d7-687ed403528d · outbound

This paper cites , year = 1983, title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , year = 1983, title =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.349715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.131388Z digest=sha256:a4db7b389a85d806106249297fd149d91353952a58938cf0f16a6065ff7efab4

Observation 185ea821-dcd0-4fa2-a70b-1c8587f17197 · outbound

This paper cites , year = 1984, title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , year = 1984, title =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.335520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.138291Z digest=sha256:68469c9b0b72452e104a5d1fac6e6397e5a677cfe94a87a11efc13d17c81923b

Observation b3d29bff-8b19-4eab-85af-8753c59bde22 · outbound

This paper cites , title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , title =

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.143686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.143686Z digest=sha256:18b15650f14bec77f389212ddd346af79ca5cbb449ad82eba477f62b70e32f34

Observation 3608bf1c-1911-4653-84ac-93be67822f81 · outbound

This paper cites , year = 1980, title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , year = 1980, title =

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.311285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.149323Z digest=sha256:e7aa2d46dd1977f613069b41e46d8607157e6d255f0698cb04554d98d3b3b355

Observation 7fe45bb2-cba3-4838-b941-ae90ab071f1d · outbound

This paper cites Clancey and Glenn Rennels , abstract =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Clancey and Glenn Rennels , abstract =

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.154451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.154451Z digest=sha256:b8acd082984d3671c4a15eabf0e2d5047422614d1a866def0f8e3befbe0990f5

Observation 0ac35a06-d4bf-4e1d-9431-1f99db176d23 · outbound

This paper cites and Rennels, Glenn R.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning and Rennels, Glenn R

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.296976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.160028Z digest=sha256:90820b93645467029391897ab284c54286199d69b86d451ff034e747010c9ea3

Observation cfc2ccef-ffbd-4879-bdb2-229b356bac7a · outbound

This paper cites an unresolved cited work.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T00:52:38.282460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.164832Z digest=sha256:415d7de835f033ed107539b95b191be0189249785527229b1c21745039242c2b

Observation 6f41fcd9-91d7-4a90-85ab-5a1651d6b9e8 · outbound

This paper cites , year = 1979, title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , year = 1979, title =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.268248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.169591Z digest=sha256:4b42421276a76f4d709bb8c5acf5d74d15665eb974f1266e28ce5a20b5bf69f9

Observation 8423e12b-d49c-4ca6-8743-803e28c5f2ff · outbound

This paper cites , title =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning , title =

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.253850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.174492Z digest=sha256:3318330cc5dd7e3ed5dd981fea83052696bd6ed085d8d924bf1fc3cceec368f6

Observation fef0cc84-d6bb-489c-bc68-e029c50a9b71 · outbound

This paper cites 2017 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2017 , eprint =

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.179241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.179241Z digest=sha256:ef52be1879e20096ee06bbf9c09ef9406a19d8ff5a7646af43eb41536764fee2

Observation f57f9d02-2e29-4399-a2a4-e07407b6d9df · outbound

This paper cites an unresolved cited work.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.183822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.183822Z digest=sha256:aa45d23dbae539a092a284b2f5db4e53b46e7a19b633cf66bbde95421fc7c391

Observation 7230285f-38ad-475e-9636-945b7b7b6387 · outbound

This paper cites The Eleventh International Conference on Learning Representations , year =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning The Eleventh International Conference on Learning Representations , year =

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.188675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.188675Z digest=sha256:c20c08646dac20a0a191ba9334602f9a647ccfe449096f3cdadfe00cf9c25f27

Observation a192e1c2-29e3-421f-ba77-5439c7648c1b · outbound

This paper cites Advances in neural information processing systems , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Advances in neural information processing systems , volume =

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.193820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.193820Z digest=sha256:369147a2d54c22e8c6588f90cb1070703064c82b27d8be4ac283039b21703c98

Observation 28f925f4-298b-49af-bc80-cafc8694a046 · outbound

This paper cites TTRL: Test-Time Reinforcement Learning , url =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning TTRL: Test-Time Reinforcement Learning , url =

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.202867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.199002Z digest=sha256:ef3fe78c9d28ff16f4036a120901d40a52c764733aabba41591c3a5393de0fb6

Observation 7a9c20f4-41d6-4a78-8f6c-cc472feeaa1c · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.188738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.204174Z digest=sha256:44afed72feb5ecada7c5cd4f165e23e75f558f2d9e1c672397e1c02d859691c8

Observation 7f668440-b608-4435-8fa4-1ee9e9964a04 · outbound

This paper cites 2025 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2025 , eprint =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.174641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.209028Z digest=sha256:8a17c266db4b7d54e4c237e735580e1e97c1c4b18d35f1e6e0eddc55c7b262be

Observation a27ffac8-fbe4-4c6c-930f-d5d2e403c0b3 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.159714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.214354Z digest=sha256:44ad587f39225442abc5561d2b7265254167a5b690302d3b4d88971f55c56a36

Observation adbaa938-1bcf-47fc-8526-8307834ead2d · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) , pages =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.144454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.218963Z digest=sha256:fde810b16ef3030efa7580067cb3fe0e18eb3e1175a1d85c0ba15f6d19973d98

Observation 20e4baa6-33a1-486a-bb8e-f7f86972d041 · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.130677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.224796Z digest=sha256:87e845e60da3f746e02655709e1215532dea95527a89fc4dec38a9a6bac9f062

Observation 79e68d40-b81d-499d-968b-c94afc4010e3 · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.116799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.229542Z digest=sha256:9c4051888c003fc80b158658e62bbf7ecb9530b90885eb9109a58b20dd4e1cd3

Observation aad71683-7505-4632-bdae-ade1e461c9f4 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.103194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.234366Z digest=sha256:4dfdb750d2fd250e69c4580b2c4f3be781706eee3341f90581237ec7d8ff4f7b

Observation 1718470a-a437-4913-a8fe-0b4324b23d6a · outbound

This paper cites Forty-third International Conference on Machine Learning , year =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Forty-third International Conference on Machine Learning , year =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.089662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.239627Z digest=sha256:5de0fef96476219b1dcdb9e74984fb6c8920b27e63a35936c7f47b3c8efac43d

Observation 3c107409-0448-475b-91d3-245a5a6e1c69 · outbound

This paper cites DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning DeepSeek-R1 incentivizes reasoning in LLMs through reinforcement learning , volume =

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.244689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.244689Z digest=sha256:6ad5a6e177ee706b85cd12283e45c5c926f0f870a0cfc9af6cde7e4aca53a83b

Observation 72850efe-d844-4d56-b9bf-396ea6340ca9 · outbound

This paper cites 2025 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2025 , eprint =

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.249652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.249652Z digest=sha256:7a58a1501cca6841c2d6ff0d9e0454269b1506fd062c8eb4a2d658ef26d2bd00

Observation 7d739666-e02b-4c69-96ea-050634c320bc · outbound

This paper cites 2024 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2024 , eprint =

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.254564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.254564Z digest=sha256:807a28c9be1be17520476edc3fa55c0f4fcda2b0a496802f5d6959c2390d2ddc

Observation fb731896-e67e-4e71-bd9b-a1385c1ed121 · outbound

This paper cites ACM Computing Surveys , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning ACM Computing Surveys , volume =

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.058301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.259146Z digest=sha256:6ed9ba2818ec863440e0265def97eefb31834bd29a49240265e45a63b72bfb58

Observation a2328cc5-d3fb-41e8-bc74-bead1257befe · outbound

This paper cites Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing , pages =

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.043693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.264159Z digest=sha256:da9bb705f7641e1f7831dd8aa456d3c7ddf74d97e416e482d9fbb53d6a9c3251

Observation b5eb781d-193a-4253-a3b9-c151273900a7 · outbound

This paper cites 2025 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2025 , eprint =

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.029749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.268766Z digest=sha256:457c08c73b84438bf28767c751ea8ba6e3b19dcb45669de46a33951a0681f2b3

Observation 11539c81-0316-421a-b678-9166245911af · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.015451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.273663Z digest=sha256:9b776c5d4a9b51cc5983e78e04220a664d9666cafa35eb410e4039773e38ab1a

Observation f788e82d-6b0d-4c88-9236-6177cfd148ea · outbound

This paper cites Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization , url =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization , url =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:38.000775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.278216Z digest=sha256:8b4cebeeb7b7617f362b94496165c82192f47508ee8a06cd916388cf429aec74

Observation 813a7dde-7224-4df4-b079-24cc74824a68 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.985176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.282965Z digest=sha256:67e74f4078ccbff5d9d27f5db572551acc1f313a0c8c092d5a1d47e6bedd93c1

Observation 19522351-5d6b-4f5a-8ea7-4483c3b21224 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.970491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.287476Z digest=sha256:0e16c1a6ceafa92c46b14bb55f648f7623681c18cf3cc840760b3485a77dc621

Observation d29273ac-bbb1-4f0f-a13b-caf93ea42e2f · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , number =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the AAAI Conference on Artificial Intelligence , number =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.955910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.292516Z digest=sha256:3c20266011b0c7e65447435c1b0daaa03142bfaa8551acfc39db6fd8e131346a

Observation 613632e7-a01d-4dda-a218-95bcc6f1caae · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.940948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.297300Z digest=sha256:658df9fbbe7f965c5ea47fd02d016a71ccbb82e2344e8cf1b014b1cda3aa4166

Observation f8802d1c-dc37-4058-8c6c-ac8382597cb9 · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.926536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.301938Z digest=sha256:18b0408d65ab0229cd82208079408ab89ca89de2ff59c8222e73a36bbad456b5

Observation d9c7f3dc-d4a5-49c1-91cc-cad23708d9e2 · outbound

This paper cites 2017 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2017 , eprint =

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.306383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.306383Z digest=sha256:8ed9507ed11e50725d0ebb8b0658e2944f6338efb63e2cc0f8c85a1ae38f4e41

Observation ae6db8fd-e887-4b58-822f-50e922298525 · outbound

This paper cites Advances in neural information processing systems , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Advances in neural information processing systems , volume =

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.311151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.311151Z digest=sha256:90c8a54ab3173132b4eca7ee7f1ede836a0d7fa091538eed0acf213916ae6071

Observation 6f09fccf-7350-4f5d-b300-9bb198b0639c · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Advances in Neural Information Processing Systems , volume =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.894932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.315756Z digest=sha256:c34f74abb6094cc4fa684a1d4695f36f98af565a8d24202654a2932d3756463d

Observation 7b77432a-8fd5-407a-99e6-c7719486bfa8 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.881012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.320316Z digest=sha256:29eb601c3768431aad3e5b92f1aba0a7f7d6e7b5402e6e08026c54bf229719d7

Observation 685415de-be4c-495f-b9df-191984cf0c50 · outbound

This paper cites 2024 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2024 , eprint =

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.866652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.324949Z digest=sha256:6af060955a86db11368448bb28c3df68721bf607efbb637c8b54cf5be33f4383

Observation c9123113-4c72-4338-bad5-3c1160a8ae38 · outbound

This paper cites Advances in Neural Information Processing Systems , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Advances in Neural Information Processing Systems , volume =

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.852473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.330636Z digest=sha256:f0e925f27616ad20c5e25f7aef2b6ca70139700cb8cc267c224f29b14ea30040

Observation c83c0caf-7d57-436f-a0d7-282ad478b756 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.838805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.335195Z digest=sha256:cd26f9d747ed6044162cbd0db9f27e6761764366fe8e2f6e09199bee632906c1

Observation 206a6009-f473-414a-b89e-7b37689ea0f4 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example , url =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Reinforcement Learning for Reasoning in Large Language Models with One Training Example , url =

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.824996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.340347Z digest=sha256:7f72b8db7aa260b04753025ae0172c5604d43571b41d9c3c301605af320e1ebc

Observation b5cfdda8-f7b2-4197-a10b-5932bf58afdc · outbound

This paper cites Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning , url =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning , url =

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.811006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.345104Z digest=sha256:1e0e7a31274f69335637e658bfb914446b4f8e0ac3c47ffd4033c1aee60c1ec8

Observation 9f72fd16-6654-4ccb-8c74-5f5c6a2c1814 · outbound

This paper cites 2026 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2026 , eprint =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.796660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.350015Z digest=sha256:901d2388724339d065a06efb1709a138537c39c1c56494ce6106097b9446fcc7

Observation c4f1566f-16b2-421e-82dd-9063565dd468 · outbound

This paper cites 2025 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2025 , eprint =

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.781622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.354797Z digest=sha256:0db06fb8150ba8fcab56b54826a9ee73fe9febe67c95bc86d36bd65125fea6ec

Observation 4c46b567-af8e-4ca3-9634-9754c596e643 · outbound

This paper cites 2024 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2024 , eprint =

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.767694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.359492Z digest=sha256:7f79fcfaf1fddaf14ceda5e68ad5ac9e8aced5ca3d2a2a727b4bdce6be2b014f

Observation 1579bc71-7677-4a15-aa62-b16ae9304b4d · outbound

This paper cites 2021 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2021 , eprint =

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.753867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.363908Z digest=sha256:a985d6d8dcd1e79302d76149aebbe2cb00c3d249ad1522b8633cba8eed37ea99

Observation 19720dca-b1e2-48e0-9c2a-1b824bad68a0 · outbound

This paper cites Advances in neural information processing systems , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Advances in neural information processing systems , volume =

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.739927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.368739Z digest=sha256:8eeb76ed6319b3cba93467c9be679b3543d2259a3a6f4740c1aaab32e6bb9697

Observation 7e02f688-fc05-453e-81b9-93c2e71db1c8 · outbound

This paper cites Hugging Face repository , volume =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Hugging Face repository , volume =

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.725796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.373259Z digest=sha256:cf0ad587befdd2b2cbd7d3d4ebe139a964afacc053e6dc343a076e16910b7715

Observation 4988a4e9-2759-4038-844c-33085495760b · outbound

This paper cites an unresolved cited work.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.377789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.377789Z digest=sha256:a7b05f95f07d723d2a7e3375bd9ed91d9452435155bf64c4c4bdccf862c380f4

Observation 06de54c9-ba28-4e50-8f3b-0211714439d0 · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2024 , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Findings of the Association for Computational Linguistics: ACL 2024 , pages =

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.701090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.382398Z digest=sha256:e45f7e96334a972d6a2fdb5283f3d001814366bea6ab9e3608ef80a92df579c3

Observation cbb5181a-3311-4f3d-8d9e-152a201fe851 · outbound

This paper cites 2025 , eprint =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning 2025 , eprint =

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.686100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.387143Z digest=sha256:e11505a61fd2e536096a6049874380368a785bea33210a099bb5ac332ee4de2e

Observation 88758300-b891-42c6-ac45-1711f3e6a1f0 · outbound

This paper cites The Hidden Link Between.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning The Hidden Link Between

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.391538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.391538Z digest=sha256:6f56318f8377f94fefae5b1a666b5125406d0fc76bd507eb93bc4359b3b631ac

Observation 306e9743-db19-4fe3-9559-470fa2c9c306 · outbound

This paper cites The Fourteenth International Conference on Learning Representations , year =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning The Fourteenth International Conference on Learning Representations , year =

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.671530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.395884Z digest=sha256:2579329e38b256185d8122446dbb73eed77b8b44a69c28ff20ac4b652a0b708c

Observation 35b06d4f-4564-49d9-83a5-ab510ec2c64d · outbound

This paper cites Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , pages =

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:52:37.656364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-08T00:52:37.400464Z digest=sha256:135d9cfdded7fba96a85911b85cd401d43e72bceae60845776d1f31afcdf43f3

Observation 7fe7ba90-5138-4f16-8f00-5575d2f34a96 · outbound

This paper cites A Survey on Human Preference Learning for Large Language Models.

Hi-TTRL: Regulating Consensus with Hints for Test-Time Reinforcement Learning A Survey on Human Preference Learning for Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T00:52:37.405393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T00:52:37.405393Z digest=sha256:c7ff6fc98f20b502aa8bfe203010dcb5ab5c0be547f274e0ea6901f6cb988424

Pith citing papers

No inbound Pith citation observations are available.