Pith. sign in

Paper Citation Record · LEDGER

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

As of 20 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 3 inbound Pith citation observations for arXiv:2506.16024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16024 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:49.180444Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T17:38:50.313305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:46:14.342643Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b9f414b3-ef0a-4c8a-b75f-41f8bef65d7d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.408301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.699129Z digest=sha256:0f41e9d0609492aabaeeb9f6fbb16c066ebd0513d95b7518f643b005dcf3475b

Observation 8cef63ee-899a-4c2a-a768-7c7e7237abcf · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.738491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.738491Z digest=sha256:dfe10700e5fd87fd91ceeb348c21c05795b1df1836ba072042556ead02596bfa

Observation 2e556eaf-749a-4e32-bb5e-5633c28f2cb8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.237578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.814403Z digest=sha256:b30eb9bfc50aa0f33018de5aa98295ec22895dfd16c535b94b89a3417fe39c36

Observation 2c265f62-e1ab-424a-ae77-ee52b5c3ca48 · outbound

This paper cites LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.860165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.860165Z digest=sha256:69bb83b63c1b40e4d4b063745abd531471c5ed3218a8a32ad17586584839cc4d

Observation 236c8de7-949f-42ca-afef-dbfbce19af36 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.052816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.907794Z digest=sha256:2e5d70e8e310b72fc9f3b0fdd0db89a95efb9178e74cc095bbdb6aab305b5d3e

Observation 4d36f76e-7a00-4171-a2b9-2da1305902de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.022270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.022270Z digest=sha256:221caec95b300738173844ae7a795d0e7c6032d3b40153cded71e4b96b5ad222

Observation 5f8acf18-99c1-4c75-817f-aa49b9edf5f5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.120598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.120598Z digest=sha256:1735af1cc6c60e4e6b7f075e70e1c850fddd4eb6ee2a684a8246be46307d324a

Observation 5f68243e-48e3-4e26-b3b3-ef1c2eeadc7c · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.272093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.272093Z digest=sha256:8e5bf05d2f25ec318a46d7b35417f31ba1ec21c093e31de90b328fbb4b4a6d3d

Observation 5f09216f-3f4f-4346-8399-c217cdace7bf · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Extending Context Window of Large Language Models via Positional Interpolation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.314588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.314588Z digest=sha256:faeadb97f7e5bf29a2a89af43bf4d7908871ef741059e441151b7f91d4ac8e55

Observation 4417de2c-c90c-4496-8e18-b2eb200b88d8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.894572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.329946Z digest=sha256:2cdfdf78cd0ee45db7887ce9ca80ade8e0ad463319cb1586ac73cc3386b9c7a7

Observation 5974bde4-aca8-44c4-92c9-6f0a835ffd91 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.735912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.396554Z digest=sha256:d78dd27eef01f9e0e600ccb429a8d464870f72956efacfbcaba245d312679929

Observation 12cae417-0717-4ead-9967-196f7ce980de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.562496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.479076Z digest=sha256:de565af6f6be3bae2bef5d9476238cd72174f12f4cc7b28d7c6d1d6120cb784b

Observation 6d4cbf06-a589-4f36-8b55-fe83b1168291 · outbound

This paper cites Human-like Summarization Evaluation with ChatGPT.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Human-like Summarization Evaluation with ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.575596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.575596Z digest=sha256:91c1cf88568b5b5b691ccf358a80acf7c5a0f8459f90359125154ee29f275e80

Observation ee2d02a8-d372-4b79-83ea-5b07b22a068b · outbound

This paper cites The Llama 3 Herd of Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.657266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.657266Z digest=sha256:533cd54b67b20368a7db42fd509033546b87649f0fd8ae638ce5a0ad81764bcc

Observation d7fb7a6a-2efb-44d7-b6cd-fb0a866d80fa · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Efficiently Modeling Long Sequences with Structured State Spaces

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.770766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.770766Z digest=sha256:2f9aa9c99aa287e0c7a5d8ce371e946e00c9bf3e0903a9ccd74e7b4689359d62

Observation 7477ff42-3194-405c-b262-185315e163df · outbound

This paper cites A Survey on LLM-as-a-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation A Survey on LLM-as-a-Judge

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.861710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.861710Z digest=sha256:875defa752e8fadb0490559f9b299c1e76656070e3246194bdf409192d1a3f87

Observation 7042a2e3-5ddb-4d6c-8f2a-26fb77cc0a4a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.944982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.944982Z digest=sha256:1c6f29e9199c511900de3f3cfe0ad401a10a468980bdbde2f3307afb39d67c7a

Observation c35b17de-6ccc-4761-af96-722dbe43503c · outbound

This paper cites MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.021883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.021883Z digest=sha256:a02ec1989944a13bdf811d75dcf26079c9f3f821c0676d1dcd08155bc47ab2fa

Observation 529850c2-aedf-4374-af1a-c1af52d8c8fd · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.103956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.103956Z digest=sha256:4a2d160d225263041da634fefdfc1626956e3886fac68555ee93319b605fbc33

Observation 34d087a0-cd4a-4e21-b2f9-d566fe2f35c2 · outbound

This paper cites LongForm: Effective Instruction Tuning with Reverse Instructions.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongForm: Effective Instruction Tuning with Reverse Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.194684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.194684Z digest=sha256:5df8e0ec7abf369a561b4d4662fe25ba084a2ede71a39ff9cbf44e3a649b23d7

Observation a5f31b34-4056-49a4-8c61-6cf39760e3e5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.402762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.242809Z digest=sha256:d157792feb1d2df51ccd4d6e5435d88e63031ca667873da3d172a7416a0af850

Observation 0d698436-d03b-4e13-a145-6c37421a7044 · outbound

This paper cites LongLaMP: A Benchmark for Personalized Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongLaMP: A Benchmark for Personalized Long-form Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.315741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.315741Z digest=sha256:92d8c50f3b1d725275565993827e7a466878458190a6526b6e94627cfe0d02f3

Observation c3c620a6-2120-4b1c-a9c2-a7dcadeb3e73 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.216011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.393122Z digest=sha256:27ec56eb4085d845692c070ed1839a1b8a9badb13970930ff332f270e3c8bee5

Observation 0179e988-ab6e-43f0-92ac-1d417b699cf5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.054656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.467796Z digest=sha256:ef8318cdb9f72194325d453f42d517c242372b28db66ad5939424b0b280f3c1e

Observation a4d8107f-edb6-4c45-b171-df8fd309b3d6 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.915515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.539371Z digest=sha256:620de718abd76d323d66132e26402641b644785f916eb05251069e290d9d8ea9

Observation dfa0033a-f9b9-48e1-b771-49d7b24b3522 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long-context LLMs Struggle with Long In-context Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.617067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.617067Z digest=sha256:fb48a09b1e9488b3af09cd34a23f3cdd0af75645ad19bc39724d7044b92036d9

Observation 5c469c74-a5be-4097-ac07-bf4a99c668c9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.753067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.687633Z digest=sha256:52d416ad615117fa3ab212a2885fca23abf3e263879b47742c20eda64cabcadd

Observation 88c43ae2-e0bf-4f22-a362-20ad6e67b8e9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.520620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.739551Z digest=sha256:4d7d4c271970deadb4234385b686d637ff36259dde5b38353f962332219faa13

Observation 13afb7af-db6b-4057-8a33-5dd80c6b2a6e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.792936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.792936Z digest=sha256:e2e2d2f17005f1417803ca6487d2d854295352176e3b8f03b44b853178f77aff

Observation a6fe5f2d-6df3-4d70-a985-0738b2b35da9 · outbound

This paper cites DeepSeek-V3 Technical Report.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-V3 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.880835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.880835Z digest=sha256:9e20a7ab8d2d17c56d87fce8272b7a843950e8e781be16fb7f384024ce28aafa

Observation 830307c8-b165-4305-bbf1-ec2a4ef09ced · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.941061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.941061Z digest=sha256:adba4873d5b0f53519bef2bccff2ce0093e6d855eaaaa1f0479912136834837e

Observation 086c634b-f677-4d9c-8c9a-9cef0dbeb295 · outbound

This paper cites LongGenBench: Long-context Generation Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongGenBench: Long-context Generation Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.997697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.997697Z digest=sha256:113a92ea3252bbbe82c8c879569a68e1cbd8da343b4cf6c035b2ec36929e34b8

Observation a592615a-3980-41cd-aced-ef607163ab2d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.378599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.093389Z digest=sha256:096a7be86f2f0b73080c766881f632f34a78c94c462e3e555994f87f1cbdf804

Observation a32eb9f3-2544-4fb3-8e49-00b627fadd6e · outbound

This paper cites Exploration of Masked and Causal Language Modelling for Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Exploration of Masked and Causal Language Modelling for Text Generation

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.759678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.175021Z digest=sha256:dd79a1484374640e6701af8daba582a2b6af1f727af7d850ac7860679bb80e94

Observation 10dafc76-c9c9-4275-8617-cbff01502816 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.222337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.255452Z digest=sha256:e00031d59bc5f64bf0945862ae9e58b1ee52d4ed504d4f3cc7a32aad07296f70

Observation f888ce7a-2b12-49b9-9897-b8e67fc92e8f · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.074672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.319450Z digest=sha256:6fbe7f935b3f189d19ef946c1f0eec3163d8d334b25a960639147c86dcdadf4f

Observation d63995fc-019c-46d9-9870-002206de8f2a · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.383781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.383781Z digest=sha256:deffb34c10608536c473e05c78d33c4e12f54160ecbf6aadf71c7b815075a03e

Observation d42f8e59-0122-4261-9780-078650f3b807 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.451664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.451664Z digest=sha256:d848075964a9de3315cd8a70f0988dae3f741ecd0139fe031f6bab4616ca8b87

Observation 92eb86ef-20e4-4d99-af77-27d2cbf87aca · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.955371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.512069Z digest=sha256:2f7caa29261cc968dd90fe280a1b94e40e56e7c5975a627d0601cc7e38d43bb8

Observation 12605ee8-c6ba-4fd6-9164-6c7aac2cafc5 · outbound

This paper cites Suri: Multi-constraint Instruction Following for Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Suri: Multi-constraint Instruction Following for Long-form Text Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.578703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.578703Z digest=sha256:27c033a26dd37a135f6370bafdec33aa103386d8407d29cf7487dd063e9d2606

Observation 51abed4d-7cb7-4fb6-b823-3491463ad441 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.856072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.677908Z digest=sha256:db482dfe21bd673a634790f6e8ba1b8fb2c0be70e8495aac5ea7c7c25d487faf

Observation 61142bbc-84ef-42df-8c1f-b9fea274b958 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.777966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.777966Z digest=sha256:9c02c68bd1725e3e6090847af9896f3dff413383551f78da26e9b05ff13db47c

Observation d6570b55-f2a4-4d3a-86ad-9cac1299f8aa · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.874265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.874265Z digest=sha256:6973ecbc0004236455445fbea484036f1bbc94e322934ffbe9a33d97361f0c62

Observation 9eed7ccb-e89a-4efb-9ae5-7a14134382ce · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.683561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.997928Z digest=sha256:0678c3eedd6543e59717b74c2c2abe8d0962e7ae4ca8bbfab31bd49ef24702b8

Observation 40630c9e-4032-4943-8cbe-acfa5b8f5174 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.087005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.087005Z digest=sha256:99b50607f0ddf5d39114d6d286c84a01d9ae9434486f9c501bac4fcb3decc3b8

Observation 5879515a-38ee-40e5-bfec-08e0afebe51f · outbound

This paper cites Preference Ranking Optimization for Human Alignment.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Preference Ranking Optimization for Human Alignment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.131971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.131971Z digest=sha256:e76d28aadeb8f8ee61ee8d44eafb8a9e8f4c329b5df351fb4ac6ac5f319ed75c

Observation d61cef29-e34b-401a-896f-fe36f61d1391 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.412239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.192922Z digest=sha256:3d76b48d32ac8ee0149f269de6b5a9aa81cf3e40f2f79f43263c27e89c994c19

Observation c63adc82-6224-4d5b-9678-88897cd54dfe · outbound

This paper cites PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.444103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.284775Z digest=sha256:25b81015aedcbfe6814b00cfa25effb74cc31c203b96dd8a63b30f16f0dd6269

Observation 6272fee1-3493-4ae4-a8b7-0fe84e220959 · outbound

This paper cites Long Range Arena: A Benchmark for Efficient Transformers.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long Range Arena: A Benchmark for Efficient Transformers

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.407269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.407269Z digest=sha256:55e37921b2ad7d99085fe2b8ca7f98bde1aae765c125a28ee83c5fbbe0d833ba

Observation b233dbd9-8c9d-474e-a345-722c41b0d59b · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.205623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.467394Z digest=sha256:afa5378eb7c3cc290ef1a0b0d65c1c2040f9044961b50b2f92543c71c5f193db

Observation 0883774c-4db3-4d91-a393-76b9d3224910 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.006117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.581044Z digest=sha256:a49c659b5f92d8d8f502d2e5214d4cb35955114959d095680494cbdbfcd4c7d3

Observation 16eff6a0-1561-4ddd-9491-e91ea99aa61c · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.732683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.732683Z digest=sha256:6c06f844d64ac16a37dbf5378a32b36c0aa2d359235849a03c45e47b4c4a6ae2

Observation 164df63e-321b-459d-a263-6131ebd555aa · outbound

This paper cites Self-Taught Evaluators.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Self-Taught Evaluators

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.837626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.837626Z digest=sha256:9a1088f820cfd4fa42e748ddce3385ff37f7cf864cebc61f2df9f21bf656df44

Observation 95be38fa-aece-4e88-8109-e4c96cdb5f47 · outbound

This paper cites Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.916012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.916012Z digest=sha256:26605614bddcc7aa5819e6d5138e38795cc4dbc60dbdaccc4e7081a8ebf31566

Observation 1d64bf2a-1c1b-4886-bc56-f5b34b402675 · outbound

This paper cites Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.044179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.044179Z digest=sha256:f0375a81acf76728ac4210661fbf85b68955a89e4226f0af8ce8832321304c77

Observation 1af95218-cebe-4276-b85b-922499cfaf73 · outbound

This paper cites Effective Long-Context Scaling of Foundation Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Effective Long-Context Scaling of Foundation Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.167491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.167491Z digest=sha256:f2145d548766b0e82b27cd043d7857ea56177cc57a03fc1ec70946cbb507a02c

Observation 94a2298e-0f4d-4ee8-ac87-a8499243abc1 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.776994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.262762Z digest=sha256:d57903ad4a4fc114163e872020acd5965db27021647a9c5d4dbe88690af1a5f2

Observation 43cd9b3f-4865-4c7d-8ad1-16fd701443bd · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.613756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.378006Z digest=sha256:129ad9bef5b6510f04c74b87e7f435f1c6236ddb84bcefd6b70ca52b91c8e032

Observation a6f978f5-fece-457e-b9b1-eb2383e21ca4 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.390405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.477481Z digest=sha256:50eaf4980a03f3955881234442f24cc0117fdb2288288ad32e504e1fa8ed3723

Observation 835d83dc-d79c-49c9-9398-323753002d8e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.209143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.564204Z digest=sha256:53f24b4cf49b9c491b3cac1b6374c594d6bad1b66dd0ac891cfe9004cf2ab4bf

Observation a45831ae-3018-49f0-b5f9-0c5dc3b00652 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.654069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.654069Z digest=sha256:210e5a0fc64fdd237722872ea8331546fb23ed1494a3963d2e9a9e1e7f5f6ee7

Observation 5c289b82-76f9-485d-9729-a6e8826b1212 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.826479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.826479Z digest=sha256:13caf0fcfc4293c5af2545bc90b7f06bd77a763613fe05a9b013005be29fdd08

Observation 3fe594dc-d118-4ddf-9c68-3c4e21e6461c · outbound

This paper cites LongReward: Improving Long-context Large Language Models with AI Feedback.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongReward: Improving Long-context Large Language Models with AI Feedback

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.916548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.916548Z digest=sha256:895efb1d006a4ae58e57d134c74b23ed4ab780dca8f54e3a27d716bdad49960d

Observation 3e0be4d5-c228-4e61-a556-972f360edc0e · outbound

This paper cites online" 'onlinestring :=.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation online" 'onlinestring :=

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.024302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.024302Z digest=sha256:dfad27e150483992696631e43f7ac49a7343ef1f094aeb50d7b3c4d63d05f59b

Observation 3fd049d2-2af4-4dbb-8589-4f3b74950240 · outbound

This paper cites write newline.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation write newline

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.180444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.180444Z digest=sha256:8bb1d3606bbd6bb134da1cdca5ee693c19582d4e109a63c7d35cfcd564ef5737

Pith citing papers

Observation 30c3a4b0-31ad-41ca-941c-9751ccf1f351 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.664012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:2eefe239d97a0e44950496d137ccb935bcd2beaa06e61a93027b36ae219c55dd

Observation 2130b475-363d-4d1c-8fd9-220de7dafc4d · inbound

SEIF: Self-Evolving Reinforcement Learning for Instruction Following cites this paper.

SEIF: Self-Evolving Reinforcement Learning for Instruction Following From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:55.949246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T01:51:09.514927Z digest=sha256:8a6ae39aedcad457771b86144084bf8657635e9825568fb054c9725a4a7b8cb6

Observation 6c6922a7-4223-4a49-87bf-676531d8ee2e · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 171

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:46:14.343964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:64fc66198a79f0f4db800e1cc031d47b9369ba3b253688984d2edb328d39d1ad