Pith. sign in

Paper Citation Record · LEDGER

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

As of 9 August 2026, this Paper Citation Record lists 65 of 65 outbound references and 3 inbound Pith citation observations for arXiv:2506.16024.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16024 v1

Coverage vector

measured 65 of 65 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:49.180444Z

measured 68 of 68 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T17:38:50.313305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T20:46:14.342643Z

Reference resolution

65 of 65 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b9f414b3-ef0a-4c8a-b75f-41f8bef65d7d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.408301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.699129Z digest=sha256:fc836894621227fcdcba3eae53810716393c2bfb91478a2e81361b7d5224ba78

Observation 8cef63ee-899a-4c2a-a768-7c7e7237abcf · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.738491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.738491Z digest=sha256:18e1b46bcd950fba5055f4f01090af9856fecf71f4b0a4322503dd99f0a87915

Observation 2e556eaf-749a-4e32-bb5e-5633c28f2cb8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.237578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.814403Z digest=sha256:f1c6f0a76e78b845f35a24abd74dd04a0d7bc143dfdbb0514504abe3f494160c

Observation 2c265f62-e1ab-424a-ae77-ee52b5c3ca48 · outbound

This paper cites LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongWriter: Unleashing 10,000+ Word Generation from Long Context LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:43.860165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:43.860165Z digest=sha256:5702618ff81c5f9d5ed80521c61353caff638ee05ca4c02a0c503a1b1f5d493f

Observation 236c8de7-949f-42ca-afef-dbfbce19af36 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:54.052816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:43.907794Z digest=sha256:70a2af6c4e9d86f755bf19e24358f72441a767217b88bff8ec379953aac8bf55

Observation 4d36f76e-7a00-4171-a2b9-2da1305902de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.022270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.022270Z digest=sha256:6fcb6b5e1d27f3f853283e3da1dd7415e5bcc7163583926a29110f782f5de8ae

Observation 5f8acf18-99c1-4c75-817f-aa49b9edf5f5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.120598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.120598Z digest=sha256:45e02510baa034f9afbe095a00048c2a0afc10ebfd9a8c3c50dc35a493826f18

Observation 5f68243e-48e3-4e26-b3b3-ef1c2eeadc7c · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.272093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.272093Z digest=sha256:70267976ead8b4876529bc785fa2472da5632e53bac3af75518bcd67c1d4419d

Observation 5f09216f-3f4f-4346-8399-c217cdace7bf · outbound

This paper cites Extending Context Window of Large Language Models via Positional Interpolation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Extending Context Window of Large Language Models via Positional Interpolation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.314588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.314588Z digest=sha256:a26fa503414a5720608b79502774b4f3f5faa771ad50b7e675f355d7ba4a947f

Observation 4417de2c-c90c-4496-8e18-b2eb200b88d8 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.894572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.329946Z digest=sha256:a565c42f1099f107a47ff1c7f34c94c5fdf279569497ecf96f26a977d17bc1f8

Observation 5974bde4-aca8-44c4-92c9-6f0a835ffd91 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.735912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.396554Z digest=sha256:4cccce82544aaf828eb1118e4989333088a8d2b9e4206eafc40f48e2d89239ac

Observation 12cae417-0717-4ead-9967-196f7ce980de · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.562496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:44.479076Z digest=sha256:9319f98ce8282d272d49040d25f48b07bc102ce9db2768b2de0df60e87d4f600

Observation 6d4cbf06-a589-4f36-8b55-fe83b1168291 · outbound

This paper cites Human-like Summarization Evaluation with ChatGPT.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Human-like Summarization Evaluation with ChatGPT

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.575596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.575596Z digest=sha256:10faaa59e1444e4db30c9956bbb199ef86e4710329ce9bf48960495fd03ca95a

Observation ee2d02a8-d372-4b79-83ea-5b07b22a068b · outbound

This paper cites The Llama 3 Herd of Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation The Llama 3 Herd of Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.657266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.657266Z digest=sha256:1284f0f01dce3a6b651131ce6758271bead6deed3317c4cf7ba4dcb2d1689de3

Observation d7fb7a6a-2efb-44d7-b6cd-fb0a866d80fa · outbound

This paper cites Efficiently Modeling Long Sequences with Structured State Spaces.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Efficiently Modeling Long Sequences with Structured State Spaces

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.770766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.770766Z digest=sha256:8edab8797f55a35530deec638a6c8a85ee479106bc03cad35055ed1c7a042a14

Observation 7477ff42-3194-405c-b262-185315e163df · outbound

This paper cites A Survey on LLM-as-a-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation A Survey on LLM-as-a-Judge

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.861710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.861710Z digest=sha256:dfebba06141f18f3afd0e6d66d1f5dd5e33bf65529cdc40367e140ff68fe2023

Observation 7042a2e3-5ddb-4d6c-8f2a-26fb77cc0a4a · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:44.944982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:44.944982Z digest=sha256:60ea5c46b819dd603469457945007595a26b4ca8926e67c5233af4a111cea5a1

Observation c35b17de-6ccc-4761-af96-722dbe43503c · outbound

This paper cites MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation MInference 1.0: Accelerating Pre-filling for Long-Context LLMs via Dynamic Sparse Attention

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.021883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.021883Z digest=sha256:98f2867ad6a132da39c081980c5fe0ddd76e18cfaf24b0a40db1b5fdf607d846

Observation 529850c2-aedf-4374-af1a-c1af52d8c8fd · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.103956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.103956Z digest=sha256:0c958209e34b580ab7f1c2a99a496d0ee22927f788ff143dda4da704364aab1c

Observation 34d087a0-cd4a-4e21-b2f9-d566fe2f35c2 · outbound

This paper cites LongForm: Effective Instruction Tuning with Reverse Instructions.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongForm: Effective Instruction Tuning with Reverse Instructions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.194684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.194684Z digest=sha256:d0d43280d3ceaf352e5b49a4b15563630d559f7d32c50a7980765fdbdb88a73e

Observation a5f31b34-4056-49a4-8c61-6cf39760e3e5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.402762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.242809Z digest=sha256:a85f8da14f10ee41b5e2736f98a2a895688ce86507c404d008ab47e79bab9904

Observation 0d698436-d03b-4e13-a145-6c37421a7044 · outbound

This paper cites LongLaMP: A Benchmark for Personalized Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongLaMP: A Benchmark for Personalized Long-form Text Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.315741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.315741Z digest=sha256:fe74ad693156f7fe686a438c051183c17c16da2d21ba40c3106a86089d440248

Observation c3c620a6-2120-4b1c-a9c2-a7dcadeb3e73 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.216011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.393122Z digest=sha256:11e423c37752e8daf847ae705a047f84a5711743354d8fd2063644b7c5950291

Observation 0179e988-ab6e-43f0-92ac-1d417b699cf5 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:53.054656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.467796Z digest=sha256:e7528f20475e63bec9ca97a4bf0b6c7020d6f1a0add9f6811e431d7a7b96c7b6

Observation a4d8107f-edb6-4c45-b171-df8fd309b3d6 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.915515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.539371Z digest=sha256:a18bac020f5c2e8c13477a55981380663a8203b6859210220a59e52905335411

Observation dfa0033a-f9b9-48e1-b771-49d7b24b3522 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long-context LLMs Struggle with Long In-context Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.617067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.617067Z digest=sha256:5590053994ce068684d86c9e5d2a378a3aaff4aa8178d097e1a24c15320a8927

Observation 5c469c74-a5be-4097-ac07-bf4a99c668c9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.753067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.687633Z digest=sha256:b100145b07208238d9292ae98b8936ccac9a1b5130bad830516113db69273830

Observation 88c43ae2-e0bf-4f22-a362-20ad6e67b8e9 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.520620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:45.739551Z digest=sha256:dd4b96af0a5774d0568a3fe3f51855dbe9fd9d7f860333943cbf898c6ffce046

Observation 13afb7af-db6b-4057-8a33-5dd80c6b2a6e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.792936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.792936Z digest=sha256:eaaed9d265377e1e59b5aa6db9cd4ff82dd6c9eecaa6dd6f6e5a0b6070320290

Observation a6fe5f2d-6df3-4d70-a985-0738b2b35da9 · outbound

This paper cites DeepSeek-V3 Technical Report.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation DeepSeek-V3 Technical Report

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.880835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.880835Z digest=sha256:31a1b92f9582141d74f6422c20c9ccea6547dd905b753a901f1a03b6e87b5fd9

Observation 830307c8-b165-4305-bbf1-ec2a4ef09ced · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.941061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.941061Z digest=sha256:03c758884e181304fe866f3f7e797b9f28673b58926499b06d836d3229ac993a

Observation 086c634b-f677-4d9c-8c9a-9cef0dbeb295 · outbound

This paper cites LongGenBench: Long-context Generation Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongGenBench: Long-context Generation Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:45.997697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:45.997697Z digest=sha256:e9b707e3c72a9c036bcf7621cc8e23f004b999ece1112572da6870b2b53f2125

Observation a592615a-3980-41cd-aced-ef607163ab2d · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.378599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.093389Z digest=sha256:9500fb6b040d105f390a749ed01ed727acfbd23682345ca6b713d9071e9c0145

Observation a32eb9f3-2544-4fb3-8e49-00b627fadd6e · outbound

This paper cites Exploration of Masked and Causal Language Modelling for Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Exploration of Masked and Causal Language Modelling for Text Generation

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.759678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.175021Z digest=sha256:1593a9db6fdca30c8266d8d2f72755e5715e6f2e53091059ad8e3b33e3b6e97c

Observation 10dafc76-c9c9-4275-8617-cbff01502816 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.222337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.255452Z digest=sha256:aee7764ee158977a0b15a94042ed0f41f94f5fcbeb4849a5ef30eb53a2f9aa5f

Observation f888ce7a-2b12-49b9-9897-b8e67fc92e8f · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:52.074672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.319450Z digest=sha256:09f9980139cd2adf00a0c3ce2be30acb7a6e228d71b1d36524d79091f3f2978e

Observation d63995fc-019c-46d9-9870-002206de8f2a · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.383781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.383781Z digest=sha256:ad38a939f5d33176c6727d8ee3a7618a50e463c5d2b92a39ceeac835440e50bc

Observation d42f8e59-0122-4261-9780-078650f3b807 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.451664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.451664Z digest=sha256:5f2ad05a58b181e581083a0ef6155558d9e27d996db5c862455bad50a8c323dd

Observation 92eb86ef-20e4-4d99-af77-27d2cbf87aca · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.955371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.512069Z digest=sha256:8ebdef318735a6ffc77b937a719c470c5592a8cb3d6a75d5f14224c847e233d9

Observation 12605ee8-c6ba-4fd6-9164-6c7aac2cafc5 · outbound

This paper cites Suri: Multi-constraint Instruction Following for Long-form Text Generation.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Suri: Multi-constraint Instruction Following for Long-form Text Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.578703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.578703Z digest=sha256:626d80c06f53fe49a6f9821cd52369ae9afc506e5e256a7495bcabe9e1b0065e

Observation 51abed4d-7cb7-4fb6-b823-3491463ad441 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.856072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.677908Z digest=sha256:2c1154b9799714eb95c2d274ec5a7565240892aaa629de2171c3afa981f1632d

Observation 61142bbc-84ef-42df-8c1f-b9fea274b958 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.777966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.777966Z digest=sha256:576a5f55a9b8e8ae5d7183427153cf77e87f78935181ba73ae461fc373a3ff00

Observation d6570b55-f2a4-4d3a-86ad-9cac1299f8aa · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:46.874265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:46.874265Z digest=sha256:aa70dae506fd72d33020337e6ccfc6b94d56666633d359d67b4b052c4a025670

Observation 9eed7ccb-e89a-4efb-9ae5-7a14134382ce · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.683561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:46.997928Z digest=sha256:5d268e894a31ccda196fa61697ebec6e7117084e88fc7cb83db90c0300d0a1b9

Observation 40630c9e-4032-4943-8cbe-acfa5b8f5174 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.087005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.087005Z digest=sha256:97ab6595ea5ff94290d28b937d5d46c3fab7d528454b3f91aeea5a2b02b978a2

Observation 5879515a-38ee-40e5-bfec-08e0afebe51f · outbound

This paper cites Preference Ranking Optimization for Human Alignment.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Preference Ranking Optimization for Human Alignment

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.131971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.131971Z digest=sha256:2e1e1b3273a3c35c01832696e5c99159dcd15649a8ab1160aa9a8a32fab311fe

Observation d61cef29-e34b-401a-896f-fe36f61d1391 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.412239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.192922Z digest=sha256:c4191a96d5165fd93cb18bb804b2b22262b339f5bc3f81b276309b8ec61f1cd1

Observation c63adc82-6224-4d5b-9678-88897cd54dfe · outbound

This paper cites PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation PROXYQA: An Alternative Framework for Evaluating Long-Form Text Generation with Large Language Models

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:49.444103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.284775Z digest=sha256:dde98c8837166278993fc8f547fde148344810d4a3912ccad7b43f03f0cf5822

Observation 6272fee1-3493-4ae4-a8b7-0fe84e220959 · outbound

This paper cites Long Range Arena: A Benchmark for Efficient Transformers.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Long Range Arena: A Benchmark for Efficient Transformers

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.407269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.407269Z digest=sha256:90b98fbeca0f7d22c30a37b435b9ff3d1f55d64c97176070ea09c50004287c4b

Observation b233dbd9-8c9d-474e-a345-722c41b0d59b · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.205623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.467394Z digest=sha256:77b8d7294508f15f4fa5c8826c1fcfbc2097c435efbbad8ae39a7827da3859b3

Observation 0883774c-4db3-4d91-a393-76b9d3224910 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:51.006117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:47.581044Z digest=sha256:db7bdec7a32e987aeed52ef9af9e1b36c6035adac6b3b67eb4d3e20e9b2025f5

Observation 16eff6a0-1561-4ddd-9491-e91ea99aa61c · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.732683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.732683Z digest=sha256:cb48db10a0e8bccda4d8b3b420aa4e8c85c1ae0b89a97d3d3291268906c4fecd

Observation 164df63e-321b-459d-a263-6131ebd555aa · outbound

This paper cites Self-Taught Evaluators.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Self-Taught Evaluators

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.837626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.837626Z digest=sha256:d10f816773afe79080a60220c24063256f2a4e62e59192766ded593ca4aff156

Observation 95be38fa-aece-4e88-8109-e4c96cdb5f47 · outbound

This paper cites Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:47.916012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:47.916012Z digest=sha256:5bb1f472a993b3b71a80d0829141ac04b52f52f92d8fbc599802fff1a67c67c5

Observation 1d64bf2a-1c1b-4886-bc56-f5b34b402675 · outbound

This paper cites Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.044179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.044179Z digest=sha256:96b14a50b2a94efa4220c71a58cfaf2b50f57b04263c9975464a4d0deb1c253d

Observation 1af95218-cebe-4276-b85b-922499cfaf73 · outbound

This paper cites Effective Long-Context Scaling of Foundation Models.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Effective Long-Context Scaling of Foundation Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.167491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.167491Z digest=sha256:879109174f5681a501751f021e7c06b578a0aa23858b661a5a6586282995282a

Observation 94a2298e-0f4d-4ee8-ac87-a8499243abc1 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.776994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.262762Z digest=sha256:f238cbd1e7533fcb87c4591452dd0d8b188c1f4b06edc50feb1afe17dee9f4d8

Observation 43cd9b3f-4865-4c7d-8ad1-16fd701443bd · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.613756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.378006Z digest=sha256:df19f8122669c923a0b534af9212973126bf4fc80f0d153b6b75bae0ba96a123

Observation a6f978f5-fece-457e-b9b1-eb2383e21ca4 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.390405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.477481Z digest=sha256:bf9542e6de522a980a95312f5154ac6ff3c6d5b60a2dea45958fe5843bfd262d

Observation 835d83dc-d79c-49c9-9398-323753002d8e · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:49:50.209143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T23:49:48.564204Z digest=sha256:91821731673cd3e12d38cad2de7b9a36a1cf209fbaf9a1c68edb11980a1fcda4

Observation a45831ae-3018-49f0-b5f9-0c5dc3b00652 · outbound

This paper cites RRHF: Rank Responses to Align Language Models with Human Feedback without tears.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation RRHF: Rank Responses to Align Language Models with Human Feedback without tears

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.654069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.654069Z digest=sha256:8e77f01f909c5fe78da28b78bb8045532ab080f4e20184a3097d94de7c3e0e34

Observation 5c289b82-76f9-485d-9729-a6e8826b1212 · outbound

This paper cites an unresolved cited work.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation Unresolved cited work

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.826479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.826479Z digest=sha256:e02b117c006caddbd2a3a3043f393a8afbde387013aa6d99bd1be64c0b9be4d2

Observation 3fe594dc-d118-4ddf-9c68-3c4e21e6461c · outbound

This paper cites LongReward: Improving Long-context Large Language Models with AI Feedback.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation LongReward: Improving Long-context Large Language Models with AI Feedback

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:48.916548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:48.916548Z digest=sha256:18222230bd6c394fffa2976591c0403c86aaf81281772620c828e01401609c34

Observation 3e0be4d5-c228-4e61-a556-972f360edc0e · outbound

This paper cites online" 'onlinestring :=.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation online" 'onlinestring :=

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.024302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.024302Z digest=sha256:bd51667ace679d0a6fed16af55de00fc19cc62838890c79347b907e3ffa6732a

Observation 3fd049d2-2af4-4dbb-8589-4f3b74950240 · outbound

This paper cites write newline.

From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation write newline

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:49.180444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:49:49.180444Z digest=sha256:f1eaf641b93722d28706cabd9f15a2f4c401226a96783dfc0eb3dde0736fa424

Pith citing papers

Observation 30c3a4b0-31ad-41ca-941c-9751ccf1f351 · inbound

A Survey of Reinforcement Learning for Large Reasoning Models cites this paper.

A Survey of Reinforcement Learning for Large Reasoning Models From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-18T00:05:31.664012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T00:02:24.352947Z digest=sha256:db38592eb69ce3f5b5761dd9e879e0d449dc16d307362375672ce934c63e6f53

Observation 2130b475-363d-4d1c-8fd9-220de7dafc4d · inbound

SEIF: Self-Evolving Reinforcement Learning for Instruction Following cites this paper.

SEIF: Self-Evolving Reinforcement Learning for Instruction Following From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:55.949246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:51:09.514927Z digest=sha256:6ad847e4939609ba8c27d2518b721de75a26ab8c1fd3b98840586600f7e872d1

Observation 6c6922a7-4223-4a49-87bf-676531d8ee2e · inbound

Trust Region On-Policy Distillation cites this paper.

Trust Region On-Policy Distillation From General to Targeted Rewards: Surpassing GPT-4 in Open-Ended Long-Context Generation

Reference 171

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:46:14.343964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T17:38:50.313305Z digest=sha256:8c695d1f6816e52e0bc6d055df767cd03b5b57fef0db926847579c4ce1fc19b9