Pith. sign in

Paper Citation Record · LEDGER

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

As of 23 August 2026, this Paper Citation Record lists 94 of 94 outbound references and 1 inbound Pith citation observation for arXiv:2505.11166.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11166 v3

Coverage vector

measured 94 of 94 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:06:21.682241Z

measured 95 of 95 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T06:07:36.830550Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T06:11:20.403487Z

Reference resolution

94 of 94 outbound references displayed

  • verified exact4
  • verified fuzzy19
  • unresolved69
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c4f53665-b4c7-46ac-85a3-0e29ffabc915 · outbound

This paper cites A general theoretical paradigm to understand learning from human preferences.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A general theoretical paradigm to understand learning from human preferences

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.287572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.287572Z digest=sha256:05960337f02dbbb576b09b53fa94a6cf7634b84c4590683003c7b510a8e9b1ae

Observation 50f25e68-8a31-40ed-88f7-f69ee8f07547 · outbound

This paper cites Unifying cross-lingual summarization and machine translation with compression rate.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unifying cross-lingual summarization and machine translation with compression rate

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.297575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.297575Z digest=sha256:5ead8871ecfb41565765c48e70fa0590d9bb20a5041ed1911f7f88ffb66d0840

Observation 321dcabc-c338-412e-9745-145c6390b12d · outbound

This paper cites CItruS: Chunked instruction-aware state eviction for long sequence modeling.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization CItruS: Chunked instruction-aware state eviction for long sequence modeling

Reference 3

Resolution
verified exact
doi, observed 2026-08-15T21:06:21.931963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.308569Z digest=sha256:f6021eab1c33776465f6a371ab0de68937505d5146b21f47bf26173f4e8badf0

Observation 11500c01-8f75-4510-b022-87f70d5f7911 · outbound

This paper cites LongAlign: A recipe for long context alignment of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongAlign: A recipe for long context alignment of large language models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.313134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.313134Z digest=sha256:364f08f3c37dc2c871da3ebbbacf9b5e9e331748360b073c485ff6d348ebe2a8

Observation 66a26a42-e044-4bf9-9533-5cd68f07d188 · outbound

This paper cites LongBench: A bilingual, multitask benchmark for long context understanding.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongBench: A bilingual, multitask benchmark for long context understanding

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.317522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.317522Z digest=sha256:91af6cf06606d6e32d2411be494d0ab60869019c4e127ac648dace0c82377835

Observation ea7e55a8-00ce-4a44-96c6-3cda48e65be6 · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.321988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.321988Z digest=sha256:e020a3530f5f23e695769a542429dffc62e8f3c520c95b636e26bb0e9184926a

Observation 56d46e58-0eb1-462e-99bd-7a7d72042d52 · outbound

This paper cites Longwriter: Unleashing 10,000+ word generation from long context LLMs.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Longwriter: Unleashing 10,000+ word generation from long context LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.326027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.326027Z digest=sha256:42a6ae959a385be00e05cbc86863372c150a764326c3266a8b52a63a54b3dfcc

Observation 51b86b48-fcf2-4f56-835e-0b333994341a · outbound

This paper cites Luna: A lightweight evaluation model to catch language model hallucinations with high accuracy and low cost.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Luna: A lightweight evaluation model to catch language model hallucinations with high accuracy and low cost

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.329421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.329421Z digest=sha256:cab09f09edbac4b4a1497e3c6fcc716792830b4736cd23a82c31cba65777199c

Observation 3ab93836-3b06-485c-b764-c8e80f7dc7b5 · outbound

This paper cites Context-DPO: Aligning Language Models for Context-Faithfulness.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Context-DPO: Aligning Language Models for Context-Faithfulness

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.333355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.333355Z digest=sha256:7b061df25aa13cda2b325c9e2ddf77e448ef37bef53b450d7311861ca8eded8e

Observation 59aa538a-0ed0-4da0-8ece-95846396d7f5 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.337415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.337415Z digest=sha256:e855a949d1107a127440464b80f97abc310391a9a1020aa90e0c39f576a74bf1

Observation b8869fc9-9c52-484c-9693-7fbcd9f45470 · outbound

This paper cites LongPO: Long context self-evolution of large language models through short-to-long preference optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongPO: Long context self-evolution of large language models through short-to-long preference optimization

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.340734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.340734Z digest=sha256:9563de3ef3ff0e490b9c179e50b06f23604e0e4106fd4262b63c28a3d4ac5739

Observation a4ca1c02-8e36-429a-9e58-8d740d7c2e31 · outbound

This paper cites LongloRA: Efficient fine-tuning of long-context large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongloRA: Efficient fine-tuning of long-context large language models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.344025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.344025Z digest=sha256:2eb6f77c1c5b337c609aaa5f6896f7ed87480d63f1de6eacd72b51f9874be1f7

Observation 0d5d4a65-3291-41c8-ab84-6cc638913e68 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.347826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.347826Z digest=sha256:2b955b9f2075742ece14f847f320764b9679b3d9b2f8ff5d3135d8e47e4ad156

Observation 2481da52-7be7-484e-b2ff-a47dec067b2f · outbound

This paper cites Smith, and Matt Gardner.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Smith, and Matt Gardner

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.351982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.351982Z digest=sha256:4daeb2a1b0dff307359bd4829f63104379ee898a77070bcaedd59da3692c98be

Observation 9e4a7a43-0673-4a1a-8247-121571271f8a · outbound

This paper cites A Survey on Long Text Modeling with Transformers.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A Survey on Long Text Modeling with Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.355585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.355585Z digest=sha256:4806c23d92c59b4544554186d89aa967af11c79b8e347a761d6d92e5caed5a81

Observation fade82b0-368b-4e18-af54-4af488960c82 · outbound

This paper cites LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.360895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.360895Z digest=sha256:6904c6261d3926881ba97ee7999e9635784fe885965feeea2a10df160d426c23

Observation 1feec261-6e05-4325-b823-57e9383eb807 · outbound

This paper cites The Llama 3 Herd of Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The Llama 3 Herd of Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.365652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.365652Z digest=sha256:2ede483fbe70e03a0ff2b914159df543ea16f547d864541130eeb9e4285ba1d3

Observation e0edc79e-c9b6-4a07-b186-bcba7152b21c · outbound

This paper cites What is wrong with perplexity for long-context language modeling? InThe Thirteenth International Conference on Learning Representations, 2025.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization What is wrong with perplexity for long-context language modeling? InThe Thirteenth International Conference on Learning Representations, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.369420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.369420Z digest=sha256:2673cf080cdd8019e906acf43426f1f5fed82d0adca87c558a8b962360adab4b

Observation a27871c9-4476-414f-8f1b-05dfe229e00c · outbound

This paper cites Open llm leader- board v2.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Open llm leader- board v2

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.373009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.373009Z digest=sha256:b7e94946b6e559799c724b2e8a89e415eb1dad15889ef5a10dbc37351c36b13b

Observation 35d1c77e-a666-48e7-b37e-0c3d58aa3571 · outbound

This paper cites Data engineering for scaling language models to 128k context.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Data engineering for scaling language models to 128k context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.376538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.376538Z digest=sha256:74708fe23e5063a88cc46b3600e74425e0b8c638a846a9ac2e1a0069e293683e

Observation b7a683b8-e35a-4cd3-9387-fd7db4173c3b · outbound

This paper cites How to train long-context language models (effectively).

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization How to train long-context language models (effectively)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.380547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.380547Z digest=sha256:c756f89295266f74664329d83fb4d358e4da30b16102f09a09f6c5cfe4901633

Observation 52c67a86-e965-4a58-abf8-9c6271cdc260 · outbound

This paper cites How to train long-context language models (effectively), 2025.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization How to train long-context language models (effectively), 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.384297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.384297Z digest=sha256:32858ca1d48b06f9159a55c9cc83193ecc79310e75c3b8e98b6820894f00287f

Observation 065e54ff-13f2-4898-ae24-09197a77bbea · outbound

This paper cites Mamba: Linear-time sequence modeling with selective state spaces.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mamba: Linear-time sequence modeling with selective state spaces

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.387800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.387800Z digest=sha256:d324e344deb851b979eb0dac788b9f57fdabc99608489324172a6e7d351f5768

Observation a26a5ab3-843c-425d-acc5-73433f8835de · outbound

This paper cites Efficiently modeling long sequences with structured state spaces.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Efficiently modeling long sequences with structured state spaces

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.817391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.391260Z digest=sha256:f916cf0517e44b1b97d3ea5294255e9ebd08cc915dd6168de97daca4e3e6489d

Observation e9a5b1c8-669d-4cec-93d2-f760c180d162 · outbound

This paper cites Two stones hit one bird: bilevel positional encoding for better length extrapolation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Two stones hit one bird: bilevel positional encoding for better length extrapolation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.804718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.394973Z digest=sha256:723534ea058feb7cbc134cd0991a693e205563f6d456113192c05a84f17aa0d8

Observation 3997cc7e-3d58-4c9e-bf2f-796c8a1d61fb · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Measuring Mathematical Problem Solving With the MATH Dataset

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.398494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.398494Z digest=sha256:28faaf4e17b56571fe1351c01b69a70c272b0b379d84e898fca8dd22e771ef83

Observation fadf7f49-889f-474b-ba40-67b0cb1fac51 · outbound

This paper cites Improving long context document-level machine translation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Improving long context document-level machine translation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.402718Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.402718Z digest=sha256:c942710fb818f2f21c49d9fa7140298123f69227ba96c4e8b5d11f9cef18bb8e

Observation 2efbadcd-b722-42da-a661-31e8210c58c5 · outbound

This paper cites Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Constructing a multi-hop QA dataset for comprehensive evaluation of reasoning steps

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.406422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.406422Z digest=sha256:04dfbbb677aa8f36e91119588af942380668a6499a449442894f7c378b20e5c7

Observation bab910bd-d6d3-403a-a56c-6e6cbdab86a0 · outbound

This paper cites ORPO: Monolithic preference optimization without reference model.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization ORPO: Monolithic preference optimization without reference model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.409833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.409833Z digest=sha256:3489e51582607f3a495fcb657824626deb003f0bff714cbf9311c7b85d5f0581

Observation 0d0a6897-d98d-4b99-b3fc-ca98eedb7816 · outbound

This paper cites RULER: What’s the real context size of your long-context language models? InFirst Conference on Language Modeling, 2024.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization RULER: What’s the real context size of your long-context language models? InFirst Conference on Language Modeling, 2024

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.413167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.413167Z digest=sha256:71f7c33d61a0415aac0140dfcf30352d2b6b890ba087eb841f0ebb0fe7a00ea1

Observation 79ac3bf2-a78e-46ff-8aaa-f542247aebf4 · outbound

This paper cites Fewer is more: Boosting math reasoning with reinforced context pruning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Fewer is more: Boosting math reasoning with reinforced context pruning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.416410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.416410Z digest=sha256:9128270f8993a7cc4e7bd01678cd17edb9ca2cb6d4f73e574652789fa39b7276

Observation 919931ce-cf80-4df9-be3c-625b4596fcba · outbound

This paper cites DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization DeepSpeed Ulysses: System Optimizations for Enabling Training of Extreme Long Sequence Transformer Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.419979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.419979Z digest=sha256:d08481450a9bb1c37e4e572f5ed918d8f31999a82453be80fedb0860aee25b1e

Observation 74afce17-8d91-4ff3-9c03-1477b9532a83 · outbound

This paper cites Mistral 7B.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mistral 7B

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.423777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.423777Z digest=sha256:1d5d2809fd45d748de7f402562a1eb3edfbf379d8bd219810410984edff41449

Observation 10423ee0-da8a-469d-9caa-44979560b076 · outbound

This paper cites The NarrativeQA reading comprehension challenge.Transactions of the Association for Computational Linguistics, 6:317–328, 2018.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The NarrativeQA reading comprehension challenge.Transactions of the Association for Computational Linguistics, 6:317–328, 2018

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.427797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.427797Z digest=sha256:3019eed749e8950cdc14049153fac07afd57117bf19932dac62514cb2fbb6293

Observation 251ae323-563d-4083-9c3a-e87e54e9b168 · outbound

This paper cites Babilong: Testing the limits of llms with long context reasoning-in-a-haystack.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Babilong: Testing the limits of llms with long context reasoning-in-a-haystack

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.786196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.432249Z digest=sha256:78f4cb813a24ace2a304964e94aece39f1c8408400bf9a88ad4ceb649933aeba

Observation 09f00ea0-30d3-401b-ac34-e5cc7d3f1e70 · outbound

This paper cites Bart: Denoising sequence-to-sequence pre-training for natural lan- guage generation, translation, and comprehension.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Bart: Denoising sequence-to-sequence pre-training for natural lan- guage generation, translation, and comprehension

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.773554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.436524Z digest=sha256:13861222ce1ff26a7fd8434d49f741053dfa5d01f9ed75f1f01795daa4756ce3

Observation 1b5592b0-42f7-42a1-8e89-6fde66ce2232 · outbound

This paper cites Fundamental capabilities and applications of large language models: A survey.ACM Comput.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Fundamental capabilities and applications of large language models: A survey.ACM Comput

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.440458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.440458Z digest=sha256:1162df74261e70d3529127343e8ecd1f1b89f9b737f08dd9ec5514f98e3af49f

Observation 625b7171-e13e-4332-9566-8c078399fdcb · outbound

This paper cites PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:06:22.181532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.444253Z digest=sha256:8812aec1cebfedf83d9512c3f2154043f87b65c352a7b2218d168998711546e3

Observation 563bd622-0fbf-4fe0-9c15-92456885cb30 · outbound

This paper cites Making long-context language models better multi-hop reasoners.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Making long-context language models better multi-hop reasoners

Reference 39

Resolution
verified exact
doi, observed 2026-08-15T21:06:21.861204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.447870Z digest=sha256:29b313840332624c6df438fc5d45c5458077cd25e3b06455c6b65ae4b6ff560f

Observation 0bfd3b9f-87cb-4c75-84fe-eaff1d4905ca · outbound

This paper cites Compressing context to enhance inference efficiency of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Compressing context to enhance inference efficiency of large language models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.451694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.451694Z digest=sha256:7f2cf0ea20d2606a2f793c5fae7b9fe07b9acaecdd9c3167ad3dbcc66ff5ac2c

Observation ff924a2d-3838-4d0e-8fbd-5e4238a6f418 · outbound

This paper cites SnapKV: LLM knows what you are looking for before generation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SnapKV: LLM knows what you are looking for before generation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.455879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.455879Z digest=sha256:06aec0211e694e2cb8fbe4295ea82a393731c5c7c692b6140982591425193d0a

Observation ea3b987e-acdf-4b55-ac8e-e83dbb7f2fc8 · outbound

This paper cites MDCure: A Scalable Pipeline for Multi-Document Instruction-Following.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MDCure: A Scalable Pipeline for Multi-Document Instruction-Following

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:06:21.842018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.459722Z digest=sha256:73c0b2acceae5979b620f232cec99402904a7049c8a09c871140d67caa183242

Observation b63d8b18-e504-41a8-ac88-5dde435dd145 · outbound

This paper cites A comprehensive survey on long context language modeling.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization A comprehensive survey on long context language modeling

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.751632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.463912Z digest=sha256:361a6ab2dd7410b9963c887572095a457ef6527cae63e9bfbef50cdc6e20569c

Observation 97ef50cd-29cf-4fc6-9e73-e8069ca1d66c · outbound

This paper cites Decoupled weight decay regularization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Decoupled weight decay regularization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.467936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.467936Z digest=sha256:72e1e8f67c1208636a518b309b2912f747c0b4734f79f8308a596778e991cb6c

Observation 6ff09e3e-5d2c-4e24-a198-efcd65399a6a · outbound

This paper cites MoBA: Mixture of Block Attention for Long-Context LLMs.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MoBA: Mixture of Block Attention for Long-Context LLMs

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.471878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.471878Z digest=sha256:ea423067e1ba0238ad44bccd51edbcab4509db62c8365b7c73b81b044aa8a72d

Observation bb756b43-6729-4e29-a0ac-6c0d4e32f8a2 · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.476220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.476220Z digest=sha256:449fddc7afb6ad4ebd0a37cc4066b9db3cefffbc9e5dadd8432fbd413339e130

Observation 3beb0357-2152-4fb4-91ac-edfd83b20d49 · outbound

This paper cites SimPO: Simple preference optimization with a reference-free reward.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SimPO: Simple preference optimization with a reference-free reward

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.731953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.480122Z digest=sha256:a0cb65be0feea474c7b8f7806025a6c61fb84a8048f0990eb586f524b768c63a

Observation 67558fb0-8f55-41cf-b6ce-77d4553746c4 · outbound

This paper cites Training language models to follow instructions with human feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Training language models to follow instructions with human feedback

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.718328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.483717Z digest=sha256:1a55f84f2eca83287ef21f773b69de8647f1a071565a9211a100a0a43e4bc5dd

Observation d8ca97d2-d082-4236-9627-0fa09848f9f9 · outbound

This paper cites Vicky Zhao, Lili Qiu, and Dongmei Zhang.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Vicky Zhao, Lili Qiu, and Dongmei Zhang

Reference 49

Resolution
malformed identifier
no resolver link, observed 2026-08-15T21:06:21.488823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.488823Z digest=sha256:eb03ada1a986f789580734d75838f85990eeef0fe9328ee922727c614ff26057

Observation 4a933ae7-cec9-4b42-a662-a2010e2ce6f0 · outbound

This paper cites YaRN: Efficient context window extension of large language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization YaRN: Efficient context window extension of large language models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.493201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.493201Z digest=sha256:c04b0ea3a35c6f1fc814f51105b78b1b90d48cc0809d89fc18b39894dd834992

Observation a8154c0d-8865-4710-9146-983fe574b3c5 · outbound

This paper cites Handling Very Long Contexts in Neural Machine Translation: a Survey.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Handling Very Long Contexts in Neural Machine Translation: a Survey

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.688508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.502109Z digest=sha256:81c4b4ddb90e235da2fddb90092d3aadcd83e90b2be99f78d9312f92a04cd6ce

Observation de4410f3-6d63-4499-853f-f7b6bf1b1f63 · outbound

This paper cites Infobatch: Lossless training speed up by unbiased dynamic data pruning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Infobatch: Lossless training speed up by unbiased dynamic data pruning

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.676606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.506022Z digest=sha256:53fb861d6f3dad5ad84221f847d196b150eaa1d7e5b005eadb2d0122ab0648fe

Observation 28999005-27d5-4b0a-a3aa-3dc4b198dfa5 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.510415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.510415Z digest=sha256:0b64c07e6bc3ddb389e08b2939eff70e2554854e854851e16e8e273152db22da

Observation ef0fe2ee-8f47-47e9-b086-1b3a71c86f96 · outbound

This paper cites ZeRO: Memory Optimizations Toward Training Trillion Parameter Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization ZeRO: Memory Optimizations Toward Training Trillion Parameter Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.522867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.522867Z digest=sha256:01ca1de10681848dcc51dcc1e2ee3c11695236662ac725974b50ea7e29d2da86

Observation ae96f388-7aa7-476b-bf48-e23d1b7b8143 · outbound

This paper cites SQuAD: 100,000+ questions for machine comprehension of text.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SQuAD: 100,000+ questions for machine comprehension of text

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.527301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.527301Z digest=sha256:2b270feb9e9d7b1b75e4c5ba9750b17e204b19bb05417e715307d9a2e7316b6c

Observation ae3ae815-e323-4bbb-86b5-782aca30583f · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.532177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.532177Z digest=sha256:aa014b436071a79a6cfba45908e097dc03082ef0f5fb0004bbcde3350a9532d0

Observation fecc4298-f24e-4cee-9d61-3d1146fd47a0 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.658195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.537001Z digest=sha256:71cabb820673dc10e1d56b54a49ad0e101b4e3a3446774a9b69187ed0ae6ce25

Observation bf4f98c5-4af6-4ebd-ad01-b5072f34c0e3 · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.541137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.541137Z digest=sha256:c4eae45d26fa8657aeed0a352cd6dc763366498462f52bb903710a1fbe533d77

Observation 8bc4674d-699f-4116-936b-361c00f14c1d · outbound

This paper cites Generalized preference optimization: A unified approach to offline alignment.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalized preference optimization: A unified approach to offline alignment

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.647333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.545856Z digest=sha256:f0534d65ad5cd602c3129bed9962e1900d5b207e51ffa6343874674838255c70

Observation 371e2cfc-b42a-4ba8-96a3-e5a0ee3f2b0b · outbound

This paper cites LOGO -- Long cOntext aliGnment via efficient preference Optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LOGO -- Long cOntext aliGnment via efficient preference Optimization

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.549930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.549930Z digest=sha256:0973c8a78da01dc9c748f43444f8a1661262e4804a11039b86f6f04ccce1643c

Observation b2108d74-9dfe-4810-823b-e5b26ba5db3b · outbound

This paper cites Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554, 2022.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Musique: Multihop questions via single-hop question composition.Transactions of the Association for Computational Linguistics, 10:539–554, 2022

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.553894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.553894Z digest=sha256:c05d64dc73cdb113edf482944ef52ccc9a65c136a6dccc03713d241b38c26ba9

Observation 43fc8ca2-f7c8-442f-825a-1c42138e4f9e · outbound

This paper cites Leave no document behind: Benchmarking long-context LLMs with extended multi-doc QA.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Leave no document behind: Benchmarking long-context LLMs with extended multi-doc QA

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.557707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.557707Z digest=sha256:7fd50f23093cf026a8ece7eb4093db00ca3457bfea91c1892915b5063d302044

Observation 443c3306-3e0a-4667-acf9-f9919d2bc10b · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.561955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.561955Z digest=sha256:85dc434d7ed2d0026cc3ab92285fe7de86b319d1ce6676ca7c14fa3cd512cbd7

Observation 8399af25-98bb-4eb4-ae1d-1f4eca0b9a84 · outbound

This paper cites Chi, Quoc V Le, and Denny Zhou.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Chi, Quoc V Le, and Denny Zhou

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.628552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.565886Z digest=sha256:2735fac905f95ccc3721173f59bbfc9a2b4752561b4610f48180021b63236f72

Observation 6059e6b7-97e0-403e-a987-fe68453342a2 · outbound

This paper cites Kullback–Leibler divergence.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Kullback–Leibler divergence

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.615703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.569471Z digest=sha256:114c374918478b734c6f13d4bdab10d806cd87a05ce532528fbfed67c86eb806

Observation 49b05a65-9a5d-454d-8bd5-7e924200befe · outbound

This paper cites Wit and Marie Gillette.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Wit and Marie Gillette

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.604784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.573105Z digest=sha256:7d157ad139fe952a197da4ff97640c00bd930f6ed2f1aced065a1d3cf9f7b455

Observation 505c72a8-e97f-420f-84dc-ec67e3eae96f · outbound

This paper cites An efficient recipe for long context extension via middle- focused positional encoding.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization An efficient recipe for long context extension via middle- focused positional encoding

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.594373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.576890Z digest=sha256:c4f22034f58b02bf0343783d66bde8b41e6becf6128cd328d8b75e1ea49a6b1a

Observation 414a2203-ee51-4946-b2e3-9f5bb1264a36 · outbound

This paper cites Effective long-context scaling of foundation models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Effective long-context scaling of foundation models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.580733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.580733Z digest=sha256:b201dda025cf16b77a58d162fe2e790cf595f00602f2c36674d1309d64592047

Observation 8cce9127-4172-4fea-83e7-ff025a3d6794 · outbound

This paper cites Large language models for generative information extraction: A survey.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Large language models for generative information extraction: A survey

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.584357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.584357Z digest=sha256:a3232ba79ee0889a00bf7eb88206c3c2f4ee942078539f1c26555c8edd807197

Observation ef386735-cb16-414a-9306-7b121d41d857 · outbound

This paper cites RECOMP: Improving retrieval-augmented LMs with con- text compression and selective augmentation.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization RECOMP: Improving retrieval-augmented LMs with con- text compression and selective augmentation

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.568096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.587957Z digest=sha256:4fb585ab7cc213446ebe26dc5fa8113811c4a331780dac46a360baaba73ed885

Observation 0832a35e-eece-4803-b931-2251f5ef98d8 · outbound

This paper cites LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.591936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.591936Z digest=sha256:a9ab527eae656d2e11fdb4f325f91de11c4256e9047e5d66d5da13e8ba82e970

Observation 1a79735a-77f7-49a6-9ef7-68c09f2c17a0 · outbound

This paper cites Qwen2.5 Technical Report.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Qwen2.5 Technical Report

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.600241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.600241Z digest=sha256:2382a59f9605a67764f8682d74dd1ffd0e1cf7769a509045e0690982ce7b2740

Observation ccc194f9-4edd-4fe2-9940-e9dd6ed15fcb · outbound

This paper cites Mindllm: Lightweight large language model pre-training, evaluation and domain application.AI Open, 5: 1–26, 2024.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Mindllm: Lightweight large language model pre-training, evaluation and domain application.AI Open, 5: 1–26, 2024

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.556789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.604178Z digest=sha256:2189a994db48f105120ebfd165a514cbd302e887b038f991799897ae65dfd21a

Observation 35f54b63-142f-4928-b67a-47f4b80787f3 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.608200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.608200Z digest=sha256:feffca4ccd433c1cc5a9f41dab60d6b9f76cded7cc4fee945ae0d72c64a5a1b4

Observation 5384a0cb-441d-4438-94b5-26ccc5be58c8 · outbound

This paper cites Longcite: Enabling LLMs to generate fine-grained citations in long-context QA,.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Longcite: Enabling LLMs to generate fine-grained citations in long-context QA,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.546000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.611974Z digest=sha256:eb3c8dfaaf01732e8948058670e59e09f0b604a424273abbc4c7170cc1b1d131

Observation 44de6843-0d4c-4884-b833-33b6f39c6000 · outbound

This paper cites LongReward: Improving Long-context Large Language Models with AI Feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongReward: Improving Long-context Large Language Models with AI Feedback

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.620099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.620099Z digest=sha256:db38fd1108a45a3e87ccb5622feea6376f339233bc3704ebe1f15689d03186d6

Observation 75363c18-7a12-48e7-97fc-8323d8c71b60 · outbound

This paper cites Extending Llama-3's Context Ten-Fold Overnight.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Extending Llama-3's Context Ten-Fold Overnight

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.625023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.625023Z digest=sha256:80036658c05e582cb062d1ec67f08f2063a8b898ff922d35b6b8a08e140f6f5e

Observation 23f5ddc1-134a-4958-8589-54d81ba8e1a5 · outbound

This paper cites IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization IOPO: Empowering LLMs with Complex Instruction Following via Input-Output Preference Optimization

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.629057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.629057Z digest=sha256:76b8f14f160b47963181c07555b67c5c02f194018b7322a098aa7fe5d8e05b7b

Observation 6b168475-3b2c-4c08-9593-503be193790f · outbound

This paper cites LONGA- GENT: Achieving question answering for 128k-token-long documents through multi-agent collaboration.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LONGA- GENT: Achieving question answering for 128k-token-long documents through multi-agent collaboration

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.633192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.633192Z digest=sha256:63fd96632d24eca32c6b887367311ea322d0482e471aaf308a1b5d4f3c242109

Observation 98b2e85e-9240-4041-a61b-dc25b16a9a96 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.534771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.616025Z digest=sha256:e1c8cae68647d32d4dff831994bfb6db1c110458865e103bd1cbc5948bf7c335

Observation 817db5e0-9104-480d-8ff4-f2ed4f5e0d81 · outbound

This paper cites LlamaFactory: Unified effi- cient fine-tuning of 100+ language models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LlamaFactory: Unified effi- cient fine-tuning of 100+ language models

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.641958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.641958Z digest=sha256:5caeef33509c44f6ec13a91a0374f5134f97452c5335975bc0b1c764cf82b999

Observation 568bcd0e-55f2-4f65-9c02-a17f075660b8 · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Instruction-Following Evaluation for Large Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.646243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.646243Z digest=sha256:bf5e7177a7857c781d66ba18ab8043c5eb963621147012ff741651bb8a90e50f

Observation bf6342b9-9d05-45ab-901d-b49bee060473 · outbound

This paper cites PoSE: Efficient con- text window extension of LLMs via positional skip-wise training.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization PoSE: Efficient con- text window extension of LLMs via positional skip-wise training

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.523288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.650411Z digest=sha256:3b8d1ead0ef454733cff7aade12464946a8f85ebadf14b7366306b812b2eee19

Observation 08d217a7-7671-4793-ac30-d7708bc2385b · outbound

This paper cites Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.653955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.653955Z digest=sha256:7a2c5f51e04cbd8ccbbe0c348bc38a16ce0b9d2151f56e0f018317aba24430bf

Observation 9f678ab7-fc80-4070-93ff-4800f562e6a4 · outbound

This paper cites SLiC-HF: Sequence Likelihood Calibration with Human Feedback.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization SLiC-HF: Sequence Likelihood Calibration with Human Feedback

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.637749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.637749Z digest=sha256:51a1b002c7c0e29ca1d8dacb863c0330e684d96e6e7918104be0542279e913e3

Observation 4150fec0-92d4-4e75-aa66-4825d206d802 · outbound

This paper cites Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.668121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.668121Z digest=sha256:6cb1a88f9b80091b9cb75660712f1f9587aa2b4a39c98c096f92dc66dece8d5e

Observation 38dae01b-6055-48ce-8958-1fe75512ae93 · outbound

This paper cites Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Generalizing From Short to Long: Effective Data Synthesis for Long-Context Instruction Tuning

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.658781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.658781Z digest=sha256:efcc0789394870ce28eb510e8a1c723934b9f3cc459493ba429acff63991cd76

Observation 23949f5e-f084-4d8b-a964-69886c64b526 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:06:22.512210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.672749Z digest=sha256:6f7a591c28aaf297cd8a599f6936624f0c2894dafa00f0214a434ec8dfd1355d

Observation b188141b-00ba-4189-99af-e8a3fb7bd7e4 · outbound

This paper cites The Collegian.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization The Collegian

Reference 94

Resolution
malformed identifier
raw_fallback, observed 2026-08-15T21:06:22.019293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.676766Z digest=sha256:a0356d628f820bf3ab58b04a3799f2548846a63a384f33f0b77add0cdbd0a6ab

Observation 2b25102c-228a-481c-b64a-cecaef3a30c6 · outbound

This paper cites In addition, we further examine the theoretical validity of the first motivation from adata-sampling perspective.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization In addition, we further examine the theoretical validity of the first motivation from adata-sampling perspective

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:06:22.499465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T21:06:21.682241Z digest=sha256:6d710576ab00acc6d923db25b3a124aeb335602f8b88da4448f26bb05a87bb52

Observation ee0517ce-950d-4a4b-8084-6f15a5b410b5 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.304223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.304223Z digest=sha256:298588e9e0dc1a4ceb349440de9d5055b3252037dcd6b2e1b4a40b6be52479ae

Observation 64d86e4e-55c4-40f4-946d-c82fdcce0c1d · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.514589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.514589Z digest=sha256:af8630206d725d31b4deaf65e68898900832e657129a36b89bd4a52e85686623

Observation 15c472a3-c4d3-4cdb-b69e-181134c4aa87 · outbound

This paper cites an unresolved cited work.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization Unresolved cited work

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.497385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.497385Z digest=sha256:e500d378919b4dd99ef2c9dadf84ea076e2fd991ece793eb6bc21d539b6ec175

Observation 23a3064a-7101-46e7-a0c4-d54932ef993f · outbound

This paper cites LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data.

SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization LongFaith: Enhancing Long-Context Reasoning in LLMs with Faithful Synthetic Data

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-15T21:06:21.595959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:06:21.595959Z digest=sha256:d20546d28b7385b7307bb71ff2f44e0a2f1750e8156bfc0d4a0f0267fcd4ae58

Pith citing papers

Observation 401592e4-5322-4dd6-81c0-5d0a9dcb878f · inbound

OPSDL: On-Policy Self-Distillation for Long-Context Language Models cites this paper.

OPSDL: On-Policy Self-Distillation for Long-Context Language Models SoLoPO: Unlocking Long-Context Capabilities in LLMs via Short-to-Long Preference Optimization

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-04T02:07:03.503226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T06:07:36.830550Z digest=sha256:6dc107c2078abd59d01fb53a33b182b6ff2aabc268852fdd0a54f3a86b022fe9