Pith. sign in

Paper Citation Record · LEDGER

LIFEBench: Evaluating Length Instruction Following in Large Language Models

As of 8 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 4 inbound Pith citation observations for arXiv:2505.16234.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16234 v2

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:08:08.453311Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T15:28:14.401581Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T00:39:35.853384Z

Reference resolution

100 of 129 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d03fcbf4-d152-4577-8b0c-e084be42cdbe · outbound

This paper cites Abedi Firouzjaei.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Abedi Firouzjaei

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.915175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.915175Z digest=sha256:aaa88d9dd90e8a3495305ebfcadfa24740807ac0e299e8c19e21591fa3350e52

Observation 52722d46-9a3d-4a13-9c58-5dec7b1fdc4a · outbound

This paper cites Alzantot, Y.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Alzantot, Y

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.980923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.980923Z digest=sha256:fe5c68794fdd53dc174d2605b4fc8863c4e2195d29472969b3fc7e7de9fa4331

Observation d62c8ae7-3dfe-454f-bb53-6b6025bae52f · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.100922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.100922Z digest=sha256:d8a18fa31c56322fd0a896326050b57198c970dae1f9a9b7c6e8d986f585e5c7

Observation 12609bb8-f46c-4f25-88c3-f655ffd5fb18 · outbound

This paper cites Claude 3.7 Sonnet and Claude Code.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Claude 3.7 Sonnet and Claude Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.243342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.243342Z digest=sha256:9d98c456bbb53e86669829b246a7be366366b05a3b8221209ebebb4e6d1760c5

Observation 75e7b2b1-f57d-478d-8d66-009a1e343575 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.370074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.370074Z digest=sha256:7a2c69c7d5d42c46cdf168e1773389ad4273e06b8ca016d2818d425bd744937c

Observation 76e6cf04-00d6-4840-bdfa-52c5db7dd6ca · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.482208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.482208Z digest=sha256:bd9e0dd72275f081b31223a27dd45e4392cf53731ebd165b5bf817aae4d962c1

Observation f427873a-8f62-4dc0-91be-8d9997e3035f · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.616769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.616769Z digest=sha256:799db9eb7ab7c6ab1d207564f6f4be607d3b5262d15386394003c0a34c82bd42

Observation c8073588-1d7d-4018-a79e-6ede488102e6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.759419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.759419Z digest=sha256:5451c6c641dcfcc8005256075dfe97b454de8ddad129485b9e4c8b44fad88cd7

Observation 85558c24-4962-4f40-9ab3-7f563662fba8 · outbound

This paper cites Bordes, Y .-L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bordes, Y .-L

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.923158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.923158Z digest=sha256:50462ce26d45875332d86fefb3ea5c60d470f735b42ad3f55561a68bfe6efa18

Observation 8a000c4b-be89-4539-878b-9cbd7123646c · outbound

This paper cites Bosselut, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bosselut, A

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.022186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.022186Z digest=sha256:22e8a1b3fc4ce85363e138fb84fd33ae7db8bd4696bf909d9ba85226b01588ef

Observation dee3ee11-3c31-4520-b503-0d57e76e40e9 · outbound

This paper cites Butcher, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Butcher, M

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.144208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.144208Z digest=sha256:b40ecda63a71dbcc9ab66a84fe1eaaf8abc7a0c608e2229e8a2738f0bb0bef03

Observation 0cec495e-48db-4aa0-aa49-185bd1bb0555 · outbound

This paper cites Doubao-1.5-Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Doubao-1.5-Pro

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.233200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.233200Z digest=sha256:da8b5dc0bbd1be597c148e528402c118b6b113d42b57c27374d808e7b9ab6334

Observation d83dcc59-f853-4d12-ad7f-9fb3c1bf25dc · outbound

This paper cites Chang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chang, X

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.312710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.312710Z digest=sha256:3d4ceba40e00d9a60d009505e477799ff882d72f59e007ebca19f9c506e2330e

Observation 2845bc06-2ac8-469d-a358-83afe998016c · outbound

This paper cites Mistral 7B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.389348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.389348Z digest=sha256:9b6f1f5e8d38f9e5f6ddc639c581301450edff0a34923eca26557add4e924f20

Observation 3a2e6bd6-f44e-4228-b0da-817dd6009ed0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.475747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.475747Z digest=sha256:19995abb94f48d699381b1015263326cd29f236c8793ac1eb91f68f3ad69fd9f

Observation 1b9c43a9-c31c-4586-b9f3-67543687f7d8 · outbound

This paper cites Chen and C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chen and C

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.576145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.576145Z digest=sha256:c619020c40af1e52b39dc2a198aeb74abdc1e74ebb534d4af9b4a9a676efa87a

Observation 7b252921-98f3-457a-b125-ee69f5ac761e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.643294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.643294Z digest=sha256:943027ea155b38572b53e1b35b76b9a88612ab6ea7aa1049a029534aff44ea12

Observation f5bcecf5-7db7-4429-9b1b-7a332a67e84a · outbound

This paper cites Chiang, L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chiang, L

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.723680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.723680Z digest=sha256:3d2db80481dd2e3572e72019b1776b69562472de012bab354440f317afc920cc

Observation 9f64694b-b587-4229-9e0e-87e2b99537da · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.821582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.821582Z digest=sha256:eb5e137d89e6e7e065eedf072225ac0a1a18eb0a7569ffa4594c876f6702204b

Observation 9ee681a9-1b69-4f20-84b7-9353dea5e07a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.907072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.907072Z digest=sha256:6fc84b730792e982dfa329decd1c912d4bf307395c43770072b02d64f4bdf2a1

Observation 5eef2f6c-3b9a-4e35-bcc9-6219b1c0f537 · outbound

This paper cites Cohan, F.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Cohan, F

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.090683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.090683Z digest=sha256:c6e73e69419fb196b6eb5300b93eaf9e431f738feae1a41370268b4a70baffc9

Observation b8a754b4-d26d-4973-8e81-433936305d43 · outbound

This paper cites Collobert, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Collobert, J

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.249374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.249374Z digest=sha256:aeda88dc46ddfde142f5e2b0bdc123cecd726224733c5727a1b5e94340721264

Observation 642495c2-3a0b-4a7a-9b68-31d725882625 · outbound

This paper cites LCFO: Long Context and Long Form Output Dataset and Benchmarking.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LCFO: Long Context and Long Form Output Dataset and Benchmarking

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:08:12.116541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:02.336424Z digest=sha256:77189e7091a07d615753e3285de50b40fb61bbc839390af2f8a327321c64fa58

Observation 8575674f-54ee-4641-8746-f565c54fe5dc · outbound

This paper cites Davidson, D.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Davidson, D

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.496551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.496551Z digest=sha256:178e9fccd783544242aef588c734e2abdc8c7c06f6f461fe19c06104efe799f6

Observation 264207e3-a4f7-49e7-9fea-48f19a70c967 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.698130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.698130Z digest=sha256:4399b494b1693b94ab448cfd7cc48492ffb6314729dd512a93fb7dd67fd424ce

Observation 207f2d8c-d40f-4223-8f31-f95caddd5bb1 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.802654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.802654Z digest=sha256:23247e0b678bfba23ad4a273eda9cf3311ad06b189a4ea3c824233c02711ba76

Observation 348202ab-4617-46ca-b372-c9179b218782 · outbound

This paper cites Dubois, C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Dubois, C

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.954068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.954068Z digest=sha256:f3c5c68e96fa5a8f61fb1db1843fea486443233bd9fe835aff473a13cbb83154

Observation 74eced30-2b84-4219-8b65-3311a6304d66 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.106742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.106742Z digest=sha256:5e4141312a7d572d9fda7f4d282d819563744c64e891a14a63b5896a43328ca1

Observation 29979e30-9658-4795-aae9-fb2f03def193 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.246808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.246808Z digest=sha256:792f02e2b098e2a1b214f6ad8a651b4a3947213bee659272d8e9bd345ff0d961

Observation 02d05053-243d-4d2a-9694-2ef27cfb62d8 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.356120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.356120Z digest=sha256:9afb9ce44ad0e55e752d6dd1761ac214aa990e2bc6a6c863dca451576ffec9da

Observation 7520112c-1822-41e5-aeb0-119846f48bbb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.507635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.507635Z digest=sha256:17b946362618d9b5f2a60caa1cc016b746a430c1f5c35429a05bca7851fe9b5c

Observation 1f012298-b49c-4271-9aa4-65d5d67af0c8 · outbound

This paper cites Foundation.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Foundation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.619331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.619331Z digest=sha256:ddab47ec18c2921a4593ef486b02dc9d807df69aa94a150bdcd55e0a140797dd

Observation c4b3fe12-b0fc-4900-b9c7-ee56d1e7a38c · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.761610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.761610Z digest=sha256:04b1c53d3734ea70417439c8a4ac7570a6fee5ce588e9b9527e56fb149a26b3b

Observation 31377b83-3942-4967-b76e-687b8a78e71f · outbound

This paper cites Gemini 2.0 Flash.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.0 Flash

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.821979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.821979Z digest=sha256:d964d76bf8b0328d08f1e032b066ce57a007ef0312558befb575edcc241e3c03

Observation ed19b53f-3d7b-4290-8c8b-879e4119f478 · outbound

This paper cites Gemini 2.5 Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.5 Pro

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.877823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.877823Z digest=sha256:dd325787079f8faaf1773f1f16d57d056ac7ce645f177ee80f1b660ce69e2b09

Observation 9e937a04-a837-4446-a22f-80a85be20f3b · outbound

This paper cites The Llama 3 Herd of Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The Llama 3 Herd of Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.923252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.923252Z digest=sha256:93784cdbdb765366ffd2123c7017fa1673885fef17ebbce01dcd5ee36ac6563f

Observation f19c070b-17d7-4cfa-9a32-0a7e8985ddd5 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.998056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.998056Z digest=sha256:2fa22245503d6e1bc8bc22a64c2e944bb90c1be1604bcd48fe00386c79cb62ab

Observation caef362d-1e6e-4407-81dc-227fe3ba992a · outbound

This paper cites A Survey on LLM-as-a-Judge.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Survey on LLM-as-a-Judge

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.044070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.044070Z digest=sha256:9c54ee82510e44813bc065069f503efda801834837f1587fceb641870de5c5d8

Observation 0a1bbbcf-23ea-4ed9-9ec2-9ab31402e39f · outbound

This paper cites Length Controlled Generation for Black-box LLMs.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Length Controlled Generation for Black-box LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.110426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.110426Z digest=sha256:dd234a037979170defcd5ea17899b7e16d54a7accd117b16d58e069b4b41434c

Observation 6e458285-5368-4628-a9a7-9cfa31d92340 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.157989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.157989Z digest=sha256:24ed0ae80f39a360425ba42cd74b1072188224296db0924aaea8f815cf43e4c3

Observation f3ab0426-6758-4aac-bd61-241b9d68ca63 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.207512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.207512Z digest=sha256:9096775e26bddcb1c1a58ef5eb3fe3bb73a41c92766a498cbd870a237ea6f316

Observation fd5c61b4-e6f6-45aa-8607-9abb2af7ac50 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.244676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.244676Z digest=sha256:1cbecca4e3b4e1d028323dbb344154250a32b1640ad38bb051f75117d7b1a4cd

Observation 48bc6929-6b5d-479b-b24e-28ae7ef193cd · outbound

This paper cites Hsieh, S.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hsieh, S

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.314289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.314289Z digest=sha256:830795d3ed96bea75b2f8df87109e9eb5a684c9e0736fb539a5f9f263632479a

Observation 0e770906-3b15-4109-8d83-d733cff41d74 · outbound

This paper cites Huang and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang and K

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.379241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.379241Z digest=sha256:7846bb4152ce907dfb0831653d16bea13210a92cebd801a09f976003491329b3

Observation eaee9ae1-cf63-48cb-84a5-8bfc7eedc978 · outbound

This paper cites Huang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang, X

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.438908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.438908Z digest=sha256:77728cca50d264cbe18a126a8b59e3b7c74570ad9807f5f3341bb98968208c67

Observation 5bcac5cd-9a07-4a6b-92fc-abaa0f1faeae · outbound

This paper cites A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.497560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.497560Z digest=sha256:b58b60c5cf14ccfd6001023456d9bdbacb893d312e1f0cd96c055069b6ffecbb

Observation 9b0ac765-2a37-4dd1-9ad1-3c987da4f752 · outbound

This paper cites The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.558589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.558589Z digest=sha256:60bbadad79c91609e99fdd12c24b4d6d718f82193a1c672fbb180fd3f23034c4

Observation 5f1ff18e-cd0c-4fee-8eae-5582c8694fa2 · outbound

This paper cites OpenAI o1 System Card.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1 System Card

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.604933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.604933Z digest=sha256:208d0003539fa2ccee05be736766c1da82fc864640dc705e922844b6fbe2ba5c

Observation 67340472-1885-4a2d-8bcb-ab51bfa796ae · outbound

This paper cites Jhamtani, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Jhamtani, V

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.656826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.656826Z digest=sha256:10558c647d6ba833a4ff10018150bb0a9dbf77b12c3db12ba5c0d27fd238ee0b

Observation 4201875f-84b9-4362-abea-85aa730b65fb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.700834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.700834Z digest=sha256:a5ae76d9b84eb365fb4c9516d285c00bf9a3b9c9d2d6b4adce14d8c413db7131

Observation f8f43d67-01fe-4215-b405-3477faba6465 · outbound

This paper cites webnovel_cn (revision 745338c), 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models webnovel_cn (revision 745338c), 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.773018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.773018Z digest=sha256:12a76d8265feb0099a5448047e7780387eb3e9d8d5345202ee12320d999a665f

Observation 01265994-1f64-48b2-b6cb-afa6fad5df13 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.828592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.828592Z digest=sha256:62c136685c3cd1c2229cc3542e565e9a862c400df298ee07832ef9c1fef06396

Observation 57ab9db9-b0ff-480c-b1f8-5985cdbab959 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.891231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.891231Z digest=sha256:2cb05e57b98320438f4d7f70db417be17275c7236cb83483d05b58e54c2b91f8

Observation b05c97b9-1322-40df-8b17-02e580e2fcc5 · outbound

This paper cites Koupaee and W.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Koupaee and W

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.954876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.954876Z digest=sha256:f09c8884f7fb88842d028d466bc6a023093ab9085583b0cd7598ffd79bdef685

Observation 0a4ae45a-d174-4dbe-82ed-da489038619a · outbound

This paper cites Kry´sci´nski, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kry´sci´nski, N

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.023199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.023199Z digest=sha256:c17e94da5fbeb341953ad42d0329badfaa32ee4ea49823d6e7288c13b8fdba41

Observation 2ce96cb9-3ce3-4822-8fd9-f5373e7b8f16 · outbound

This paper cites Kuratov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kuratov, A

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.089583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.089583Z digest=sha256:d743277b77bea708428d1e0bc9d29a2ab6f42ccd315fe0f9fa19cca018c6516e

Observation 06ebd402-d64f-402d-a331-cf01a447796f · outbound

This paper cites Lample, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lample, M

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.137728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.137728Z digest=sha256:d4447b9d990abd0c437eacad880a11c57fa183a709f06954c293d0c0a8e97714

Observation f5181dd5-d25f-48fb-83d8-73d647c30973 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.197847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.197847Z digest=sha256:bbf0ea75b15d82236a29384c42626624c371e8db7c9d6dc0cc51ed1cc98f7378

Observation 49645e82-b652-4441-9933-cf3cd5d6f730 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Long-context LLMs Struggle with Long In-context Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.259340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.259340Z digest=sha256:b277ce609ea25eca8815623264104389de9a0bc573ed8e87e67d92debd91f1b7

Observation 9555014a-566b-40ee-927f-818b14f6d62b · outbound

This paper cites AI Awareness.

LIFEBench: Evaluating Length Instruction Following in Large Language Models AI Awareness

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.339186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.339186Z digest=sha256:1c4067393f33d19061122f20e6d51af2d3a7445805d27b305f66ab287cf24044

Observation 24f626ea-adfb-44ea-94c5-37195056c79d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.453724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.453724Z digest=sha256:8a3edf232d735328bc37f4c61a3cee58ad8463e52655a271bae72131581b4841

Observation 4d97b2c8-357a-470f-8ada-2521b0871e72 · outbound

This paper cites Controllable Text Generation for Large Language Models: A Survey.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Controllable Text Generation for Large Language Models: A Survey

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.544649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.544649Z digest=sha256:32c7f7b0e02e36a539b64d326e6f2e6f195860a82b4502e305e7baa5d6bc9a4f

Observation 730f358e-1ce1-4107-bdd1-0a7bb8fc9a1c · outbound

This paper cites Lightman, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lightman, V

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.644559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.644559Z digest=sha256:28511ffb590b449d6eb76ec324e1874bf698d59c21fc6f46da2c78049e4fbf70

Observation bda9e038-2d63-4e11-90c4-3171e0cd5aa6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.718550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.718550Z digest=sha256:42d7e5468cff4f5799a3c15914f5c854642249f6f112b003161604d05440c65f

Observation 2fe3fbff-a7f4-4f7c-a39c-f29a07718948 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.827379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.827379Z digest=sha256:876f752dc717660a7376607194015087b492751da6b3e0026931d4dd351e7e76

Observation 9f8e5921-e614-4fe2-9239-c1d4d2eca7e8 · outbound

This paper cites DeepSeek-V3 Technical Report.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-V3 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.916363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.916363Z digest=sha256:2bc6a40cd4f8cb70fe51422d7caf458d0cb6f2a8212de449c01ba22653358c7a

Observation cd80fcd9-43ab-461a-9333-a5737480603a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.968830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.968830Z digest=sha256:179182363ecabff7d56ce89128c628a90edccd769e7fcea8f353dc3daaf7d024

Observation a8ce4bf1-95c6-4c87-aaa5-dc73ae8255f3 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.040972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.040972Z digest=sha256:4fd1bc4183971016f8dbb369e0cfd6415bd2898e73261da799cd421467779b7a

Observation f44b67bb-10ee-43a8-ac99-29f04933732a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.965757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.138421Z digest=sha256:8b4088a5b7adca3aed1937309bd5be5fdaceca3065c7076d0257032179675bce

Observation fe4da9eb-02ca-4ea7-9620-ebff6fd0f76e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.237974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.237974Z digest=sha256:3cd4c2a5a80c28747750f79c2f06555d28535e25aecd6016f7db470020d2b7be

Observation 5380913f-d36a-42ef-b2c7-db46a4ade9d3 · outbound

This paper cites ExpertQA: Expert-Curated Questions and Attributed Answers.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ExpertQA: Expert-Curated Questions and Attributed Answers

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.311393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.311393Z digest=sha256:05172e1bd437bd02eb4007336daf2ba40943e7d0fc90d189d44384a9d63a7aa4

Observation 76fa2654-0be0-4af4-8c2c-89fa5fe9345d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.823549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.397371Z digest=sha256:605ff60633454a92711fe5f6cc8f2757f5fbf8bce9b3b78f497dd23d5236e5ef

Observation 69b59cad-1a91-4d25-a19e-dc830b707bc7 · outbound

This paper cites Chinesenlpcorpus.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chinesenlpcorpus

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.743188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.513962Z digest=sha256:6d7c0e81a35b1b339e624f62d5cd6971d4e2decc927f3334c80eb9a773f17866

Observation 24b62dbc-691b-4772-9348-0ed39bdcd484 · outbound

This paper cites Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.639920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.674885Z digest=sha256:e9f3080cb3982e32d2c0ed8b1216d75264eb22cb74003c5b19c5360b2548e028

Observation 1173264d-9c54-422a-93b9-da8a23ce010e · outbound

This paper cites Mostafazadeh, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mostafazadeh, N

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.537826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.786681Z digest=sha256:c4562e5918023742e84bbc285816894c62c48c7f55e0dd16e35480daf6874b3f

Observation 85b58e44-50d7-4733-8ca1-173a4139db81 · outbound

This paper cites Nallapati, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Nallapati, B

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.369749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.882313Z digest=sha256:9eb9ebdc4f7dfeffc6f806a82e3677c580519c231dcef64de351dc3e5a58fe21

Observation f6c223a5-af30-4fcf-a99f-2bbb5af790b9 · outbound

This paper cites GPT-4o mini: advancing cost-efficient intelligence.

LIFEBench: Evaluating Length Instruction Following in Large Language Models GPT-4o mini: advancing cost-efficient intelligence

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.180851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:06.933239Z digest=sha256:5c911f07297a80f66160e25f89850e4244b640f1ca9f5e1875712a56f987a7b5

Observation 0dd40658-09f2-4836-9104-014091f5039e · outbound

This paper cites Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.032956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.032956Z digest=sha256:0df8ec5ba36478d3b26a95b7da4c0b9429161cb0ec4f48ccb6bb77f84d8de139

Observation c04d3ce7-0fc1-4c56-890c-8de0610a4e57 · outbound

This paper cites OpenAI o1-mini: Advancing cost-efficient reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1-mini: Advancing cost-efficient reasoning

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.018525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.136231Z digest=sha256:7632fc566906a7fc5f1d1e1068e9111fb5b7428c03dfc2154b5085bff7512350

Observation a5e6597d-4497-42ea-8a4e-8e43c9b6123b · outbound

This paper cites OpenAI o3-mini: Pushing the frontier of cost-effective reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o3-mini: Pushing the frontier of cost-effective reasoning

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:19.781683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.212823Z digest=sha256:4666629270d815cc87dc51da6ae278b6df6d80fa743c3611613665a018c2c1f5

Observation e2b1c336-0ac1-4c52-8c46-6131af0bed4a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.563974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.302010Z digest=sha256:0d27b8caab7a8d125d2d7a3968306b26eedfae1ce4afe2d18b25365fc61db4b1

Observation 5f1e47e0-eb8d-4b31-9ccb-2af426340b81 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.417697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.421996Z digest=sha256:cd85ca8cf267729f4006e7113a2c08eee47821556f69c9243fcc8cd9a3ad9d8d

Observation 84f16566-6a20-4804-866e-4ec8307b499e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.248048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.521225Z digest=sha256:619c7e8b838dcf33ff1203c6a7304ca7618790d934adcb94c75dcda04e53433b

Observation 5a92f655-6357-4763-aa38-2c13fc4c6f1e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.112561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.583818Z digest=sha256:f23a8debf52be208aac0f7317ccb30e752ac15ff0a15b77f42fe036adb417cf5

Observation bb0e9050-1346-4443-b9b1-b22197d61123 · outbound

This paper cites Language Models can Self-Lengthen to Generate Long Texts.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Language Models can Self-Lengthen to Generate Long Texts

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.651762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.651762Z digest=sha256:c73eb2e0784aadb2ecbcb1577fed09e7b3d8fd8f83c568b250c2b998eff1b364

Observation 27da7725-b122-4352-8b43-795015fecb70 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.698689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.698689Z digest=sha256:05708dba6add7b3a714b41453ed61e11071267400843e9ed310a50b14d4fefbc

Observation cfa6af53-cd98-4c87-8776-4bde43e9b3ee · outbound

This paper cites Radford and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Radford and K

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.997721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.792526Z digest=sha256:69a83d4f2250a36af949a42dc6c902665dcb9a22b63fb7c735c2d4765fdab63f

Observation ae6662e7-ee89-48b2-be05-5437d31a5453 · outbound

This paper cites Rafailov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Rafailov, A

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.862390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.833724Z digest=sha256:5e90db0e9b78bc61a76d14f70efcb728d5b87e11cff1d382801d6469bacc8de6

Observation 91c869c8-9537-4c21-a830-09f45af784a0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.746199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.872975Z digest=sha256:849646f0a6e82bd51b2d94bc87e33086f7029a39c184f2200ed10d8773212b44

Observation dd45e7ff-fbd0-4514-b49d-f0ff703a619c · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.635380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.912824Z digest=sha256:af98777c80acf0b8864a8457a59a110659f6fa8408b7ce7d2df93a4e22c16c98

Observation 84318624-9748-4bb5-86af-f06cfb3c21e0 · outbound

This paper cites Sennrich, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sennrich, B

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.512717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:07.978742Z digest=sha256:d6d18978d14e916d89d2c5e79d894b78606ef70a0adfd8b1c3927b640ce2c83f

Observation 75dad7df-64c9-439f-ad22-0c65def8afee · outbound

This paper cites Shaham, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Shaham, M

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.362523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.027546Z digest=sha256:cd0caeb6017defa74a1fc365c4ad205f2fe23aff703989dbc35d9856afb90d49

Observation bc065065-870c-4664-b9d5-a5dc81738cb4 · outbound

This paper cites Socher, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Socher, A

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.204207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.073224Z digest=sha256:a699b400173f3abf5ca1bf3994c4e1b2f5c2d4aa03484ee33e3d13c4c95a1e99

Observation d8ffbb68-2cad-46e2-b7a1-28e802e14dff · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.065546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.114897Z digest=sha256:7b3928f18e106cc5e8aadd90af37de45c076f6b971b7009435d88190d13ea572

Observation c2cd78d9-fbae-46fa-9bd9-170b91d2a54a · outbound

This paper cites Sutskever, O.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sutskever, O

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.152779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.152779Z digest=sha256:61740dc1dc0cee1add4c3c6f372b1baad10d00283f3e3972502de5e8d925bed9

Observation d8b7c369-252c-4771-ba80-ece7033421e9 · outbound

This paper cites Talmor, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Talmor, J

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:17.933821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.199068Z digest=sha256:324f99870610e4e93015701b8445a036db8410d708290f45789d5a250c94e13a

Observation aeea2fa9-9cbd-49b9-8125-dfbc067f22bc · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.705498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.252432Z digest=sha256:a9eac7221cb9a8660fd8fbb0a201f5c1bc794b033a7b1000251783d196e45152

Observation 4b274721-20f3-4893-9441-212289a8c5cb · outbound

This paper cites CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis.

LIFEBench: Evaluating Length Instruction Following in Large Language Models CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.312320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.312320Z digest=sha256:9951ea444984289924879fcbb1b1b30c71cd634550d26dfc1c1dbd1272d46cba

Observation fc9de4cb-0846-4cd2-a62e-5f5ab5bb5382 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.405637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:08:08.393222Z digest=sha256:d6757f574e1bea09c3de737527dcecd207c8a7fed80a7bf34fd82efeeabf48bc

Observation 617e84f1-45a8-486b-bc09-a595734d39da · outbound

This paper cites A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.453311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.453311Z digest=sha256:080d13aace081e1655911dd7ed94d55dbf65f6772dfa288b14943e04336ac395

Pith citing papers

Observation 30905453-4f45-4e4b-b4e7-d3288f6a18c7 · inbound

TiCo: Time-Controllable Spoken Dialogue Model cites this paper.

TiCo: Time-Controllable Spoken Dialogue Model LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:39:35.854869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T00:38:52.182973Z digest=sha256:3040a819a8ea029a7313842505c7f1cfa73246b5f066099f21903ea7a1ec43c8

Observation eacdae65-3aa9-48a7-84d2-fe4f586cb22c · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:31:26.140879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-07T10:42:27.644514Z digest=sha256:ca1888e771e3c6e37f2be331ecd552aa54390609198e838c974f6875634cb99d

Observation 463c613c-2a32-44e3-bfe1-3cbabdb90551 · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T15:28:14.401581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:28:14.401581Z digest=sha256:50f83a1d983b2ed8fe119835c3d522ef98db23bebd8be7219533d65e038dff6d

Observation 9003ed95-f141-480b-8560-443b02ae4816 · inbound

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation cites this paper.

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:03:36.857585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:03:36.857585Z digest=sha256:38b18ffda69eeff28b20745eb80c23185003970e014f2f257d356c6f344795ca