Pith. sign in

Paper Citation Record · LEDGER

LIFEBench: Evaluating Length Instruction Following in Large Language Models

As of 20 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 5 inbound Pith citation observations for arXiv:2505.16234.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16234 v2

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:08:08.453311Z

measured 105 of 105 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:16:53.580999Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T00:39:35.853384Z

Reference resolution

100 of 129 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved86
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d03fcbf4-d152-4577-8b0c-e084be42cdbe · outbound

This paper cites Abedi Firouzjaei.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Abedi Firouzjaei

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.915175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.915175Z digest=sha256:15886b58b8e38f363cc831ff01a2002486fe2ade7060f79613fc143d9f3f1b8e

Observation 52722d46-9a3d-4a13-9c58-5dec7b1fdc4a · outbound

This paper cites Alzantot, Y.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Alzantot, Y

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:59.980923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:59.980923Z digest=sha256:1b62433372fbdd4b9d996ac3de63d97cc113539ca872a91beac33bc41388a5c4

Observation d62c8ae7-3dfe-454f-bb53-6b6025bae52f · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.100922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.100922Z digest=sha256:547b4466c19bc5b992e231c16d4b286c3145ca81ae4bb824fd2f22a36d3d9eba

Observation 12609bb8-f46c-4f25-88c3-f655ffd5fb18 · outbound

This paper cites Claude 3.7 Sonnet and Claude Code.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Claude 3.7 Sonnet and Claude Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.243342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.243342Z digest=sha256:cb05e27f7d955a5fc3ee9ffb2c3f968ec0cf8ea80a8c22127b855f60a50bfaac

Observation 75e7b2b1-f57d-478d-8d66-009a1e343575 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.370074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.370074Z digest=sha256:a647c528179eda3f13332c58ed2c7b50ace9e03d40f552b47780382a5325932a

Observation 76e6cf04-00d6-4840-bdfa-52c5db7dd6ca · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.482208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.482208Z digest=sha256:f33d51de5efdfeba98ca2c5b25eaf672079df362d5be0846ae77a8c513202f13

Observation f427873a-8f62-4dc0-91be-8d9997e3035f · outbound

This paper cites LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.616769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.616769Z digest=sha256:239bad8568fc2f848ba975bc93b9a40aae839f74e15328a27323f431fa06914a

Observation c8073588-1d7d-4018-a79e-6ede488102e6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.759419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.759419Z digest=sha256:bf730b19789560e2f4ec77d6702456fee2e6bb6b18d35b344048c847a5406424

Observation 85558c24-4962-4f40-9ab3-7f563662fba8 · outbound

This paper cites Bordes, Y .-L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bordes, Y .-L

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:00.923158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:00.923158Z digest=sha256:3732b0f21b4cad3e38c9fb29c288c555d4300c08d353bb80ad25ed6089afbff1

Observation 8a000c4b-be89-4539-878b-9cbd7123646c · outbound

This paper cites Bosselut, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Bosselut, A

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.022186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.022186Z digest=sha256:0460194df2b3b4bd9150206d9f2f9983432edecb6fed33ac3943c087b16a8f93

Observation dee3ee11-3c31-4520-b503-0d57e76e40e9 · outbound

This paper cites Butcher, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Butcher, M

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.144208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.144208Z digest=sha256:7cf2c92901501f18e64113be989cd2ec8d4f065da99a57fd2cd3bf0e3c75d510

Observation 0cec495e-48db-4aa0-aa49-185bd1bb0555 · outbound

This paper cites Doubao-1.5-Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Doubao-1.5-Pro

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.233200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.233200Z digest=sha256:933bf1ee5baa26d817ee494a5dc7fc33a59aa9901230f79c9e652cec038d9bf9

Observation d83dcc59-f853-4d12-ad7f-9fb3c1bf25dc · outbound

This paper cites Chang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chang, X

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.312710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.312710Z digest=sha256:8fc5c08e0c218cdfa296b1caa1deb7cc23e0e1cef2de83b947be330e4ef6434a

Observation 2845bc06-2ac8-469d-a358-83afe998016c · outbound

This paper cites Mistral 7B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.389348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.389348Z digest=sha256:d5461844582f4757dcb0895a3369a14457dca523f77b34c41da50503e1286ea2

Observation 3a2e6bd6-f44e-4228-b0da-817dd6009ed0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.475747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.475747Z digest=sha256:676e1ae93f43028d4e87f68be355f5e739783eb0f373e3e84eab7ba42d394079

Observation 1b9c43a9-c31c-4586-b9f3-67543687f7d8 · outbound

This paper cites Chen and C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chen and C

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.576145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.576145Z digest=sha256:9cd2fe066d410d63835b260810e602afd12777c91c18b2ee87622abdeeea329d

Observation 7b252921-98f3-457a-b125-ee69f5ac761e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.643294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.643294Z digest=sha256:ae3d917abde8633fb1423506e9ad6685f11eef0bd641cd24f15394bee2d7f919

Observation f5bcecf5-7db7-4429-9b1b-7a332a67e84a · outbound

This paper cites Chiang, L.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chiang, L

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.723680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.723680Z digest=sha256:b87758957539d909d1eea72f637e7c0fe30276462404793803d002377977c4f9

Observation 9f64694b-b587-4229-9e0e-87e2b99537da · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.821582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.821582Z digest=sha256:6567aa1fd83859765ef9f0254d325ed47e123b26e3b398a69f11bdb68c65cf93

Observation 9ee681a9-1b69-4f20-84b7-9353dea5e07a · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Training Verifiers to Solve Math Word Problems

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:01.907072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:01.907072Z digest=sha256:db4f193401d61886e59a7ee34a9605181cd51a1e43f9c08e4c2929547c0ed4f4

Observation 5eef2f6c-3b9a-4e35-bcc9-6219b1c0f537 · outbound

This paper cites Cohan, F.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Cohan, F

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.090683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.090683Z digest=sha256:d94813f7d659016c1d67ea3451399f3ca8cabc19d86c983d0b70c57c10a41e82

Observation b8a754b4-d26d-4973-8e81-433936305d43 · outbound

This paper cites Collobert, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Collobert, J

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.249374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.249374Z digest=sha256:ab9e0c8eaf946aaea4a13512c164f420a561bf2779ca7eb2f3d434017f220309

Observation 642495c2-3a0b-4a7a-9b68-31d725882625 · outbound

This paper cites LCFO: Long Context and Long Form Output Dataset and Benchmarking.

LIFEBench: Evaluating Length Instruction Following in Large Language Models LCFO: Long Context and Long Form Output Dataset and Benchmarking

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:08:12.116541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:02.336424Z digest=sha256:2bb71d2f1d78cf7a5982201a6457aa14fe73879b8e5ceaa4b6ea080958230e39

Observation 8575674f-54ee-4641-8746-f565c54fe5dc · outbound

This paper cites Davidson, D.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Davidson, D

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.496551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.496551Z digest=sha256:b02e668eca6c4c1101509f961dc795973ab0b4d725b631332f1e2e85032a37a7

Observation 264207e3-a4f7-49e7-9fea-48f19a70c967 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.698130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.698130Z digest=sha256:d057dbf09470b95c411420482cbfc41d23ebd7c190d6774b0aa10c1ba53c21bc

Observation 207f2d8c-d40f-4223-8f31-f95caddd5bb1 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.802654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.802654Z digest=sha256:cb096dc94bfd73834bc8f540928c28cf233fa7331985373e482266b019bdeb71

Observation 348202ab-4617-46ca-b372-c9179b218782 · outbound

This paper cites Dubois, C.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Dubois, C

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:02.954068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:02.954068Z digest=sha256:f5a4b540a38c0e4e642e53c0cdf54aa18037a8d66e19f5ea438d45fe98841b6c

Observation 74eced30-2b84-4219-8b65-3311a6304d66 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.106742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.106742Z digest=sha256:4795257461b074dbef6e8a1deb8c119c35f5b3ad90fe6872fac15f0887fc85bc

Observation 29979e30-9658-4795-aae9-fb2f03def193 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.246808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.246808Z digest=sha256:59313ddec7f6b4177dee07a7b4a1b5b7f2664d1a0d2fde24bc0cd0129c89439f

Observation 02d05053-243d-4d2a-9694-2ef27cfb62d8 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.356120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.356120Z digest=sha256:d52d3f631ce90de061b179b12ab95ed250b458e5745332f45ff82cdadda03bd3

Observation 7520112c-1822-41e5-aeb0-119846f48bbb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.507635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.507635Z digest=sha256:82e6f0fad53552976ec5473f20f4b64c6a864d57b3d779c144a565e39ffb0954

Observation 1f012298-b49c-4271-9aa4-65d5d67af0c8 · outbound

This paper cites Foundation.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Foundation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.619331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.619331Z digest=sha256:e3067666238b09897b71c39e05a7f8dcd7562a0f906aec84606726eb67781d28

Observation c4b3fe12-b0fc-4900-b9c7-ee56d1e7a38c · outbound

This paper cites ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.761610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.761610Z digest=sha256:ccf1b24f7abc7d749e169261cd6568473c2ecbf1f388ca95bf905cef6a53f93b

Observation 31377b83-3942-4967-b76e-687b8a78e71f · outbound

This paper cites Gemini 2.0 Flash.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.0 Flash

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.821979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.821979Z digest=sha256:b9a8454db8cdb31ae529fbf041d6d1c54c4493dd742c003375962e40d0063e7b

Observation ed19b53f-3d7b-4290-8c8b-879e4119f478 · outbound

This paper cites Gemini 2.5 Pro.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Gemini 2.5 Pro

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.877823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.877823Z digest=sha256:b8b0003a18256f7f139725aa3acb50c323bdc5fc5d363884a5fe2959fec4defe

Observation 9e937a04-a837-4446-a22f-80a85be20f3b · outbound

This paper cites The Llama 3 Herd of Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The Llama 3 Herd of Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.923252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.923252Z digest=sha256:a7317bed871c3059b20c36ca57dfc1fe35cca3cf63d3b83a11af31c08667e480

Observation f19c070b-17d7-4cfa-9a32-0a7e8985ddd5 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:03.998056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:03.998056Z digest=sha256:14dd78b1756deaa9cb74590e8a6f7d24b09caa511416d0f289ceaf29c2cb4424

Observation caef362d-1e6e-4407-81dc-227fe3ba992a · outbound

This paper cites A Survey on LLM-as-a-Judge.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Survey on LLM-as-a-Judge

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.044070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.044070Z digest=sha256:9370a407b6e2b3cffc90619850a4fac4fb56d341bcc32e40f019cd688a28052b

Observation 0a1bbbcf-23ea-4ed9-9ec2-9ab31402e39f · outbound

This paper cites Length Controlled Generation for Black-box LLMs.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Length Controlled Generation for Black-box LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.110426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.110426Z digest=sha256:6f45a7b07e14209a4f0b775ddf28dfb81b2cd1264f4c72fe494fe96d43a21984

Observation 6e458285-5368-4628-a9a7-9cfa31d92340 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.157989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.157989Z digest=sha256:c61285723e26e5f57168ba7a839a92b60fd2698abf56f9943a3069401607b400

Observation f3ab0426-6758-4aac-bd61-241b9d68ca63 · outbound

This paper cites Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Multi-IF: Benchmarking LLMs on Multi-Turn and Multilingual Instructions Following

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.207512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.207512Z digest=sha256:c18df77ded62bd78702cf49036093b451291e9fdb31d0ac60ec0c822aca0e057

Observation fd5c61b4-e6f6-45aa-8607-9abb2af7ac50 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.244676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.244676Z digest=sha256:07b9a1b9b2c79a61c56a4fb0852ad5d2e66d6e80823bb5b3924e89c2e57235a3

Observation 48bc6929-6b5d-479b-b24e-28ae7ef193cd · outbound

This paper cites Hsieh, S.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hsieh, S

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.314289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.314289Z digest=sha256:6719cd7d8ec1bf6e2ab6b39638ac7ecf0dcd235026bdf812090b2da9bb26123f

Observation 0e770906-3b15-4109-8d83-d733cff41d74 · outbound

This paper cites Huang and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang and K

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.379241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.379241Z digest=sha256:5430412f83e98900eb8a87990c63227ba03f52864e67639af2da0f249f20f387

Observation eaee9ae1-cf63-48cb-84a5-8bfc7eedc978 · outbound

This paper cites Huang, X.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Huang, X

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.438908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.438908Z digest=sha256:74dc77828d9ff4be35571f11a492e44fd7a65b4f803bcd2d8ffcf18c92bd5ad1

Observation 5bcac5cd-9a07-4a6b-92fc-abaa0f1faeae · outbound

This paper cites A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey on Evaluating Large Language Model Applications in the Medical Industry

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.497560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.497560Z digest=sha256:a2de7fd5bfbbd9473b4310033fa53bc7e36cb23e746c370bf070cc3160a3d543

Observation 9b0ac765-2a37-4dd1-9ad1-3c987da4f752 · outbound

This paper cites The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input.

LIFEBench: Evaluating Length Instruction Following in Large Language Models The FACTS Grounding Leaderboard: Benchmarking LLMs' Ability to Ground Responses to Long-Form Input

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.558589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.558589Z digest=sha256:5eae03bc3b0bf5a73c7ac2c58dfcfb74f1caa89848eeeb778f8d562c6e870291

Observation 5f1ff18e-cd0c-4fee-8eae-5582c8694fa2 · outbound

This paper cites OpenAI o1 System Card.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1 System Card

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.604933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.604933Z digest=sha256:6619e57a4d5beedf9e1b0dfb2498bd9bd8ec92f9d80b76151defaddd6cc771c9

Observation 67340472-1885-4a2d-8bcb-ab51bfa796ae · outbound

This paper cites Jhamtani, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Jhamtani, V

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.656826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.656826Z digest=sha256:1c110b43247ef0f714577083f87398fd01794760b51766432a939d7cb095a29a

Observation 4201875f-84b9-4362-abea-85aa730b65fb · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.700834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.700834Z digest=sha256:796169b84d12247fc3545535d5372e90c6836ece843ff7fe234e85d1ad5a2d8c

Observation f8f43d67-01fe-4215-b405-3477faba6465 · outbound

This paper cites webnovel_cn (revision 745338c), 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models webnovel_cn (revision 745338c), 2023

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.773018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.773018Z digest=sha256:c16b616dc238533f32238a19ce92f3436aadeac039a95b478d28f735178b8de8

Observation 01265994-1f64-48b2-b6cb-afa6fad5df13 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.828592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.828592Z digest=sha256:98c93d4e04b5ce8c05848f061f3fa98f23d933c2cf65e1c8afa4ab753ddc9f4d

Observation 57ab9db9-b0ff-480c-b1f8-5985cdbab959 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.891231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.891231Z digest=sha256:4dc9adff01b39155fd61d6560d0f366b1f4daddab370c69274e13360315ec396

Observation b05c97b9-1322-40df-8b17-02e580e2fcc5 · outbound

This paper cites Koupaee and W.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Koupaee and W

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:04.954876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:04.954876Z digest=sha256:322f592a71fc5b067517fcc30b45a3118f687fc59907074eb30f7cdc4d5a4c3c

Observation 0a4ae45a-d174-4dbe-82ed-da489038619a · outbound

This paper cites Kry´sci´nski, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kry´sci´nski, N

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.023199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.023199Z digest=sha256:1ba4dabf8a581c9990d9a55335285a04b883fc8ed6fb3010f2ac533c744e54bc

Observation 2ce96cb9-3ce3-4822-8fd9-f5373e7b8f16 · outbound

This paper cites Kuratov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Kuratov, A

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.089583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.089583Z digest=sha256:a79cbece90e699466d6d4ab713940ad9a5250765b1c71ea4db2b85bece5762dd

Observation 06ebd402-d64f-402d-a331-cf01a447796f · outbound

This paper cites Lample, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lample, M

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.137728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.137728Z digest=sha256:f04654949c79e11f0d56b2a507fbfc6d59e28aafd3a2cd78707e4ec6dd5f1532

Observation f5181dd5-d25f-48fb-83d8-73d647c30973 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.197847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.197847Z digest=sha256:ecc77cb8a2470310e1f6a5e796ac811ce5f322a1144fa8c9bcad60d0b80ee210

Observation 49645e82-b652-4441-9933-cf3cd5d6f730 · outbound

This paper cites Long-context LLMs Struggle with Long In-context Learning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Long-context LLMs Struggle with Long In-context Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.259340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.259340Z digest=sha256:1a4637aef1855a04ef921b0058b6e84f315bc9ef291cf372b11ba59b8580252c

Observation 9555014a-566b-40ee-927f-818b14f6d62b · outbound

This paper cites AI Awareness.

LIFEBench: Evaluating Length Instruction Following in Large Language Models AI Awareness

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.339186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.339186Z digest=sha256:5b84f4ac8fe93f1a40c0b1b4603fbbfec2d3c93e35b72d1b7792aac504056839

Observation 24f626ea-adfb-44ea-94c5-37195056c79d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.453724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.453724Z digest=sha256:613374e00a6ef660faa4db1fc0939513a4ca329d1d4567b14f2fed9aeb6178cf

Observation 4d97b2c8-357a-470f-8ada-2521b0871e72 · outbound

This paper cites Controllable Text Generation for Large Language Models: A Survey.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Controllable Text Generation for Large Language Models: A Survey

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.544649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.544649Z digest=sha256:b522833b555ba0531b952047d50f747df7cd4a6c3696ac92d261a7f8c12c7972

Observation 730f358e-1ce1-4107-bdd1-0a7bb8fc9a1c · outbound

This paper cites Lightman, V.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Lightman, V

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.644559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.644559Z digest=sha256:e85e771a22317a9056981fc004f5a298f659e4e0ec1c28f199ec4937fa90b651

Observation bda9e038-2d63-4e11-90c4-3171e0cd5aa6 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.718550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.718550Z digest=sha256:5ea8886fc2c191996df8b8f8138126aeffc7681ca4d5c6e60b6d3de92102aff1

Observation 2fe3fbff-a7f4-4f7c-a39c-f29a07718948 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.827379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.827379Z digest=sha256:f80959e8eea1d840e11040af25c17ca3cb6db556876b8c506e977efc39ff8e5e

Observation 9f8e5921-e614-4fe2-9239-c1d4d2eca7e8 · outbound

This paper cites DeepSeek-V3 Technical Report.

LIFEBench: Evaluating Length Instruction Following in Large Language Models DeepSeek-V3 Technical Report

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.916363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.916363Z digest=sha256:8524f3a5ce2c6fb261620086d36418b663f501c22bda724b91c5bce96ca3251a

Observation cd80fcd9-43ab-461a-9333-a5737480603a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:05.968830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:05.968830Z digest=sha256:875c8b3b2b33d026fa20cc03df50b6e454ef9d2b8c21fbb2aa6ce640af2d88b0

Observation a8ce4bf1-95c6-4c87-aaa5-dc73ae8255f3 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.040972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.040972Z digest=sha256:11547a148ee6147e058f5c6dbfad9946b9e790a12889771fcd93e266449c8dc1

Observation f44b67bb-10ee-43a8-ac99-29f04933732a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.965757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.138421Z digest=sha256:4c75923e98467b950ad5406b0b892ff1e1788ee99227091c100ea8307a8d9851

Observation fe4da9eb-02ca-4ea7-9620-ebff6fd0f76e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.237974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.237974Z digest=sha256:989fbafbdcb6ae6cfa22f5f4a296482eb104e0b5b15e84d188fb57c01aa52067

Observation 5380913f-d36a-42ef-b2c7-db46a4ade9d3 · outbound

This paper cites ExpertQA: Expert-Curated Questions and Attributed Answers.

LIFEBench: Evaluating Length Instruction Following in Large Language Models ExpertQA: Expert-Curated Questions and Attributed Answers

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:06.311393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:06.311393Z digest=sha256:53ba8a49c16da5ff5c0a0d9f82aeb825c24aaf7cddbd1d8ca1b251df084b6ca8

Observation 76fa2654-0be0-4af4-8c2c-89fa5fe9345d · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:20.823549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.397371Z digest=sha256:f29a21e4b666e6c39bbdfe0598289ecab69098466123bc6ff09b3a07a71cd0b1

Observation 69b59cad-1a91-4d25-a19e-dc830b707bc7 · outbound

This paper cites Chinesenlpcorpus.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Chinesenlpcorpus

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.743188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.513962Z digest=sha256:11a3916d4dc4cbc44c4c75d89507073ae09b2f993b7f773b1117ba4c3ec75ac5

Observation 24b62dbc-691b-4772-9348-0ed39bdcd484 · outbound

This paper cites Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mnbvc: Massive never-ending bt vast chinese corpus.https://github.com/esbatmop/MNBVC, 2023

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.639920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.674885Z digest=sha256:d57eb862e2028332f28c16b07ce9b28848d18407c1478f893bb29f0dc4688570

Observation 1173264d-9c54-422a-93b9-da8a23ce010e · outbound

This paper cites Mostafazadeh, N.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Mostafazadeh, N

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.537826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.786681Z digest=sha256:cccd7e32f5a5096b32ae767b84d2b1078caf3a0055d066f012d59107643a3ccb

Observation 85b58e44-50d7-4733-8ca1-173a4139db81 · outbound

This paper cites Nallapati, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Nallapati, B

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.369749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.882313Z digest=sha256:3ff24e64652de2075e15d8827a76052b344b11c20ef2697226363203a32dad0e

Observation f6c223a5-af30-4fcf-a99f-2bbb5af790b9 · outbound

This paper cites GPT-4o mini: advancing cost-efficient intelligence.

LIFEBench: Evaluating Length Instruction Following in Large Language Models GPT-4o mini: advancing cost-efficient intelligence

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.180851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:06.933239Z digest=sha256:e76b4f521feb5eb1f76ffd0b979de243c57eb0c9e789b4b19e7c1c7bdf6c668f

Observation 0dd40658-09f2-4836-9104-014091f5039e · outbound

This paper cites Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Hello GPT-4o.https://openai.com/index/hello-gpt-4o/, 2024

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.032956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.032956Z digest=sha256:d9037ab26fa686acd0a7776f7040591d7f5cae96506096831e17cc98cc3f8f18

Observation c04d3ce7-0fc1-4c56-890c-8de0610a4e57 · outbound

This paper cites OpenAI o1-mini: Advancing cost-efficient reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o1-mini: Advancing cost-efficient reasoning

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:20.018525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.136231Z digest=sha256:dce7a94f93728b53e30a2ce2003e1a1785a930665ccb91335244a15fafc147c8

Observation a5e6597d-4497-42ea-8a4e-8e43c9b6123b · outbound

This paper cites OpenAI o3-mini: Pushing the frontier of cost-effective reasoning.

LIFEBench: Evaluating Length Instruction Following in Large Language Models OpenAI o3-mini: Pushing the frontier of cost-effective reasoning

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:19.781683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.212823Z digest=sha256:3a1ea50b8380cf02b44e35cb10da7666c297cba88f7f633e0910d515f335457a

Observation e2b1c336-0ac1-4c52-8c46-6131af0bed4a · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.563974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.302010Z digest=sha256:fd060a53567cc517a60fa82169604f1c87f86623f4b3d2fcc936bbf0705ff328

Observation 5f1e47e0-eb8d-4b31-9ccb-2af426340b81 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.417697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.421996Z digest=sha256:fbee3217cfc6be829be6c16331e1f4af7d5f6d6ee359814baf0e9240a0dbd350

Observation 84f16566-6a20-4804-866e-4ec8307b499e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.248048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.521225Z digest=sha256:5348f956c30301da58e525154c6676b05274f9ea4f7fe8cc42cc80fbcc0fcf55

Observation 5a92f655-6357-4763-aa38-2c13fc4c6f1e · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:19.112561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.583818Z digest=sha256:a9d9fbb221e73674cd60068329bab04b81dd6dd5962b399da75d7a4448376193

Observation bb0e9050-1346-4443-b9b1-b22197d61123 · outbound

This paper cites Language Models can Self-Lengthen to Generate Long Texts.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Language Models can Self-Lengthen to Generate Long Texts

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.651762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.651762Z digest=sha256:4c79ab294d80278bb9bf1d1648e35306e5b9a212398dd70bc8132b849aadfa0c

Observation 27da7725-b122-4352-8b43-795015fecb70 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

LIFEBench: Evaluating Length Instruction Following in Large Language Models HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:07.698689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:07.698689Z digest=sha256:bfb9a9bd6faaa8225ff57b2a94c00920063b1d15693bd050a1a7aef3ea64f53e

Observation cfa6af53-cd98-4c87-8776-4bde43e9b3ee · outbound

This paper cites Radford and K.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Radford and K

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.997721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.792526Z digest=sha256:386bf3696f9d1479459fffc3b2679a581a45184f3698f1b4a32cd7bea3394241

Observation ae6662e7-ee89-48b2-be05-5437d31a5453 · outbound

This paper cites Rafailov, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Rafailov, A

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.862390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.833724Z digest=sha256:2599881c6caa360849394d6ae4b29a54cf407b7cb7eb93ea66591a40055807db

Observation 91c869c8-9537-4c21-a830-09f45af784a0 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.746199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.872975Z digest=sha256:dae16011f1f728bec9daf2f9167737c3372f54335a87ee0c44a89c12618590c4

Observation dd45e7ff-fbd0-4514-b49d-f0ff703a619c · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.635380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.912824Z digest=sha256:a0672d51064812fafe7eca628433bbe63ec07ac879b30c3d451a635cc12c96f6

Observation 84318624-9748-4bb5-86af-f06cfb3c21e0 · outbound

This paper cites Sennrich, B.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sennrich, B

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.512717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:07.978742Z digest=sha256:2982f3f83e6f66caa00380fb16d0f2c188dccb77977180fe3f0735c2cd0a9e89

Observation 75dad7df-64c9-439f-ad22-0c65def8afee · outbound

This paper cites Shaham, M.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Shaham, M

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.362523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.027546Z digest=sha256:ad06f6cf6fd457204c167bdbea00cc67a25f6b5f4c0743e6332d467b1a95b75a

Observation bc065065-870c-4664-b9d5-a5dc81738cb4 · outbound

This paper cites Socher, A.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Socher, A

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:18.204207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.073224Z digest=sha256:d09c4847c8fda861214ad6b9ae97bab91211ed61fa8d1d95c5dd22ca96dc4b50

Observation d8ffbb68-2cad-46e2-b7a1-28e802e14dff · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:18.065546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.114897Z digest=sha256:d82cfb64c5071e86b19b94a85ed44e2c8d4f13b6ad30b38effe913c5919aacf9

Observation c2cd78d9-fbae-46fa-9bd9-170b91d2a54a · outbound

This paper cites Sutskever, O.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Sutskever, O

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.152779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.152779Z digest=sha256:cfc09689c268041c9151b4a69183dda2157c395c1c9c64eaaddbae810d5bc44e

Observation d8b7c369-252c-4771-ba80-ece7033421e9 · outbound

This paper cites Talmor, J.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Talmor, J

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:08:17.933821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.199068Z digest=sha256:1aea74c57c54de7a0883e81eeab604fb025d7bb2d2579464948c3af1303d0234

Observation aeea2fa9-9cbd-49b9-8125-dfbc067f22bc · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.705498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.252432Z digest=sha256:186d360a56818d14a7d91994001a3335139762110c81fc59be4541834520354f

Observation 4b274721-20f3-4893-9441-212289a8c5cb · outbound

This paper cites CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis.

LIFEBench: Evaluating Length Instruction Following in Large Language Models CollabStory: Multi-LLM Collaborative Story Generation and Authorship Analysis

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.312320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.312320Z digest=sha256:a0b44f1cc775417e2ac912582d4f2da5c7dbc8f9320f0c2c89a1d14b7267f34e

Observation fc9de4cb-0846-4cd2-a62e-5f5ab5bb5382 · outbound

This paper cites an unresolved cited work.

LIFEBench: Evaluating Length Instruction Following in Large Language Models Unresolved cited work

Reference 99

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:08:17.405637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T15:08:08.393222Z digest=sha256:1cb0c90c33f1989366b951a34196079175137302cb20bafd04f551e6fc1e8ab7

Observation 617e84f1-45a8-486b-bc09-a595734d39da · outbound

This paper cites A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment.

LIFEBench: Evaluating Length Instruction Following in Large Language Models A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.453311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:08:08.453311Z digest=sha256:f47cc3ddee0e25f3e70bee550a9dbfa7a74ce4b578f528ed07669a312a8597e7

Pith citing papers

Observation 6ff77394-c99f-42d7-af36-3730cab6fc15 · inbound

Prompt-Based One-Shot Exact Length-Controlled Generation with LLMs cites this paper.

Prompt-Based One-Shot Exact Length-Controlled Generation with LLMs LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:16:53.580999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:16:53.580999Z digest=sha256:51d380137575f0dbb86f042e2b5215c4ad36b54c5d7cfbb579a303904f331295

Observation 30905453-4f45-4e4b-b4e7-d3288f6a18c7 · inbound

TiCo: Time-Controllable Spoken Dialogue Model cites this paper.

TiCo: Time-Controllable Spoken Dialogue Model LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:39:35.854869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T00:38:52.182973Z digest=sha256:f1e31a7285d9df1b44cb8ea0bb4248d47f646add213917242a0a01b839eaf562

Observation eacdae65-3aa9-48a7-84d2-fe4f586cb22c · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:31:26.140879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-07T10:42:27.644514Z digest=sha256:892aa874e7ea651b4540d0d3bbf9e673d5e294cd2284bb0348d7655beaaa6f45

Observation 463c613c-2a32-44e3-bfe1-3cbabdb90551 · inbound

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling cites this paper.

Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T15:28:14.401581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T15:28:14.401581Z digest=sha256:826518a8cd80a9e5b7c0a47b6a965d4cd736aace639208c699e0a003f7897c56

Observation 9003ed95-f141-480b-8560-443b02ae4816 · inbound

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation cites this paper.

The Librarian Who Refused to Code: Model-Dependent Identity Enactment in LLM Code Generation LIFEBench: Evaluating Length Instruction Following in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T18:03:36.857585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T18:03:36.857585Z digest=sha256:05677030742e47daed9376134970000e342aed00b2d65fb6249a563ffa6e1fce