Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2401.00595.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T16:59:49.571352Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 9c1c3d5e-eab5-40a0-b9de-b33195ebaef3 · inbound
Revisiting Sentiment Analysis for Software Engineering in the Era of Large Language Models State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a1f4b398-d6c4-4504-8d34-9ade7a41a5a4 · inbound
CRUXEval: A Benchmark for Code Reasoning, Understanding and Execution State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 003918c0-7574-4de8-9d64-a19d8f777951 · inbound
LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 106
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e7fe15f2-ee82-42fd-9a94-9449eb690554 · inbound
Holmes: A Benchmark to Assess the Linguistic Competence of Language Models State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 004717d1-fbf6-4a2a-b95b-438caa3ec91d · inbound
Lessons from the Trenches on Reproducible Evaluation of Language Models State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 67f2ded3-90f4-4e78-aa56-2b349ed0a4ee · inbound
JuStRank: Benchmarking LLM Judges for System Ranking State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b4d1b1f-c45e-4d0d-bf19-4c528f644464 · inbound
Drama Llama: An LLM-Powered Storylets Framework for Authorable Responsiveness in Interactive Narrative State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810e7897-981a-4a76-afef-92a24701ff33 · inbound
LCTG Bench: LLM Controlled Text Generation Benchmark State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebf17e8a-803a-46f9-9a86-1f73a7e839db · inbound
Evalita-LLM: Benchmarking Large Language Models on Italian State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fff97b18-4589-4e30-a1df-0d7823c72029 · inbound
Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f9febc8-c580-4f34-857e-c867ebda099f · inbound
Personalizing Education through an Adaptive LMS with Integrated LLMs State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 853fbbd7-28bc-4a13-adbd-0e86dd15e784 · inbound
Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7d7e74f0-daba-4236-9275-bd8ee43cafac · inbound
Predicting Performance of Symbolic and Prompt Programs with Examples State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6d86b257-40e5-4b84-81d0-4713d9f46cc5 · inbound
SafetyRepro: Configuration-Conditional Rank Instability on Alignment Benchmarks State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7d68ccd6-9171-44d5-b05f-9d322032e90c · inbound
Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 51592a10-a247-47d2-a4b1-a9a8a46924ea · inbound
Latent Confidence Alignment for LLM Self-Assessment State of What Art? A Call for Multi-Prompt LLM Evaluation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.