Pith. sign in

Paper Citation Record · LEDGER

EvalAgent: Discovering Implicit Evaluation Criteria from the Web

As of 17 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 2 inbound Pith citation observations for arXiv:2504.15219.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.15219 v2

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:34:42.445714Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:28:02.231013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:28:02.879377Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy14
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dc80ca97-e0d2-4389-88bf-08dd822a47f8 · outbound

This paper cites Art or Artifice? Large Language Models and the False Promise of Creativity.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Art or Artifice? Large Language Models and the False Promise of Creativity

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.413201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.210971Z digest=sha256:14e2df8d2cfd83f7c6a2bfe35de5899000f6af2f9fc385790c7a99ea0dd89d49

Observation 8bf1413d-b26a-4b87-867f-348f66932095 · outbound

This paper cites Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Can AI writing be salvaged? Mitigating Idiosyncrasies and Improving Human-AI Alignment in the Writing Process through Edits

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.222740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.222740Z digest=sha256:a05c2a3a4fd2504cad9e4c274380c7eb71016aa3297b6504df347fe1960864fe

Observation 69546b1e-1f75-4132-827f-40495f4dd5a5 · outbound

This paper cites Complex Claim Verification with Evidence Retrieved in the Wild.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Complex Claim Verification with Evidence Retrieved in the Wild

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.230404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.230404Z digest=sha256:1333c2bc440a6c3a6ac3dac77db4590b3dc8c48ad43c432296227dff4235a56e

Observation 1d06759a-6ef0-478d-a0c1-1fdeb0a93276 · outbound

This paper cites Angelopoulos, Tianle Li, Dacheng Li, Banghua Zhu, Hao Zhang, Michael I.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Angelopoulos, Tianle Li, Dacheng Li, Banghua Zhu, Hao Zhang, Michael I

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.390375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.244411Z digest=sha256:92228401b72ab9278365f8dfcb36deb9eb7887478b13f6935898062261e4248b

Observation 3985c6c8-bb01-497a-b7bf-2406a0c5be53 · outbound

This paper cites TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.252534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.252534Z digest=sha256:0daca6900778012ee28814de97e758cf90ac66b4719c59870261b79f00121c32

Observation c962e89d-71ce-44e6-829c-f7db689d7c1b · outbound

This paper cites Gonzalez.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Gonzalez

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.371846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.263813Z digest=sha256:f4d33c195fae1451f730721d5195b91a692ca6e9a07cebef659edb42454108d5

Observation 52971542-9699-47ac-9a56-320483853afa · outbound

This paper cites The BiGGen Bench: A Principled Benchmark for Fine-grained Evaluation of Language Models with Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web The BiGGen Bench: A Principled Benchmark for Fine-grained Evaluation of Language Models with Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.280569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.280569Z digest=sha256:ca2908dabdae470f30bf1257c161e517d3debe014a276c4798e67dc38954beef

Observation adba927f-252c-4890-b8ab-a746d39eef5b · outbound

This paper cites Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.286451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.286451Z digest=sha256:8d786427b034e673a789e30b27020a1cd1bf59b0935be73f22c491ec1e15ae7a

Observation 4c348ca0-2464-4ca0-913a-9373a1a9281b · outbound

This paper cites Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.298607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.298607Z digest=sha256:dc3d0e76928a28c9cf24197a6af08cb30840597e4fa8a82f35d9719de92fd526

Observation 99eacea3-4d3c-4d4c-8e35-5aac3f5fdfda · outbound

This paper cites Linguistically-Informed Specificity and Semantic Plausibility for Dialogue Generation.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Linguistically-Informed Specificity and Semantic Plausibility for Dialogue Generation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.316933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.304270Z digest=sha256:719222768a94a1db054583d5f7ec84a7ff21e417e618d9f01358c7280ffc88b5

Observation 12ccbb68-75cb-41a7-afaf-d1e0d590d972 · outbound

This paper cites Large Language Models Are State-of-the-Art Evaluators of Translation Quality.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Large Language Models Are State-of-the-Art Evaluators of Translation Quality

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.291782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.309014Z digest=sha256:1e94a5544c0cd273ee3de0775596defa28327cc3580945932212645787778cf7

Observation 7342278c-4d72-44ab-8575-f2f90ce1aacc · outbound

This paper cites Hashimoto.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Hashimoto

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.263699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.315778Z digest=sha256:c28d65343a955bce9f0d4af9e82fb8c2a9d709d15563a2872846aaff74572e69

Observation 90a8c738-2023-4469-90ea-8bf2bb399a19 · outbound

This paper cites Wildbench: Benchmarking LLM s with challenging tasks from real users in the wild.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Wildbench: Benchmarking LLM s with challenging tasks from real users in the wild

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.325583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.325583Z digest=sha256:b2f7f0a77c94b4329bf0a65236d81aeddf18f0419aeb6c793d74ccde244de0dc

Observation a2b6920b-4535-4644-b467-0f4012614289 · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Self-Refine: Iterative Refinement with Self-Feedback

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.181610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.337084Z digest=sha256:db01b6c3e23f1b2f434bd60f701cdf1d7c5aa68550e516535d6d395af2c2a01a

Observation 5d9e88f1-b382-4fac-ba89-99bf33df1689 · outbound

This paper cites ExpertQA: Expert-Curated Questions and Attributed Answers.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web ExpertQA: Expert-Curated Questions and Attributed Answers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.341801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.341801Z digest=sha256:c7f66fb2cd9f4efdebb43448332809ef566c4c30a2ed153525377c62a15da15a

Observation 70159c1a-3ed7-4fe4-b530-5500477f9a25 · outbound

This paper cites Dolomites: Domain-Specific Long-Form Methodical Tasks.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Dolomites: Domain-Specific Long-Form Methodical Tasks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.347058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.347058Z digest=sha256:28f4c86b502f69fda8e7133452bededd15f41ac2da1419db8158b99905b53e83

Observation 0f429166-e12e-4055-bc4d-b0f351b5269b · outbound

This paper cites Contextualized Evaluations: Judging Language Model Responses to Underspecified Queries.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Contextualized Evaluations: Judging Language Model Responses to Underspecified Queries

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.353884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.353884Z digest=sha256:a6462e1f0b67cb0772c6cdc34f08675057ef0c2c020529fd825e3588f88be912

Observation 4c71e040-0d1a-411e-9b3d-05ecea94e92f · outbound

This paper cites Olausson, Jeevana Priya Inala, Chenglong Wang, Jianfeng Gao, and Armando Solar-Lezama.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Olausson, Jeevana Priya Inala, Chenglong Wang, Jianfeng Gao, and Armando Solar-Lezama

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.164788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.359344Z digest=sha256:064c956dccd354b12891133005a99c3c926590aa5e067db3e2e93217d2be7f8f

Observation 3c210ad4-14a1-4870-89f0-2571ad844cf3 · outbound

This paper cites InFoBench: Evaluating Instruction Following Ability in Large Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web InFoBench: Evaluating Instruction Following Ability in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.364569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.364569Z digest=sha256:6330646e66db9f7565dba3d647c4618c139c3df15ac1cf79ebbbf9dbd1a9ecde

Observation f1ff29d4-0a29-4ec6-9522-9d03deb9ac29 · outbound

This paper cites Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMs.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Do LLMs Plan Like Human Writers? Comparing Journalist Coverage of Press Releases with LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.370900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.370900Z digest=sha256:43f458c77608b63a9f7418d7a6d1d5c2ce6656e779c5f27e311409ff51184c8d

Observation 6b088437-5a3f-4471-9a81-7898f9261f3b · outbound

This paper cites an unresolved cited work.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.375275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.375275Z digest=sha256:8a1dd55c8c89cd29e98922a882c2b1af28a50eae6b6a6554ff74a70fa0501c5b

Observation b0386a8b-126c-45c7-ba52-a148b8c25294 · outbound

This paper cites Using natural language explanations to rescale human judgments.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Using natural language explanations to rescale human judgments

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.145289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.380031Z digest=sha256:15b51679aeb498b838900cc1fca841b71f2f81447186fa48482063e48cac1863

Observation 266acaca-ff34-41b3-84b9-c5aedf4f34f5 · outbound

This paper cites Learning to Refine with Fine-Grained Natural Language Feedback.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Learning to Refine with Fine-Grained Natural Language Feedback

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.384836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.384836Z digest=sha256:a890e9408f9a6e8bee52885809cfb6ab7c2f94ea90429d95404eb29462295334

Observation 61b8e228-8783-4e71-88c8-65c7bee1ed19 · outbound

This paper cites WritingBench: A Comprehensive Benchmark for Generative Writing.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web WritingBench: A Comprehensive Benchmark for Generative Writing

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.389900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.389900Z digest=sha256:fdafc9b528673adcecc1c792049c96775ba07853a48de907b854955ba3b5e67a

Observation d6f90d70-e7da-4f5f-9053-a5fe3b0fe9b3 · outbound

This paper cites Fine-Grained Human Feedback Gives Better Rewards for Language Model Training.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Fine-Grained Human Feedback Gives Better Rewards for Language Model Training

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.108754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.394755Z digest=sha256:d387339076d8c9c3436bc9f5b40fabdfb0409978b507535b1aef97ab1a350307

Observation be34f127-a47e-4eab-a3cd-84903821029b · outbound

This paper cites Hanjie, Runzhe Yang, and Karthik R Narasimhan.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Hanjie, Runzhe Yang, and Karthik R Narasimhan

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.089717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.402219Z digest=sha256:f8fce9d926cd0d07249149e64af67b865e7655aab31cc33618dbf618eec501eb

Observation 4f45d2a8-3615-4530-a97f-2505229d2333 · outbound

This paper cites Self-Rewarding Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Self-Rewarding Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.406977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.406977Z digest=sha256:441b9e5dbaba3796b0bc4c484548c530ca91f20bc9d744bc3b656c8a6b4d336f

Observation 7e16b302-6e24-4253-b436-5ebc178351e8 · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.055058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.412226Z digest=sha256:87325a83c312acad1788ba3947dfaf8a8287913b92e5e6b88bea41daf4122b26

Observation 82d1d276-15e5-4e96-becb-8f4aca36dc77 · outbound

This paper cites Learning to Control the Specificity in Neural Response Generation.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Learning to Control the Specificity in Neural Response Generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:43.024570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.416655Z digest=sha256:e6d7e03aad770e02b81a9b9f2a9f5a310388c96a5a1d66277a2f90e49d84d25f

Observation 22ccd00a-76de-4e38-90c6-d54f38d16858 · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T11:34:42.997395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T11:34:42.421285Z digest=sha256:0698545e3850e2c2bc58eba643dfbfc30a1eca99b0d539894a2801e36d4aa15b

Observation 841621a3-eabb-478c-bcb5-bca8d6fbebdc · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Instruction-Following Evaluation for Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.425730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.425730Z digest=sha256:85296a84a14f8d33bf04e9de609993f1887051eb0c91479731b17494425715aa

Observation b70ff4d5-af34-457f-9fad-2c108ef8d972 · outbound

This paper cites write newline.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web write newline

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.430515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.430515Z digest=sha256:4766e76f209f5b4c30d680314ec4ee08ce2446fcc223e77ed7818542d9a6d0f8

Observation 2ab15910-6e9c-4742-8b63-95b0ea2be823 · outbound

This paper cites @esa (Ref.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web @esa (Ref

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.436029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.436029Z digest=sha256:5f6242faedec2c3b451d8e354bedccb631e11beab1ca19178e710dd74f135545

Observation b725652b-dc17-47ff-9468-f79dd0acbfac · outbound

This paper cites an unresolved cited work.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.441005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.441005Z digest=sha256:15f6708e5cb5c370a8ccb7bf968635d701b585ccc2dd3967863f792f5687803a

Observation 749d0ce5-4339-44c9-bda6-780be7e3582a · outbound

This paper cites HF f? -`U w DZYH+ RX AI 0 6g:B0ʁEڑD @Q Vx 6y!.

EvalAgent: Discovering Implicit Evaluation Criteria from the Web HF f? -`U w DZYH+ RX AI 0 6g:B0ʁEڑD @Q Vx 6y!

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T11:34:42.445714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:34:42.445714Z digest=sha256:6229e9c790b82dd00a21e4bbd8dda44316f72f7406f91fc2217eb1ae0c39409d

Pith citing papers

Observation dece3487-360c-4aff-851e-49ebc34f9eb9 · inbound

BehaviorSFT: Behavioral Token Conditioning for Clinical Agents Across the Proactivity Spectrum cites this paper.

BehaviorSFT: Behavioral Token Conditioning for Clinical Agents Across the Proactivity Spectrum EvalAgent: Discovering Implicit Evaluation Criteria from the Web

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:28:02.922311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T13:28:02.231013Z digest=sha256:b137745175930b5654c582903da6c588e16dedf29ad756122e6050098546c6fd

Observation ce7d7a5e-20c1-46ec-be8d-f92522bbf4ee · inbound

Self-Improvements in Modern Agentic Systems: A Survey cites this paper.

Self-Improvements in Modern Agentic Systems: A Survey EvalAgent: Discovering Implicit Evaluation Criteria from the Web

Reference 1948

Resolution
unresolved
no resolver link, observed 2026-08-02T06:35:59.589082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:35:59.589082Z digest=sha256:ed6cc39090004c005ba12b13a84f61d592709fbc1d6987dcb7c107dd1aa8cad3