Pith. sign in

Paper Citation Record · LEDGER

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research

As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2507.13300.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13300 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:30:31.148470Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact3
  • verified fuzzy1
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a639b420-0009-4b18-9092-cfeb21672a31 · outbound

This paper cites URL: " 'urlintro :=.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.863012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.863012Z digest=sha256:e861a965277673a42081938ea48d7e51c36df0c29e2b6a159811277a0fa03164

Observation f011d55d-8f72-4317-aec8-67d9cde1df15 · outbound

This paper cites write newline.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.867039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.867039Z digest=sha256:3c720e065ffb7484971dd62d3c1718b7de40b9f3fff44989d5626e4d0ac4746a

Observation f62fbbe1-67c3-4ba0-9888-9dbece2476eb · outbound

This paper cites LitLLM: A Toolkit for Scientific Literature Review.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research LitLLM: A Toolkit for Scientific Literature Review

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.874417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.874417Z digest=sha256:a5d5ededeb5a0c2fc9b60742df0ac879f87179953634eaf0f57088c180392a6c

Observation 1e0f122c-bade-4270-9f43-b2e7b4882fb0 · outbound

This paper cites The Llama 3 Herd of Models.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.878617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.878617Z digest=sha256:d3ef52a37e87d57ff6712f3a457fbeb1949951e9a4b2c5b2d89accbb65360b8e

Observation c5f3c8dd-ddfc-4b47-9730-c0127cfe2a50 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.882304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.882304Z digest=sha256:8eff55d86f47404b3174359934c8f59c6b09ba953ce97a55ef02466755c85d19

Observation a99d8233-c32a-458a-a8d9-a266c4f16682 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.885445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.885445Z digest=sha256:f965749ba04ab47b7438755e166cef05c5a5b88c958d8866c96d0aa53b0ec05f

Observation 4300e2b8-afb0-47d2-9086-758216bafeca · outbound

This paper cites AI4Research: A Survey of Artificial Intelligence for Scientific Research.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research AI4Research: A Survey of Artificial Intelligence for Scientific Research

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.888470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.888470Z digest=sha256:aeb9cd4efdbb0027ce86a4cc61c3c9ede4f9e01fbadba976e1ef014043dc1e27

Observation 806d738b-fb99-454e-b8b4-287598e09d3c · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.891644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.891644Z digest=sha256:e14d147b22c92630f2a7a9d6c019c56fd3c01c5d841dcf2ef852e62f5f88a29d

Observation 2d8c2b77-5c02-4ba8-a27f-920a0e0d3738 · outbound

This paper cites MARG: Multi-Agent Review Generation for Scientific Papers.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research MARG: Multi-Agent Review Generation for Scientific Papers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.895146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.895146Z digest=sha256:f4f00096774075e77b8f0e40cf2995589bd9db9be21efe03d962d858046459cc

Observation c7fc8dbc-2a70-4e48-b4b7-58745b49fab2 · outbound

This paper cites Smith, and Matt Gardner.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Smith, and Matt Gardner

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:30.997967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:30.997967Z digest=sha256:5e1856479b74a127709da6d3f291c5042a5e5a8d5d4314b437187e35b8888d3e

Observation 9193c18d-bee3-486d-9062-18ffdb7b4672 · outbound

This paper cites DeepSeek-V3 Technical Report.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research DeepSeek-V3 Technical Report

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.001582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.001582Z digest=sha256:5a85c98c9f49a076e018aa608fe540e4b315b181f5f521e92a82dff969a91022

Observation 7417691f-2df2-4db1-a102-e9ec3e23dcbc · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.005307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.005307Z digest=sha256:e59e679e039ff5a2fffe6caa29640bfce7ae7a1435e5c13f17823fe517f65a8a

Observation 1ee6742d-998c-47f2-9df6-b13f388f3c8d · outbound

This paper cites LLMs Assist NLP Researchers: Critique Paper (Meta-)Reviewing.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research LLMs Assist NLP Researchers: Critique Paper (Meta-)Reviewing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.008892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.008892Z digest=sha256:11ad02929118dd887f3e96dc27f94416d5646300fe0b6958cbc3a4d82d689780

Observation 437f1899-b34d-40ac-b34b-7e818cbf294b · outbound

This paper cites Fabbri, Wojciech Kry \'s ci \'n ski, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Fabbri, Wojciech Kry \'s ci \'n ski, Bryan McCann, Caiming Xiong, Richard Socher, and Dragomir Radev

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.012316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.012316Z digest=sha256:691266110b2fe5d0b699731af4cb07e1ee98069f55757c29fa83e298debd6789

Observation 083ce5ba-32b1-41d0-ba6a-70f0f2857533 · outbound

This paper cites Sengamedu, and Christos Faloutsos.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Sengamedu, and Christos Faloutsos

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:30:31.822196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.015875Z digest=sha256:6052419535fb974bcc35769dc0015e4a6a88d98db1eead2bd585d97dbd3df1d2

Observation 5d8e6790-81d9-416a-a3a5-397944bb2b61 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.019280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.019280Z digest=sha256:c5d641d07a5028aa67583f4c904a9f26a27381de59302e72bd154e58da256e2b

Observation 5bff5dd6-063f-4307-b1f1-4df0fd3d6c46 · outbound

This paper cites Mixtral of Experts.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Mixtral of Experts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.022702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.022702Z digest=sha256:8f4f1cdf8182b98fa0b1ce229baee73d3b548003902d9b83c86b755bf8b56093

Observation 4ce7d73f-2297-4606-859d-76ffa23b4283 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.811334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.026139Z digest=sha256:b1c541e6ad8ad43f6a2d6214e671c47f96028407f456bcbc73967fa000d4f020

Observation 4833b901-f4ff-4c94-9853-57b9bafc2ea0 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.029330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.029330Z digest=sha256:c954bfb0efd6cd782b8210e6be9a486dea049422bdcafaf0d551312091857d86

Observation 85bf35d1-8570-41fd-9c0c-a0c74d35e5c0 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.032766Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.032766Z digest=sha256:bfc866e6b310b86a994ccb0c3b22b4f99f5087833c54f55e60fdeb7544eefdcb

Observation 185b38a5-74db-4f83-b4fb-eea0ac12af20 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.036070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.036070Z digest=sha256:8891873412157e743920a26a40d71236ec6265ab34b5e0bf2e54a17cf0c53e9b

Observation c6f21ed6-1294-43ac-9006-709465b70578 · outbound

This paper cites ML-Bench: Evaluating Large Language Models and Agents for Machine Learning Tasks on Repository-Level Code.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research ML-Bench: Evaluating Large Language Models and Agents for Machine Learning Tasks on Repository-Level Code

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.039397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.039397Z digest=sha256:507c5a0148351673cea5ab5610bca50f7bbc6a913944763d43223348d60e7ea1

Observation 665073b8-e824-4ccf-b649-b982a6ed0339 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.042531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.042531Z digest=sha256:7207c4924018b63986edcc0a972358453ea588ad39c7dfe6cfc2241f269215a2

Observation f0f220d6-e3a8-4051-87ac-fe2221f36dcc · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.798588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.046147Z digest=sha256:d649fbeceff618d8db8098a35d44b44139114f532de6bd1d6f69a7193f32dfde

Observation 25cdec61-3801-45a5-a4a7-23e2d8b87f94 · outbound

This paper cites The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.049248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.049248Z digest=sha256:348aa90f20ae55d07324070140c66742ab7292509ad12fa3b6c9ae833fdb132e

Observation 29b932a9-8b9e-4da5-9616-39c6dfbc5871 · outbound

This paper cites Data-driven Discovery with Large Generative Models.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Data-driven Discovery with Large Generative Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.052687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.052687Z digest=sha256:c52dd377510055f74b6dfba99862e52e3be5859ebed898f04c836e6d2e86c362

Observation 60c78680-721c-4a0d-89bd-42765d2a061b · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.787440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.055903Z digest=sha256:28999cd25152a0ed0a4c614cf2e8773d1bd4f5ddc34aea890f84071fb6f8643b

Observation 808d6a98-936f-4299-b0e2-f39aab016ff8 · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.059038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.059038Z digest=sha256:36ce78597547aa689636df18aa5141072644304a2aede11ecec328c2fc669370

Observation 688d34cc-d254-4706-ab4e-4408e45bc378 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.062412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.062412Z digest=sha256:080666d8ff0250e02e1322bc6ffaac50dbd5ce1a2224d718ff83a5615a3517de

Observation 8f0da463-dfff-4cca-86f1-0e4f962c02e8 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.770380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.065403Z digest=sha256:5375b050decc48aef880052cd10dd236f24c282e4c81716dff1659659feee334

Observation a0f174d5-69ae-4218-8f7f-464315482aa1 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.759968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.068701Z digest=sha256:ceb847c8803921b23cdf353bd88f305202c36f4d9fab9e838fe9ede83d498cfd

Observation fce17fcd-67f8-4626-92bf-62c5ddc75395 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.748859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.071806Z digest=sha256:7df5485212080084f8c380d616e59b04ce37db1090c8e95492a028cc367c316c

Observation 0e27f756-18d3-408c-ab87-faeb6d0df65f · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.075285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.075285Z digest=sha256:0b89bbeb0035577bff5da175e739fa31e5ba8f2ad841d2f05fa82ba2b79283a9

Observation f1399d7f-bb78-4f69-acec-02d4b62a02d0 · outbound

This paper cites Table Meets LLM: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Table Meets LLM: Can Large Language Models Understand Structured Table Data? A Benchmark and Empirical Study

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.078561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.078561Z digest=sha256:0c4ed02b3155984e026187319ab25cfc9b4d392c079ab4ab3c7fe6a1b994430f

Observation ddecfe4f-6646-4ec9-be3e-e20959d4e546 · outbound

This paper cites Peer Review as A Multi-Turn and Long-Context Dialogue with Role-Based Interactions.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Peer Review as A Multi-Turn and Long-Context Dialogue with Role-Based Interactions

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.082021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.082021Z digest=sha256:54ded244d3f4c601885e05721691c5c92cb5309f6c333cdb74cc9d964f3bfa7b

Observation 96dbbd43-b39a-4731-8ece-5fe516d7e906 · outbound

This paper cites Gemma 3 Technical Report.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Gemma 3 Technical Report

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.085794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.085794Z digest=sha256:83f9522bd155ec9419ee942a8c787387192cdbccf16843008da149025ce0073e

Observation c70b9c7c-b688-4624-9d70-de818087afb0 · outbound

This paper cites Qwen3 Technical Report.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Qwen3 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.089625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.089625Z digest=sha256:32d26e1088297e9787fea173bc799d4ddde7ca48206bac05a5d452101cb77eb8

Observation 9586f3c8-ffe0-4981-932a-9520f2580e60 · outbound

This paper cites SciVer: Evaluating Foundation Models for Multimodal Scientific Claim Verification.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research SciVer: Evaluating Foundation Models for Multimodal Scientific Claim Verification

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:30:31.399169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.093234Z digest=sha256:957e89dd3763ce52d9841efba2465a915a1af7acbe8b274af7c09933dd8c99eb

Observation 19a818ec-0e27-4e39-886a-f8ac522d13db · outbound

This paper cites SciMON: Scientific Inspiration Machines Optimized for Novelty.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research SciMON: Scientific Inspiration Machines Optimized for Novelty

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.097289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.097289Z digest=sha256:6841ea271b8d20879aba2e8aabdc52e6def84f2984759054c46f2c32634a16cc

Observation da8c7d2f-da3d-4060-b1ab-d90fe9923a1c · outbound

This paper cites SurveyAgent: A Conversational System for Personalized and Efficient Research Survey.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research SurveyAgent: A Conversational System for Personalized and Efficient Research Survey

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.101369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.101369Z digest=sha256:7f5ea102654a2fc2915247ba6d552f523152079bd0964d2a997a0dcb57a32309

Observation 2b409e93-3539-4596-9a51-bfa40875a554 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.730606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.105540Z digest=sha256:0438d5428e159afd4f2901762ae6daf8e1027d505cf409e2b67b188a68dc7e5a

Observation 5d53540b-a46a-4849-98c9-a0370e2977e5 · outbound

This paper cites KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.111256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.111256Z digest=sha256:e00e518c456e395ea53f46f3cd9438b2a44ddab95c7bcca7e2dc583188c6dad6

Observation 725bddb8-efc5-4e3d-9f2f-b603cf1cc2e0 · outbound

This paper cites Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Can LLMs Identify Critical Limitations within Scientific Research? A Systematic Evaluation on AI Research Papers

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:30:31.354775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.115395Z digest=sha256:c9858e626a10acbb83b30fc7e1435beb0edf4af3d5a305647b9157bbebef9227

Observation 428c9335-fac2-40f6-9ff1-21fe7f418772 · outbound

This paper cites Qwen2 Technical Report.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Qwen2 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.119728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.119728Z digest=sha256:613a382f3825f9583961bab49cc90342fdc76a7a1edb5fbb36979192e7d9aad3

Observation f86dadd1-e385-4c20-bacc-35486fc1a6cb · outbound

This paper cites SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.123653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.123653Z digest=sha256:56b2a809c81bbcd0a4c73b63ee7b3410867740c7d549b1d9132d22efd40eef4f

Observation 1957df79-f98e-4708-a777-28755761d020 · outbound

This paper cites Griffiths, Yuan Cao, and Karthik R Narasimhan.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Griffiths, Yuan Cao, and Karthik R Narasimhan

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.127880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.127880Z digest=sha256:9025e2f2f8d0762b53bf5b5abea3444412d3361ea539b7e68f14a73ebee10b76

Observation 8a19b5a3-592f-4cd4-9ae9-fe986ef0618a · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.131703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.131703Z digest=sha256:688bca87420ff3193a1341f97267087f0768ae66967681f7cebf68a31e238e70

Observation ead4df80-43c5-4fa2-9cd1-94d2e2f83fb5 · outbound

This paper cites Can Multimodal Foundation Models Understand Schematic Diagrams? An Empirical Study on Information-Seeking QA over Scientific Papers.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Can Multimodal Foundation Models Understand Schematic Diagrams? An Empirical Study on Information-Seeking QA over Scientific Papers

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:30:31.315034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.136261Z digest=sha256:db58102dd9200ab8deb252b1e366d6f25d7eae77b6d4aaa892e325010e90fe30

Observation a10a905c-0f4f-43c5-908f-2607ed074f1d · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.140179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.140179Z digest=sha256:fb615196bc5927ad553956fe5bbf5a0b38629df8e97af8e8ccc56aa8535ed16c

Observation c5de9c4c-c617-430d-aa44-2ee374e800c4 · outbound

This paper cites an unresolved cited work.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:30:31.707216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T16:30:31.144243Z digest=sha256:c00f37c63b8ffd5119f90a069b8b125f111d98fb027a504047a963c4b794cddf

Observation 87a32eb7-089c-40a5-b0f8-85786feeb8a2 · outbound

This paper cites Hypothesis Generation with Large Language Models.

AbGen: Evaluating Large Language Models in Ablation Study Design and Evaluation for Scientific Research Hypothesis Generation with Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:31.148470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:31.148470Z digest=sha256:1a64b17cb119eca7e3e792c15b0cdf185384628b685691db153b73a656ac4820

Pith citing papers

No inbound Pith citation observations are available.