Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:23:32.774387Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 3 inbound Pith citation observations for arXiv:2507.11059.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:23:32.774387Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T23:01:35.774091Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T07:04:44.553235Z
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 522319be-0f16-4023-a4d8-71ab6f5f9fd5 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab8db74e-8677-4fcd-aa93-0816f2d8ef0f · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d597f088-f028-4633-961e-c630e8acdb65 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks SWE-Bench+: Enhanced Coding Benchmark for LLMs
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05a5ca85-3c73-4f89-91e0-8424e04323e2 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88c76691-db89-4dd7-8a6c-9620c427d861 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Qwen2.5-Coder Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12e7bcb-9458-4b80-b0d9-2145d51f405d · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a49d633-9b25-4c0f-827a-bdc085364b3b · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks SWE-bench: Can Language Models Resolve Real-World GitHub Issues?
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d40b88ce-95d8-41bb-b4e0-dba096edcfa4 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Training Software Engineering Agents and Verifiers with SWE-Gym
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 385ade8c-74be-42a5-93d6-d511c1c86999 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaa81217-c354-40cf-bc03-7d6e0cc5a7bc · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Qwen3 Technical Report
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b50e8b67-0ba7-4a18-b4b6-31a3cd5258ab · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Qwen2.5 Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba76c1e7-a3b9-4e89-b264-d8b8dfc199b5 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks SWE-smith: Scaling Data for Software Engineering Agents
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47d99099-fcc2-423c-96d7-957e3e954dc6 · outbound
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks Multi-SWE-bench: A Multilingual Benchmark for Issue Resolving
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 361e4e2a-fec5-4709-b57d-e0f611e9d605 · inbound
RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ca14c0ea-849c-4e29-b863-c20cd752f562 · inbound
RuBench: A Repository-Level Agentic Coding Benchmark with Natively Authored Russian Task Specifications SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ace0f74-a7fd-4d09-97ad-8c1d4715aed4 · inbound
The Horizon Gap: Planning, Memory, Execution, Training, and Evaluation for Long-Horizon LLM Agents SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.