Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:15:40.443879Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 5 inbound Pith citation observations for arXiv:2504.16027.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:15:40.443879Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:10:06.520779Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-16T07:22:31.204263Z
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fb8905c4-dcd8-4010-98cd-93dbaf7f9ed4 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Coding by Design: GPT-4 empowers Agile Model Driven Development
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24f86a0b-968c-4614-a071-d36255375d0b · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Evaluating Large Language Models in Detecting Test Smells
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8d06b7e-9688-49bb-bc04-b9eb0b08fdf5 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 How Propense Are Large Language Models at Producing Code Smells? A Benchmarking Study
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a07f0e5-c5e9-4509-a6be-83b625b0e0b7 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Are sonarqube rules inducing bugs? In 2020 IEEE 27th international conference on software analysis, evolution and reengineering (SANER), pages 501–511
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 791c30f9-3841-4da4-9c8d-9b99a15567a3 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Code smells
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c8a25a31-156c-4489-9809-05d30c0b66c8 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Token-Hungry, Yet Precise: DeepSeek R1 Highlights the Need for Multi-Step Reasoning Over Speed in MATH
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fccf7bbb-24b7-4e93-8749-f80df37bc54f · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 DeepSeek-V3 Technical Report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e845e1-9557-4a24-a77e-4580133b5488 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Code Smells in Machine Learning Systems
Reference 2006
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d07daf0-8c55-4afd-a766-2a805a40ee2e · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Causes, impacts, and detection approaches of code smell: a survey
Reference 2016
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 80fd2e32-b144-4b00-9c33-776fb8c35c64 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Unresolved cited work
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c6b4c137-0512-471f-b1b3-601e0ac00e62 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 Unresolved cited work
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1ba3ae49-a05f-4ab7-9734-54e5ba929900 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 ChatGPT as a Software Development Bot: A Project-based Study
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3697bb99-00f1-4bae-848b-541a080f89d3 · outbound
Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3 GPT-4 Technical Report
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ea1dce1-7027-46f5-ac13-12847ff76092 · inbound
Are We SOLID Yet? An Empirical Study on Prompting LLMs to Detect Design Principle Violations Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b2aeaaf-6b14-4f49-a423-6fe8031865a7 · inbound
Towards Generalizable Reasoning: Group Causal Counterfactual Policy Optimization for LLM Reasoning Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4bef57b6-c9e3-4791-bd69-b67b2f4bfe64 · inbound
Model Context Protocol (MCP) Tool Descriptions Are Smelly! Towards Improving AI Agent Efficiency with Augmented MCP Tool Descriptions Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ab5abea-5ee6-4819-88d6-1acc5e1c087d · inbound
DynamicsLLM: a Dynamic Analysis-based Tool for Generating Intelligent Execution Traces Using LLMs to Detect Android Behavioural Code Smells Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c91ab388-4f3d-49ae-8a5b-ecda5e0ea9db · inbound
Mitigating LLM Sycophancy in Code Smell Detection Using Evidence-Guided Reasoning Prompts Benchmarking LLM for Code Smells Detection: OpenAI GPT-4.0 vs DeepSeek-V3
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.