Pith. sign in

Paper Citation Record · LEDGER

CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.07599.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.07599 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:55:16.115892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:24.102448Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6d295178-0456-4dd9-83e5-f2fb80e74d11 · inbound

SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis cites this paper.

SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:16.115892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:16.115892Z digest=sha256:10bbd903b38ba8bbc4eac66bde664f3ebc84241b09f0f11bdad499eb0579c07b

Observation 9160eeab-babb-4fc3-99d2-a6de7021e998 · inbound

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure cites this paper.

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:40.069631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:27:40.069631Z digest=sha256:281b102c84037726905cf079a590f2a68dbf52f6c7c9cc081cc5f1c8ad6c8b71

Observation f6ef7597-a1e8-4265-aae1-a122bea82018 · inbound

A Practical Guide for Evaluating LLMs and LLM-Reliant Systems cites this paper.

A Practical Guide for Evaluating LLMs and LLM-Reliant Systems CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:55.150430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:41:55.150430Z digest=sha256:2ff4bb012087c8c2bfb304b8a213517cc787f92d437db2695a14deb1e7a485aa

Observation f012fec6-dbe4-4299-8217-d0d6fef50081 · inbound

False Alarms, Real Damage: Adversarial Attacks Using LLM-based Models on Text-based Cyber Threat Intelligence Systems cites this paper.

False Alarms, Real Damage: Adversarial Attacks Using LLM-based Models on Text-based Cyber Threat Intelligence Systems CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:55:32.994446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T07:54:51.925530Z digest=sha256:23184cdd0d6ee2b36cf767d65f253e7eb7f2153d4122429404cc8ab336e5877f

Observation 3ca6a682-1b67-478e-a1cf-776abf18eac4 · inbound

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation cites this paper.

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.763400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T04:37:33.942379Z digest=sha256:ca82631249d618889003743644717a5beb31ba437d6c229709541683d4a0ee9e

Observation 084bd0b0-62d6-400c-bd10-d6b3261f90af · inbound

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting cites this paper.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.064442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.064442Z digest=sha256:3f05e31f82662dab46aab6c765f7db55d7bffe6d84c45cf95ecab0d242392295

Observation f36749e2-0a77-47fa-b9e9-3261b9bb0c94 · inbound

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence cites this paper.

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:47.986398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:47.986398Z digest=sha256:edb816d41d7a025cfb954af0aa950a6ee1706f308fedc19045f8be463ca2a5f8

Observation a701c1a1-be47-4628-a3aa-b149cbabf5f4 · inbound

Dynamic Cyber Ranges cites this paper.

Dynamic Cyber Ranges CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:52.816852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T03:04:03.611481Z digest=sha256:03ac1d326c3fd992519a82c46d7c3706aee38536205aab93488f8b73e748f7ed

Observation 0028edf9-0d23-4139-b133-f812aff1a3ec · inbound

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents cites this paper.

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:05:02.502191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T05:00:26.307473Z digest=sha256:93da55f3e5c8613933814ac5945684c57edd251607a601a7a5504cd38068c619

Observation 6a284f94-2916-4a7f-9a6f-b2d187e1342e · inbound

Cybersecurity AI (CAI) Dataset cites this paper.

Cybersecurity AI (CAI) Dataset CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.887469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T12:07:49.656453Z digest=sha256:6bc9d5b76f758b1f411afabbf94093c53e852032ea1514a83609139a71ae3b49

Observation b523346f-d1aa-492a-99e5-00b393894364 · inbound

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning cites this paper.

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.105488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T14:03:41.903170Z digest=sha256:ee9c7a1b32aa82c2f3acb2c70a7e1fb415073bf2b0acfa93bec3aef9d756756f

Observation 3073133f-b7f5-4a69-a307-ee116e01a7b7 · inbound

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense cites this paper.

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T10:23:20.740447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:23:20.740447Z digest=sha256:5c0546acdca06e224ce16b2c32ff4d97c98572471d5caf98293301a3c217dca8

Observation 8848ad00-0978-448f-8d2f-90257fdb5331 · inbound

Antares: Foundation Models for Agentic Vulnerability Localization cites this paper.

Antares: Foundation Models for Agentic Vulnerability Localization CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:50:50.560574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:50:50.560574Z digest=sha256:07176abd164f201ad2e685769a7f6892aeeb96faaa774707bbc11c0a0dc3a060