Pith. sign in

Paper Citation Record · LEDGER

Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.13499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.13499 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T00:35:56.017831Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T19:45:36.483341Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5a9d68c8-e368-4009-814b-395c74a988e7 · inbound

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision cites this paper.

FlashAttention-3: Fast and Accurate Attention with Asynchrony and Low-precision Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:45:36.485256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-20T19:45:36.337956Z digest=sha256:bc9092b043dcda53d2c6b9dd7304f2117f966f5a2c15bdb950bdaf88cce83970

Observation 61ec8c13-443f-4801-8fe9-0deeedd598cc · inbound

FlashAttention on a Napkin: A Diagrammatic Approach to Deep Learning IO-Awareness cites this paper.

FlashAttention on a Napkin: A Diagrammatic Approach to Deep Learning IO-Awareness Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T22:38:43.152367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T22:38:43.152367Z digest=sha256:275c56834aa7330aedaacb9be326d64631f0f5cf4164773535b5f30f48e27746

Observation 8747a86c-45c4-4c0b-8518-82bbe5afae16 · inbound

Apple vs. Oranges: Evaluating the Apple Silicon M-Series SoCs for HPC Performance and Efficiency cites this paper.

Apple vs. Oranges: Evaluating the Apple Silicon M-Series SoCs for HPC Performance and Efficiency Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T19:49:45.811059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:49:45.811059Z digest=sha256:38e338c158d23bd40551aee7ba0f49defe16324547a5aa0e783fb49fd9f272e9

Observation 0ae02f5b-1f98-41cb-a991-bf2123def3f3 · inbound

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model cites this paper.

DECA: A Near-Core LLM Decompression Accelerator Grounded on a 3D Roofline Model Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:56.263912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:56.263912Z digest=sha256:79daffa819f7fe1daaac0df0e9caea032fa79ccd9bfbd4f17c9986d2fb45e62c

Observation 4b94aedc-74d3-4318-afc3-1ef6256e14b6 · inbound

TriADA: Massively Parallel Trilinear Matrix-by-Tensor Multiply-Add Algorithm and Device Architecture for the Acceleration of 3D Discrete Transformations cites this paper.

TriADA: Massively Parallel Trilinear Matrix-by-Tensor Multiply-Add Algorithm and Device Architecture for the Acceleration of 3D Discrete Transformations Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T22:04:36.684526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:04:36.684526Z digest=sha256:e4c7edbe11909f3e8a7b7fb9a8debc5b4b753a5b05e59b2d1df6da2cd5f3b372

Observation 8ea17b5b-9e43-4cde-b566-158abff5b0e5 · inbound

Spec Sheets Are Not Kernels: An ISA- and Source-Level Audit of INT8 Availability on NVIDIA Blackwell Ultra cites this paper.

Spec Sheets Are Not Kernels: An ISA- and Source-Level Audit of INT8 Availability on NVIDIA Blackwell Ultra Benchmarking and Dissecting the Nvidia Hopper GPU Architecture

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T00:35:56.017831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:35:56.017831Z digest=sha256:2eb1d018291e59ada97e44ada479107b549ab2f1b00e18ff176abc257c41f5bb