Pith. sign in

Paper Citation Record · LEDGER

A Comprehensive Study of Deep Video Action Recognition

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2012.06567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.06567 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:03:08.926468Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:23:51.204442Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 735d3440-32fe-49d3-b768-2f8d059bb947 · inbound

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance cites this paper.

Low-Latency Video Anonymization for Crowd Anomaly Detection: Privacy Versus Performance A Comprehensive Study of Deep Video Action Recognition

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-23T18:28:19.001700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T18:25:44.799112Z digest=sha256:6cb984bbfd81d57398abfedcbd4a293be8b0b00015e1e7420967509a73803c3d

Observation a61ab16e-617c-4862-a10d-3ac3da5b4032 · inbound

What to Do Next? Memorizing skills from Egocentric Instructional Video cites this paper.

What to Do Next? Memorizing skills from Egocentric Instructional Video A Comprehensive Study of Deep Video Action Recognition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:08.926468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:03:08.926468Z digest=sha256:6c266569a257dc3e2247c88553992c526b1edd3a858f16857e394e074caa17b0

Observation a1115ff8-dc7e-4c5d-b44e-ff31a2438b41 · inbound

A Survey on Video Temporal Grounding with Multimodal Large Language Model cites this paper.

A Survey on Video Temporal Grounding with Multimodal Large Language Model A Comprehensive Study of Deep Video Action Recognition

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T23:32:17.622367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:32:17.622367Z digest=sha256:dfdc7b3613aa0610f9c394b36e207c66b0eea5e28e3f06663a51eac0d4dc2a57

Observation 168a7747-f835-4501-906a-29f09152926e · inbound

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization cites this paper.

Leveraging Vision-Language Large Models for Interpretable Video Action Recognition with Semantic Tokenization A Comprehensive Study of Deep Video Action Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T05:12:54.460881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:12:54.460881Z digest=sha256:f040de1484f9908e8ab36d5d0c214fcf0d702a8c36a65855117cd7affc26f9d1

Observation 9d01e54a-3ce2-40d0-aff1-53a055647e03 · inbound

AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning cites this paper.

AVATAAR: Agentic Video Answering via Temporal Adaptive Alignment and Reasoning A Comprehensive Study of Deep Video Action Recognition

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:20:11.813025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T20:18:19.580156Z digest=sha256:c6653ce31e7a70b96571d4cf2795c7dd016a19f31a580d4e58fe8d7a408b7dcd

Observation d9a1d30a-67d0-4183-bb3c-e1b2329784ec · inbound

TRUST: Efficient Abdominal Trauma Recognition via Image-to-Ultrasound-Video Transfer Learning cites this paper.

TRUST: Efficient Abdominal Trauma Recognition via Image-to-Ultrasound-Video Transfer Learning A Comprehensive Study of Deep Video Action Recognition

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:23:51.205822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T05:08:02.242341Z digest=sha256:112ec14bcae302d1d766d9aaf62367290064b934dc3c831a99aa63f0a575d003

Observation 8c62c160-30cc-4b2b-b040-a835b7ec13d9 · inbound

Efficient Video Dataset Distillation via Cluster-Guided Prototype Blending cites this paper.

Efficient Video Dataset Distillation via Cluster-Guided Prototype Blending A Comprehensive Study of Deep Video Action Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T22:10:41.654548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:10:41.654548Z digest=sha256:7c12a4850b458a31acc751ed8128967f2554b56dc198bf53c60725b3bc530ac6