Pith. sign in

Paper Citation Record · LEDGER

Improving Keystep Recognition in Ego-Video via Dexterous Focus

As of 17 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2506.00827.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.00827 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:59:06.328656Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3abed9b5-24a3-452f-9d78-d2c639b88a83 · outbound

This paper cites Introducing hot3d: An egocentric dataset for 3d hand and object tracking, 2024.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Introducing hot3d: An egocentric dataset for 3d hand and object tracking, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:10.167863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:03.398227Z digest=sha256:106b6903ef5d8974d1fa3b0856f701942787c06f2f491489b039c8884f38e86f

Observation 3e4ffd7f-5cf5-4159-a741-d0c9a61e8198 · outbound

This paper cites Is space-time attention all you need for video understanding? In Proceedings of the International Conference on Machine Learning (ICML), 2021.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Is space-time attention all you need for video understanding? In Proceedings of the International Conference on Machine Learning (ICML), 2021

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:09.936909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:03.487431Z digest=sha256:2fa932c2fb2ea2d990b41ee59e51c61ecd5f460ca1549f1875cb4782c6b6ef2f

Observation cd0d9544-8ffb-4cc7-b313-fee242e9d6d8 · outbound

This paper cites A dense-sparse complementary network for human ac- tion recognition based on rgb and skeleton modalities.Expert Systems with Applications, 244:123061, 2024.

Improving Keystep Recognition in Ego-Video via Dexterous Focus A dense-sparse complementary network for human ac- tion recognition based on rgb and skeleton modalities.Expert Systems with Applications, 244:123061, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:09.728003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:03.617735Z digest=sha256:5986e3f58c9b44deabbd04f7decac86d96629e5b2a0d0e31b2afe4ea1cfccc96

Observation d01aa204-5528-421e-970c-266d7a7ec5b4 · outbound

This paper cites Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Rescaling egocentric vision: Collection, pipeline and challenges for epic-kitchens-100

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:09.409102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:03.727388Z digest=sha256:6f7ec5925bc6fa508cbb448204c74bbab3be19e51c69f7c0cd445d84e6913d62

Observation 55c20920-77fb-49d5-9731-6115f10a76eb · outbound

This paper cites Activitynet: A large-scale video bench- mark for human activity understanding.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Activitynet: A large-scale video bench- mark for human activity understanding

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:09.175037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:03.863529Z digest=sha256:8a7a13e584a5a716cb46f90438e3a1f04737db2a97adc8768fb841ef62941802

Observation d53db83b-735f-422f-9d84-18dc87ed6150 · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

Improving Keystep Recognition in Ego-Video via Dexterous Focus The” something something” video database for learning and evaluating visual common sense

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:03.974014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:03.974014Z digest=sha256:7aa52478de8f55c531076fdb48b6d1164ad23a6eff90c7a4478e8e9cf4b93d61

Observation 66932c3b-2b11-4283-aeec-d96c5049da21 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Ego4d: Around the world in 3,000 hours of egocentric video

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:08.993075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:04.117291Z digest=sha256:348d9301edfaece70279d0aa89133061ad808ccdca26b58448f4bee496fa44e5

Observation 4c95153d-6583-4f16-9ede-e22832de8963 · outbound

This paper cites Ego-exo4d: Understanding skilled human activity from first- and third-person perspectives.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Ego-exo4d: Understanding skilled human activity from first- and third-person perspectives

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:08.753227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:04.282107Z digest=sha256:3820791309f0ebbf8bff5706b5d6c85b70819104f2286fe63cff43ea7abbe894

Observation 1efd576f-2e94-47ee-95b1-7a71c17383e7 · outbound

This paper cites Jiang, J.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Jiang, J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:08.535971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:04.419006Z digest=sha256:a3d9494fd54331165a123c22dbe8bcaf6a2df7c3ee9681eee677620beb9255fd

Observation 48a33909-d9a7-4168-a6ad-a43e964f66c7 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Improving Keystep Recognition in Ego-Video via Dexterous Focus The Kinetics Human Action Video Dataset

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:04.567274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:04.567274Z digest=sha256:b1d0567ccef735a373ca7a04ea8ebb86dc5329c99c163e031b93c1ed320badbb

Observation 8451dedf-ce70-46ab-a700-ecaed6225350 · outbound

This paper cites Epic-fusion: Audio-visual temporal binding for egocentric action recognition.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Epic-fusion: Audio-visual temporal binding for egocentric action recognition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:08.297259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:04.723588Z digest=sha256:8e15507097635c06d84fb669764c3921fa3747958f048f5f6cd5b9a697157857

Observation 40a7b2ac-a119-4eff-8dc0-629831ba756b · outbound

This paper cites Human action recognition and predic- tion: A survey.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Human action recognition and predic- tion: A survey

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:04.874486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:04.874486Z digest=sha256:d95c4cfe9ea7582a874e75571a773c664cacd258fc587e9e8a4de48fa258ff1d

Observation b0aeda98-7a95-4851-9142-0e617663e98c · outbound

This paper cites X-mic: Cross-modal instance conditioning for egocentric action gen- eralization.

Improving Keystep Recognition in Ego-Video via Dexterous Focus X-mic: Cross-modal instance conditioning for egocentric action gen- eralization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:08.080916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:04.959598Z digest=sha256:3a2cdd607c492bace203a7756eaeed0c36ea43a8bd9ab9fe142439ced4d86def

Observation fadc45a9-2ed8-401f-ba91-ae8cc6ed7a39 · outbound

This paper cites Ego-exo: Transferring visual representations from third-person to first-person videos.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Ego-exo: Transferring visual representations from third-person to first-person videos

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:07.845496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:05.102608Z digest=sha256:15c7cfa840535ad33be473ea6b0fcef1de822fa002762c2c0370dcffb8a694f9

Observation 346dd980-3330-42c3-9f7d-44302af0302c · outbound

This paper cites Egocentric Video-Language Pretraining.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Egocentric Video-Language Pretraining

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:05.218954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:05.218954Z digest=sha256:081f3b23d59cf90d07798e01881837753626fb08b9d31b7398fc1d3b184f6f6a

Observation d544b563-dcdd-462c-892d-56b189b820b2 · outbound

This paper cites Where a strong backbone meets strong features – action- former for ego4d moment queries challenge.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Where a strong backbone meets strong features – action- former for ego4d moment queries challenge

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:07.646633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:05.321906Z digest=sha256:7fb2c579cf77829aaffb298d202a7eeb5b777bd97b2aa174ee6dc1100eecc552

Observation 3d699295-a9ef-48c6-a957-15ee3edeb865 · outbound

This paper cites Egoenv: Human- centric environment representations from egocentric video.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Egoenv: Human- centric environment representations from egocentric video

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:07.399174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:05.427141Z digest=sha256:d62448b6d26473f5f27dbe13db7defe6632dd4b2846ea4aca5b1ea07c1cccbf9

Observation b028e65e-5af7-44d2-847d-cc2e36b35e00 · outbound

This paper cites Project aria: A new tool for ego- centric multi-modal ai research, 2023.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Project aria: A new tool for ego- centric multi-modal ai research, 2023

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:07.179449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:05.530517Z digest=sha256:56f2ccfea93139d652e0843258836c3bd4baeecdb12ef49acb8d736b55e90e70

Observation 947f55c1-1528-4471-91e6-07c289a7e6b9 · outbound

This paper cites EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation.

Improving Keystep Recognition in Ego-Video via Dexterous Focus EgoVideo: Exploring Egocentric Foundation Model and Downstream Adaptation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:05.646267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:05.646267Z digest=sha256:30b1111a2e549b29533d9bf3506575e39c746cc40bab6bb09eecb61196aa8bab

Observation fb824526-af44-4294-9793-660ddbe35d51 · outbound

This paper cites EgoVLPv2: Egocentric Video-Language Pre-training with Fusion in the Backbone.

Improving Keystep Recognition in Ego-Video via Dexterous Focus EgoVLPv2: Egocentric Video-Language Pre-training with Fusion in the Backbone

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:05.781867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:05.781867Z digest=sha256:3b6f577fc19cae39093bc737d236504e708f4e5b65acd556effd26ebb5b5aecf

Observation 79cdaa7a-0386-48ac-9509-6de7f4c6d0d8 · outbound

This paper cites Understanding human hands in contact at internet scale.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Understanding human hands in contact at internet scale

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:06.899474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:05.899839Z digest=sha256:97cb379d7b779fb2a2eda15dea4b2ac5895e69ebd70928cd1d13e7e0b5761d49

Observation e00bff9a-e979-4536-b5e7-5081ae2ed929 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Representation Learning with Contrastive Predictive Coding

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:06.003000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:06.003000Z digest=sha256:d7d26c3c8f86ae0cb51b39408f0317a01f1b9d2ea1eafba4915dca639a72baf9

Observation 348e432f-cd03-416b-8af0-23bf91de4165 · outbound

This paper cites InternVideo: General Video Foundation Models via Generative and Discriminative Learning.

Improving Keystep Recognition in Ego-Video via Dexterous Focus InternVideo: General Video Foundation Models via Generative and Discriminative Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:06.099409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:06.099409Z digest=sha256:8eda666f969d504a669913b10cb4d817050fad2feaa75b14efd2fb98f76837a2

Observation adfd9535-5c8b-4070-a0fc-537de4b1c8a5 · outbound

This paper cites M&M Mix: A Multimodal Multiview Transformer Ensemble.

Improving Keystep Recognition in Ego-Video via Dexterous Focus M&M Mix: A Multimodal Multiview Transformer Ensemble

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T11:59:06.205021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:59:06.205021Z digest=sha256:c659dd4cf0d6334ad3c9c357a01227c4001d7c797a4799a0a57f81d28f5807e8

Observation 25ad6a03-c343-4276-a302-0a57316e7302 · outbound

This paper cites Actionformer: Lo- calizing moments of actions with transformers.

Improving Keystep Recognition in Ego-Video via Dexterous Focus Actionformer: Lo- calizing moments of actions with transformers

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:59:06.664550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T11:59:06.328656Z digest=sha256:06464bacf8a3a94cdcc9efe9ab292194eef4bbf6759a92684645754ac32bb885

Pith citing papers

No inbound Pith citation observations are available.