Pith. sign in

Paper Citation Record · LEDGER

Comparing Learning Paradigms for Egocentric Video Summarization

As of 10 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2506.21785.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.21785 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:23:43.243274Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact4
  • verified fuzzy1
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9c8b10c0-3408-4d37-b3b5-e2fca4dcb349 · outbound

This paper cites UniVTG: Towards Unified Video-Language Temporal Grounding.

Comparing Learning Paradigms for Egocentric Video Summarization UniVTG: Towards Unified Video-Language Temporal Grounding

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:44.366891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T22:23:41.171290Z digest=sha256:9b57fcfcfe60517f74ddd82510ed2d09cf5c01d17386245891e88f1a6a748a1c

Observation 90c0f3b1-5916-4edf-8bfe-b087c15b290a · outbound

This paper cites VideoLLM-online: Online Video Large Language Model for Streaming Video.

Comparing Learning Paradigms for Egocentric Video Summarization VideoLLM-online: Online Video Large Language Model for Streaming Video

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.274752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.274752Z digest=sha256:542fe8f62ae284e82369c6967463fc985e47e97d96b1ce55494d8fc31b8b53a3

Observation 60fd24f9-7f8e-4d8d-bade-a4d8a55cc9a8 · outbound

This paper cites VideoMamba: State Space Model for Efficient Video Understanding.

Comparing Learning Paradigms for Egocentric Video Summarization VideoMamba: State Space Model for Efficient Video Understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.396361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.396361Z digest=sha256:b551ca39c8fd443801788212488fe1b2f207b50a3e9d94b633073733830a8bd0

Observation bdf647d9-3f6f-4ac8-bb9d-c1c7908b92a9 · outbound

This paper cites Mamba: Linear-Time Sequence Modeling with Selective State Spaces.

Comparing Learning Paradigms for Egocentric Video Summarization Mamba: Linear-Time Sequence Modeling with Selective State Spaces

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.495758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.495758Z digest=sha256:eec757aab0ec6f4e6396518fb8638694063f6abedbe7c3610777119e147c4a47

Observation bd7e80c6-f1a6-4542-be12-6e0b683c42a7 · outbound

This paper cites Is Space-Time Attention All You Need for Video Understanding?.

Comparing Learning Paradigms for Egocentric Video Summarization Is Space-Time Attention All You Need for Video Understanding?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.597659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.597659Z digest=sha256:3da0c1c7845cd9a8ef7fefcbdc2d49718a588b0af130c86cf76b0615ecb4abc8

Observation 3ae6917d-ea86-4819-bad3-0222acd15ed4 · outbound

This paper cites Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives.

Comparing Learning Paradigms for Egocentric Video Summarization Ego-Exo4D: Understanding Skilled Human Activity from First- and Third-Person Perspectives

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.740633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.740633Z digest=sha256:eb5b239e78bf4a6b80e9cdb19cd9fb2a828f26b4ce461b0d0970965520943b81

Observation 248d10b3-b2d1-4c0d-b58e-6c6220c682b4 · outbound

This paper cites Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization.

Comparing Learning Paradigms for Egocentric Video Summarization Shotluck Holmes: A Family of Efficient Small-Scale Large Language Vision Models For Video Captioning and Summarization

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:44.044371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T22:23:41.864051Z digest=sha256:1e509030bdbbe51d7454029a7049c55660df39cdc5ecbdf1f6225e2ef7219334

Observation e5aed2b2-69a5-42fe-87bd-b4dce7efe8ad · outbound

This paper cites Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos.

Comparing Learning Paradigms for Egocentric Video Summarization Shot2Story: A New Benchmark for Comprehensive Understanding of Multi-shot Videos

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:41.969901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:41.969901Z digest=sha256:8418cc8e239152ddd333ff8409d35d0c11d4b875fcf37f2cd016b6706ef11f84

Observation 5e87f939-0cdb-4a1c-bef7-4e88cf736c47 · outbound

This paper cites Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions.

Comparing Learning Paradigms for Egocentric Video Summarization Advancing High-Resolution Video-Language Representation with Large-Scale Video Transcriptions

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:43.860557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T22:23:42.126947Z digest=sha256:5320b3ec77cbb8d81c6c3ff62ed16adb4f489eb424626f9a0dcd3b63e3313c0d

Observation 504e0a5d-bbdc-46a6-9602-16acb08fd1e5 · outbound

This paper cites TransNet V2: An effective deep network architecture for fast shot transition detection.

Comparing Learning Paradigms for Egocentric Video Summarization TransNet V2: An effective deep network architecture for fast shot transition detection

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.246470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.246470Z digest=sha256:175154e46970dfd0aecaaf44ecd0c011e4c7c38e479b1a7ff5cac001526a7865

Observation 0335ffe1-ba69-4b56-90af-e15f5577fc2e · outbound

This paper cites Ego4D: Around the World in 3,000 Hours of Egocentric Video.

Comparing Learning Paradigms for Egocentric Video Summarization Ego4D: Around the World in 3,000 Hours of Egocentric Video

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.360607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.360607Z digest=sha256:d9853fd5ad28e38046ba281b70c95c449edbd8b3db764277d558a6e435d6e38b

Observation 2f91c3c5-a537-41c3-a653-86e22eebc81f · outbound

This paper cites From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting.

Comparing Learning Paradigms for Egocentric Video Summarization From Sparse to Dense: GPT-4 Summarization with Chain of Density Prompting

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:23:43.644497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T22:23:42.496883Z digest=sha256:cda50d78058146e0643002c52784b51e636fe8ecec4b391225310767cd97b5fc

Observation ece7af2f-5675-4b0a-8fb9-4d57d673dfd8 · outbound

This paper cites GPT-4 Technical Report.

Comparing Learning Paradigms for Egocentric Video Summarization GPT-4 Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.657182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.657182Z digest=sha256:cf4d281af3308a01a697db8390f14f174970cc94c34a616265a70f736b7f84d1

Observation 998b9a53-fd36-4a54-a714-cae86b08c33a · outbound

This paper cites Cluster-based Video Summarization with Temporal Context Awareness.

Comparing Learning Paradigms for Egocentric Video Summarization Cluster-based Video Summarization with Temporal Context Awareness

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.816049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.816049Z digest=sha256:16e36e810d2f11f612f89813d8ff76f14955066cda63067baf9151b59cd1a812

Observation a2d84fed-fae2-4413-bc87-bb203fad4ccf · outbound

This paper cites Creating Summaries from User Videos.

Comparing Learning Paradigms for Egocentric Video Summarization Creating Summaries from User Videos

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.917546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.917546Z digest=sha256:c7b857ceb3b751d9b6f4014096d87da9dce785a2a341ede6621be6efa4803969

Observation 9d10f3e3-ba32-4d81-88d6-3cf103de00d9 · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

Comparing Learning Paradigms for Egocentric Video Summarization Sigmoid Loss for Language Image Pre-Training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:42.985153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:42.985153Z digest=sha256:b4023421323d3e68d8adcea95159618f6a09320344c4826b341acf1705d09ed3

Observation e64a774c-1eea-46ea-987e-18ebdce0fd69 · outbound

This paper cites BIRCH: An Efficient Data Clustering Method for Very Large Databases.

Comparing Learning Paradigms for Egocentric Video Summarization BIRCH: An Efficient Data Clustering Method for Very Large Databases

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:43.101087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:43.101087Z digest=sha256:ea4128e286d2493144a705a4012c91134ae2c1b47b6b01cdcb04a073e28714f6

Observation 46b74914-c675-4384-9cba-1e2f60511c7d · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Comparing Learning Paradigms for Egocentric Video Summarization Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:23:43.159750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:23:43.159750Z digest=sha256:cedbb5e4a27deb2518ae97365a1dcff67497914b5f83702fa1d1d27bb30c89e7

Observation 6b7f6f11-5185-4657-9eca-d3c950b1ec78 · outbound

This paper cites Nov 4, 2024, https://build.nvidia.com/nvidia/video-search-and-summarization 13.

Comparing Learning Paradigms for Egocentric Video Summarization Nov 4, 2024, https://build.nvidia.com/nvidia/video-search-and-summarization 13

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:23:44.649217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T22:23:43.243274Z digest=sha256:feefa9b82cb39ee3b7d869ac1b1e92222f9234044ff32c3be8debdda2d0a5a1a

Pith citing papers

No inbound Pith citation observations are available.