Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:37:09.891480Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.03549.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T04:37:09.891480Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e39f4487-2931-45ab-926b-e28f2997a924 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The low-rank bottleneck is vital, because it may lead to that many rows of the attention map are seriously homogenized
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 30e31daa-6937-46b5-b0b1-25fcb9582777 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1e078e8e-1313-4eae-abd3-59f0b552184b · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners 24 Published as a conference paper at ICLR 2025 5
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29884104-94be-499e-8b19-d477871bdfb4 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Question to ChatGPT: 1 Cutting in the kitchen
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cee28ea6-30d4-404f-8dd0-d0beb9fd9a48 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Open-vocabulary Object Detection via Vision and Language Knowledge Distillation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf103ac1-24f3-4d49-a753-fbcf689133ba · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Ssan: Separable self-attention network for video representation learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bf8c7b70-2d00-44d5-886a-5f08865214a8 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Probing Image-Language Transformers for Verb Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c7360e-8155-4571-994f-905f8d27a155 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Frozen clip models are efficient video learners
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d4a05d5-0734-4cce-81a5-001d1e0bbe43 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Fine-tuned clip models are efficient video learners
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 17f4bbeb-d622-4727-9fd4-a82cd8d8c95d · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ba2c31b-bb45-4bbd-88b7-a88902eb8966 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Learning Video Representations using Contrastive Bidirectional Transformer
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 189d7143-2681-42f8-b109-8eeb513f617b · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners LLaMA: Open and Efficient Foundation Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c9f1f86-30b4-46ae-b74c-958f828cef35 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Clipasso: Semantically-aware object sketching
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1233edc1-6f33-470a-ad29-d29257a74c55 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Alternative semantic representations for zero-shot human action recog- nition
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48be1796-333d-469a-afc5-bb86bdcd62a5 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68564d2e-5f15-4512-b3a0-539c2841273a · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Multi-task zero-shot action recognition with prioritised data augmentation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 379eb698-d417-405d-9f65-11510a95ebc1 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners AIM: Adapting Image Models for Efficient Video Action Recognition
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2161c4e0-612f-4dcc-a496-766d5770fb8f · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Florence: A New Foundation Model for Computer Vision
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff8d952-8172-4892-8a76-e4246122f55e · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Co-training Transformer with Videos and Images Improves Action Recognition
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67f53610-4f3e-43d2-8685-a23343577c2d · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Learning a deep embedding model for zero-shot learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4ee53428-791e-44c5-b502-350d806c42d2 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners always be full rank
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 200f7ef2-5a07-4254-a8fd-6c20240743a9 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cd7a408c-c186-4d38-9883-ff73944cb9e5 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners For cases generated by Imagen, in Fig
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e18cf98f-99a8-4f0d-8d28-d574c381e5b1 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The evaluation is conducted three times
Reference 600
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 16a8fbbd-67f1-4c20-a0a6-08fdd3b4d172 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2dfe5d4-36b7-4d08-9b0b-344002d1ea88 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners UniFormer: Unified Transformer for Efficient Spatiotemporal Representation Learning
Reference 2011
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d69a116-17cb-4aaa-96e2-85e47d273369 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Reference 2012
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e430e668-dc8c-474e-b363-2c660d5577ae · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The Kinetics Human Action Video Dataset
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04503e34-8e11-4846-8e67-8a304247919c · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners ActionCLIP: A New Paradigm for Video Action Recognition
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d063b408-6141-46ae-866c-f2d5f89d9bb1 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners A Short Note about Kinetics-600
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04db9eee-fad7-499b-953e-b5d3a774339c · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Unresolved cited work
Reference 2018
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7282830b-f4e7-44d8-b555-e2eb1ac68c8b · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners All About Knowledge Graphs for Actions
Reference 2019
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18879d0e-89d6-4d49-8471-9459c318a785 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Ost: Refining text knowledge with optimal spatio-temporal descriptor for general video recognition
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d9cb6af-c5f0-49bf-8617-bc48dcad3c95 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners GPT-4 Technical Report
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34eb9f1f-4ddb-4be1-9c7f-4ddebeba97b8 · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Imagenet: A large-scale hi- erarchical image database
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63b0ca20-0bfd-4aea-9f86-eb9b99f9e19e · outbound
Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The Llama 3 Herd of Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.