Pith. sign in

Paper Citation Record · LEDGER

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners

As of 9 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2502.03549.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.03549 v3

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:37:09.891480Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy15
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e39f4487-2931-45ab-926b-e28f2997a924 · outbound

This paper cites The low-rank bottleneck is vital, because it may lead to that many rows of the attention map are seriously homogenized.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The low-rank bottleneck is vital, because it may lead to that many rows of the attention map are seriously homogenized

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.576953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.867432Z digest=sha256:216250ef67442aead93dad3b4200f4b9ec6283e3ab25dd6d641e0258cd2f929e

Observation 30e31daa-6937-46b5-b0b1-25fcb9582777 · outbound

This paper cites Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, et al

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.744593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.602158Z digest=sha256:ed3bcb4c064914a2e245f5708b4e9f9692cfd92d6b967f55215cf19851cef457

Observation 1e078e8e-1313-4eae-abd3-59f0b552184b · outbound

This paper cites 24 Published as a conference paper at ICLR 2025 5.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners 24 Published as a conference paper at ICLR 2025 5

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.534172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.882418Z digest=sha256:0ba457faed467691c6cb121e1f1dbcebd8d9f3c89adea6ee64c914cf92a89004

Observation 29884104-94be-499e-8b19-d477871bdfb4 · outbound

This paper cites Question to ChatGPT: 1 Cutting in the kitchen.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Question to ChatGPT: 1 Cutting in the kitchen

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.385360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.886807Z digest=sha256:076107c47eff3d9be8fc6710d77a9642f7cd4019d0e8dbf376fed4f7695d9cfe

Observation cee28ea6-30d4-404f-8dd0-d0beb9fd9a48 · outbound

This paper cites Open-vocabulary Object Detection via Vision and Language Knowledge Distillation.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Open-vocabulary Object Detection via Vision and Language Knowledge Distillation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.633834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.633834Z digest=sha256:dd8774a5e67b94c59dc915d7316a2f5006ac00ac495388fe8fc8a54e30a65eca

Observation bf103ac1-24f3-4d49-a753-fbcf689133ba · outbound

This paper cites Ssan: Separable self-attention network for video representation learning.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Ssan: Separable self-attention network for video representation learning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.707354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.639269Z digest=sha256:b3ef54a913defa5b2c5588c9f5c59bc2274f5b2a5b49c1fa267415bd66e6927a

Observation bf8c7b70-2d00-44d5-886a-5f08865214a8 · outbound

This paper cites Probing Image-Language Transformers for Verb Understanding.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Probing Image-Language Transformers for Verb Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.644104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.644104Z digest=sha256:66d6cfd70f062029997a7d55da3ef020968e501e1adbf748c0a36cf2996791c5

Observation d8c7360e-8155-4571-994f-905f8d27a155 · outbound

This paper cites Frozen clip models are efficient video learners.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Frozen clip models are efficient video learners

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.693708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.659340Z digest=sha256:6c1a5c41cdca9deeaee28e368f0764bad1b33c94f51e8b769e29ee6b45702ec1

Observation 8d4a05d5-0734-4cce-81a5-001d1e0bbe43 · outbound

This paper cites Fine-tuned clip models are efficient video learners.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Fine-tuned clip models are efficient video learners

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.679530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.688091Z digest=sha256:9ebac7eb73af294eb5407027868edc7595049d43c37dea2cd71e3678a72cd713

Observation 17f4bbeb-d622-4727-9fd4-a82cd8d8c95d · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.706572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.706572Z digest=sha256:ab1c467f0daa7d91c9ba8049187dc12ac6c30ec675f8fecdfdee637be8acc75a

Observation 4ba2c31b-bb45-4bbd-88b7-a88902eb8966 · outbound

This paper cites Learning Video Representations using Contrastive Bidirectional Transformer.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Learning Video Representations using Contrastive Bidirectional Transformer

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.765177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.765177Z digest=sha256:449b8a0b108262370b263a1952c42d9885bcfc9910bb257d522bdbcd377d2161

Observation 189d7143-2681-42f8-b109-8eeb513f617b · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners LLaMA: Open and Efficient Foundation Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.784435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.784435Z digest=sha256:e633f69c4a59fc2acce463d25cb223f20f3d2c0051735fff8b03efcc974b9cb7

Observation 4c9f1f86-30b4-46ae-b74c-958f828cef35 · outbound

This paper cites Clipasso: Semantically-aware object sketching.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Clipasso: Semantically-aware object sketching

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.664290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.804493Z digest=sha256:1bacfcffe70da27036cb355863d4eb59805a376446feb19eec68d05b2ff24105

Observation 1233edc1-6f33-470a-ad29-d29257a74c55 · outbound

This paper cites Alternative semantic representations for zero-shot human action recog- nition.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Alternative semantic representations for zero-shot human action recog- nition

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.649678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.824240Z digest=sha256:f814f2238306a7dd5cc04e7b7c39ee70607f38ef6de804a1097910ba168cfb86

Observation 48be1796-333d-469a-afc5-bb86bdcd62a5 · outbound

This paper cites VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners VideoCLIP: Contrastive Pre-training for Zero-shot Video-Text Understanding

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.828306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.828306Z digest=sha256:22b67f45bfe88734cb8f1e5ceba54dd65764f0697b078621eec174c75f3eb928

Observation 68564d2e-5f15-4512-b3a0-539c2841273a · outbound

This paper cites Multi-task zero-shot action recognition with prioritised data augmentation.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Multi-task zero-shot action recognition with prioritised data augmentation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.634794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.832884Z digest=sha256:3dfdd1f178665349636fbf6e2fc914dd194908d55d36997be87f76ac524ee1f7

Observation 379eb698-d417-405d-9f65-11510a95ebc1 · outbound

This paper cites AIM: Adapting Image Models for Efficient Video Action Recognition.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners AIM: Adapting Image Models for Efficient Video Action Recognition

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.837333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.837333Z digest=sha256:1f92ae281774c6caa48bca08ba15debf7885d376367fde7a1f32f2cdb8223715

Observation 2161c4e0-612f-4dcc-a496-766d5770fb8f · outbound

This paper cites Florence: A New Foundation Model for Computer Vision.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Florence: A New Foundation Model for Computer Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.841758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.841758Z digest=sha256:bc6c17fc1dcb583e21a856844c21392e81dff4677fa1b860b6eaf64301b642e7

Observation 8ff8d952-8172-4892-8a76-e4246122f55e · outbound

This paper cites Co-training Transformer with Videos and Images Improves Action Recognition.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Co-training Transformer with Videos and Images Improves Action Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.846651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.846651Z digest=sha256:809d9caf0888be06fa8c2d6d8cfce067d34a06c3e75e4c63a8390d4b41213848

Observation 67f53610-4f3e-43d2-8685-a23343577c2d · outbound

This paper cites Learning a deep embedding model for zero-shot learning.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Learning a deep embedding model for zero-shot learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.619242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.851597Z digest=sha256:d84a7b3289b0b689ee5053cce2cce9108274eafd7167074712c8064c17d0434e

Observation 4ee53428-791e-44c5-b502-350d806c42d2 · outbound

This paper cites always be full rank.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners always be full rank

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.590705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.862379Z digest=sha256:8e21b2212229e6ecc397c1fdf4bb31480010cfab71ce6fa84bd3700b2a3911a9

Observation 200f7ef2-5a07-4254-a8fd-6c20240743a9 · outbound

This paper cites an unresolved cited work.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:37:10.561677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.872839Z digest=sha256:f909357de3b3e3b218da9bc388e6a0a5b89007a9739d0558f3c6bd60e5870dd1

Observation cd7a408c-c186-4d38-9883-ff73944cb9e5 · outbound

This paper cites For cases generated by Imagen, in Fig.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners For cases generated by Imagen, in Fig

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.242416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.891480Z digest=sha256:fa92dd7ec9a9075126abea031f6accdec108dd92046b3101d3a4150c7403f110

Observation e18cf98f-99a8-4f0d-8d28-d574c381e5b1 · outbound

This paper cites The evaluation is conducted three times.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The evaluation is conducted three times

Reference 600

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.548244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.877802Z digest=sha256:c863ba0b480a0154795a24654ef33bc83e59138c150b66296b81a8167e1b2302

Observation 16a8fbbd-67f1-4c20-a0a6-08fdd3b4d172 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.623418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.623418Z digest=sha256:15299ce8654d068efb88d59c52c0c65bf6386fbcd3197498a59aae3ea758ebbd

Observation f2dfe5d4-36b7-4d08-9b0b-344002d1ea88 · outbound

This paper cites UniFormer: Unified Transformer for Efficient Spatiotemporal Representation Learning.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners UniFormer: Unified Transformer for Efficient Spatiotemporal Representation Learning

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.654476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.654476Z digest=sha256:e5e735fcae27ee5db96a6c0cff4d73a19427121deddeb6fc355c4e9ae3f1ebde

Observation 5d69a116-17cb-4aaa-96e2-85e47d273369 · outbound

This paper cites VL-BERT: Pre-training of Generic Visual-Linguistic Representations.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners VL-BERT: Pre-training of Generic Visual-Linguistic Representations

Reference 2012

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.737902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.737902Z digest=sha256:f732dcc05cad2fef3c7e677b13f49a6a7203d58aa84d58171e478025bb8ca8be

Observation e430e668-dc8c-474e-b363-2c660d5577ae · outbound

This paper cites The Kinetics Human Action Video Dataset.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The Kinetics Human Action Video Dataset

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.649514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.649514Z digest=sha256:41d712ec827057898b1d51b660eed96db043d418524d0e6ae5365d6c6ffcc54c

Observation 04503e34-8e11-4846-8e67-8a304247919c · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners ActionCLIP: A New Paradigm for Video Action Recognition

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.819109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.819109Z digest=sha256:a7a9694c821ee86bc02907d40f612d87e68a1c1948434c558ac06e069df30e9b

Observation d063b408-6141-46ae-866c-f2d5f89d9bb1 · outbound

This paper cites A Short Note about Kinetics-600.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners A Short Note about Kinetics-600

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.606955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.606955Z digest=sha256:f3f9a84d318384739833a37ac6d3e1ef8a0bda9a628e8a36d8e06c1365b32b8f

Observation 04db9eee-fad7-499b-953e-b5d3a774339c · outbound

This paper cites an unresolved cited work.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Unresolved cited work

Reference 2018

Resolution
unresolved
raw_fallback, observed 2026-08-09T04:37:10.604170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.856868Z digest=sha256:22428af8fc606ebc11165cecaf5aa761c1c03714860de0f49afcc1f758f8b8d0

Observation 7282830b-f4e7-44d8-b555-e2eb1ac68c8b · outbound

This paper cites All About Knowledge Graphs for Actions.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners All About Knowledge Graphs for Actions

Reference 2019

Resolution
verified exact
local_arxiv, observed 2026-08-09T04:37:10.148462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.628515Z digest=sha256:26db26f0b6309920114dc2ccae868fafdde5ae0a169a47405ace28d36580ac9c

Observation 18879d0e-89d6-4d49-8471-9459c318a785 · outbound

This paper cites Ost: Refining text knowledge with optimal spatio-temporal descriptor for general video recognition.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Ost: Refining text knowledge with optimal spatio-temporal descriptor for general video recognition

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T04:37:10.730280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-09T04:37:09.612768Z digest=sha256:bf4469e9b6d3e3efc131c8cb91137d6441916cf7f54156f649b170ea69f56987

Observation 8d9cb6af-c5f0-49bf-8617-bc48dcad3c95 · outbound

This paper cites GPT-4 Technical Report.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners GPT-4 Technical Report

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.663562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.663562Z digest=sha256:3fb0649f732db0be09ccd7090497f67ed6cbc7c0dc26cfd075654dfc8e0a6193

Observation 34eb9f1f-4ddb-4be1-9c7f-4ddebeba97b8 · outbound

This paper cites Imagenet: A large-scale hi- erarchical image database.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners Imagenet: A large-scale hi- erarchical image database

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.618483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.618483Z digest=sha256:3e59e589ade9beb5fe0a7eec3f7ad88b4552411f823a775c84e389eeb0945f10

Observation 63b0ca20-0bfd-4aea-9f86-eb9b99f9e19e · outbound

This paper cites The Llama 3 Herd of Models.

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners The Llama 3 Herd of Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T04:37:09.596768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:37:09.596768Z digest=sha256:d3dd49cf2f4a76b0f2a51722188c5c682f4b349c0e8a989e35883940d31de214

Pith citing papers

No inbound Pith citation observations are available.