Pith. sign in

Paper Citation Record · LEDGER

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification

As of 19 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.12585.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12585 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:51:56.884183Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved9
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b67e59df-72e3-4759-8e5a-2bb81e2eef98 · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.633975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.633975Z digest=sha256:0d21feea58b0f35d5e3f571c8530c9232166921995ec37cbb89f2872c313122a

Observation 8ab1f47e-37e1-492d-ba82-f1c039c19797 · outbound

This paper cites Timesformer-base-finetuned-ssv2.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Timesformer-base-finetuned-ssv2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.525211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.642321Z digest=sha256:cb67f49f5ea3ead136a9bae483d7d213d0b77c60fa88ac7ce38af7ee5d767674

Observation 658d9a26-9c3b-4d8d-aab3-e2817319650d · outbound

This paper cites Vivit: A video vision transformer.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit: A video vision transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.648106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.648106Z digest=sha256:e2e51148457c28fb59b93bbaee4ed1989169cfc7a68ed9cbed0192064a35e5fc

Observation 0e6c9a0d-4b2b-4063-b9ca-347ec6f49e00 · outbound

This paper cites The UEA multivariate time series classification archive, 2018.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The UEA multivariate time series classification archive, 2018

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.654307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.654307Z digest=sha256:dd6376f907d1e0f2b91002eb446569cafc83d5ae1d96d38f090a5009765d33ff

Observation 5bbfc546-cf26-442b-a620-80a9224747a0 · outbound

This paper cites Using dynamic time warping to find patterns in time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Using dynamic time warping to find patterns in time series

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.499531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.662898Z digest=sha256:b579876825a9ad620520b5ec3e328f43b9af1a0f65cb2c35aa618500f498bf88

Observation 1d43921f-e6c8-4bf0-80f9-0808086775e4 · outbound

This paper cites Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.484306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.669434Z digest=sha256:ad22f3b000583f7dd77f392751ea05914027c91ad088ed3da551e20456ffec93

Observation cf12174a-9605-4cc6-94ec-25092be3717a · outbound

This paper cites Dtwnet: a dynamic time warping network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dtwnet: a dynamic time warping network

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.468495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.675610Z digest=sha256:74eeb3c757cd6fb145009729c150670f02a47a390b0c47cd32d981a1a7044b1c

Observation b4090e4c-e43c-4145-bfb2-a4b89ad4c28f · outbound

This paper cites Few-shot video classification via tem- poral alignment.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Few-shot video classification via tem- poral alignment

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.453503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.683460Z digest=sha256:a948ee7f8f018b67ac72649e758592f26daab4eb98eea49846ca45b79cd51d67

Observation 09578d25-841b-4bdb-bc52-a50e17681f84 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Quo vadis, action recognition? a new model and the kinetics dataset

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.437285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.689715Z digest=sha256:60a2c146970bcb3d2b10157d2baa5532a10a5c294152690b976c3afbf9486ee9

Observation 315ad924-60b3-4429-8890-1adac7358998 · outbound

This paper cites D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.421567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.696692Z digest=sha256:fc714a73f633d905770a4462f188dbe87f74bd6a941be3993ea3f4e3f256df95

Observation 72fb712a-6387-4227-928a-1b4aa814a023 · outbound

This paper cites Soft-dtw: a differen- tiable loss function for time-series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Soft-dtw: a differen- tiable loss function for time-series

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.405107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.702587Z digest=sha256:af560cd2bd8e77066e8f62ac8689a85d019565816dc280dee1737bfcad02d23c

Observation 49b76c07-5636-4dec-89ea-cc453d97ef29 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.708585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.708585Z digest=sha256:258c54fc768e52524c72e8c6090af92d0539b569db73b98638315a329b4f0e26

Observation 459a09dc-828e-4b8a-b759-e4ff64da2bcb · outbound

This paper cites Slowfast networks for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Slowfast networks for video recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.388333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.714671Z digest=sha256:685fb5f278285de69a7f8746906b9c9a7eae7cb24cfb2bdcef41dfd745de386c

Observation 7f00da72-1183-46ff-a8af-06f48485d2fe · outbound

This paper cites Fine- grained temporal contrastive learning for weakly-supervised temporal action localization.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Fine- grained temporal contrastive learning for weakly-supervised temporal action localization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.372011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.720485Z digest=sha256:4fe43ea64b4d12907569bb043be950dac1d7ee1a9a66f779d3cb0985850e18c1

Observation bfe652ce-9907-4a39-815b-03c21ce5086d · outbound

This paper cites Video action transformer network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Video action transformer network

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.355791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.727998Z digest=sha256:19946c85a9cef7e2972cb3e89fc4b3c6a237a3f25a803064d6a6fcfe8bd4d59e

Observation 5b31ee90-1947-4b84-ad02-bb3848c92ffa · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The” something something” video database for learning and evaluating visual common sense

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.340875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.734194Z digest=sha256:1066a1ed6bdc69e3f2f1a61519c8ba523394690d8ed30f43487287d2599211a9

Observation d22e2b74-cc59-435e-b28f-a9c623386500 · outbound

This paper cites Large-scale video classification with convolutional neural networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Large-scale video classification with convolutional neural networks

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.325267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.740138Z digest=sha256:d0d7a3134b618939055820bf040f382b27e23bab99ac8110748af25b7f901913

Observation 9f767076-8961-46d2-b541-d5fefcb44829 · outbound

This paper cites The Kinetics Human Action Video Dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The Kinetics Human Action Video Dataset

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.746645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.746645Z digest=sha256:b546dbed5fe3c8d9890c978d109cec089b614a6e00e0401e2d98c48ceb23f13c

Observation 10b678ad-36a8-44d2-9eda-b472ec2b4779 · outbound

This paper cites Imagenet classification with deep convolutional neural net- works.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Imagenet classification with deep convolutional neural net- works

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.753566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.753566Z digest=sha256:8631a507433dd96532261ea5bb00aac2e4f5721d245a6da43f5bcbf4fe98c5bd

Observation aca566b0-fae4-45de-9ee6-402a446ed581 · outbound

This paper cites Hmdb: a large video database for human motion recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Hmdb: a large video database for human motion recognition

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.299267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.761485Z digest=sha256:7eeeb7bf1c06eedf655a39730632c44b8a9ee807911c1ec72fd4eee9e001880b

Observation b22a03c0-d157-4d0d-a430-b76faa09ebf0 · outbound

This paper cites Tam: Temporal adaptive module for video recog- nition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Tam: Temporal adaptive module for video recog- nition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.283560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.767773Z digest=sha256:7b27d79c5bdcd35c4e178431c808ec7f183eb9f2bd1166ef6d1b8b1ea9acc74a

Observation 7b9f8c0d-00dc-4b12-90ab-d2b9395b7587 · outbound

This paper cites Action recognition on something-something v2 leaderboard.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Action recognition on something-something v2 leaderboard

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.267961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.773604Z digest=sha256:573c43bcfd4d91a494d09b5cb9362124d707e0c4925983b6dd77c6605d43561d

Observation b2429d20-ce5a-4c9a-a0f7-4a670b37a1d8 · outbound

This paper cites A global averaging method for dynamic time warping, with ap- plications to clustering.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification A global averaging method for dynamic time warping, with ap- plications to clustering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.235229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.786601Z digest=sha256:2a021157323b072896d7327c141c4a0ae984037baa10700e9daff38c694882aa

Observation 7b01e21c-c369-4531-94f8-37af1feec1cf · outbound

This paper cites Re- thinking video vits: Sparse video tubes for joint image and video learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Re- thinking video vits: Sparse video tubes for joint image and video learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.220040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.793063Z digest=sha256:473ce108afd95eec07385890ee643c3544d14e0deee4a319de8127959919ddb5

Observation d893f334-bfab-4f78-a694-2d4300b22767 · outbound

This paper cites Vivit-b-16x2-kinetics400.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit-b-16x2-kinetics400

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.204661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.799549Z digest=sha256:f322b296564bd0e3d83c29acdb4d4eca8708bca5e4fe032f6a0a9926fcbfd245

Observation 1d9f18c2-6219-4efa-8109-9b06d1675cbd · outbound

This paper cites Dynamic time warping algorithm review.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dynamic time warping algorithm review

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.189675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.805951Z digest=sha256:9008c0d2ee7186404804d684deb4122dadfb84ec7c80c274ab98f8a42fff37ae

Observation d83b47c6-d347-44ad-bc5b-bf406b750fb4 · outbound

This paper cites The move-split-merge metric for time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The move-split-merge metric for time series

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.173908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.813505Z digest=sha256:1a256c59c690d5012145d9af208a77ee5d96eda8509c8ac8668b759325162dee

Observation ae3fccc3-ab4b-497c-936e-f4bdbf6f40cc · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.820397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.820397Z digest=sha256:545719e30011462e07aeba000bd75ebd0adeed1adcc13bc0b4a3b4565c28bceb

Observation 4ca14107-601d-42f1-8128-b6e559cacf50 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Learning spatiotemporal features with 3d convolutional networks

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.148604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.826935Z digest=sha256:c0ee8a3e9b6c8baffc975b7e42e136c412b5aba0bb10613234e691d7abee4cd5

Observation a39310a2-bc76-415a-9f7a-b977ed6debc5 · outbound

This paper cites Implicit temporal modeling with learn- able alignment for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Implicit temporal modeling with learn- able alignment for video recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.132257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.833801Z digest=sha256:4c45a2311931a465ea9bc1580df72131214e77a26e25c6665b377c68005a2080

Observation 9a368ee8-8e1b-4778-a037-9852857ffb9b · outbound

This paper cites Videomae v2: Scaling video masked autoencoders with dual masking.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae v2: Scaling video masked autoencoders with dual masking

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.116076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.839546Z digest=sha256:94a3b33061f25f0cd4838686a3d47e3f7273d6910f815463db1cdc8d3cfffa11

Observation 5838d60d-65ef-43e8-9612-0def7432d396 · outbound

This paper cites Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.099790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.845472Z digest=sha256:adb615953d8e333a08fa86032276a1063748dcb239322399adf7fade1d34feb4

Observation 30471e61-9c6b-4718-a4bb-197fbf6e0fd0 · outbound

This paper cites InternVideo2: Scaling Foundation Models for Multimodal Video Understanding.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.851307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.851307Z digest=sha256:51571445d71edae955f6cefad8baed05593953260cc626d65ae9c879075b3f4c

Observation cc167c16-f907-46f5-bfd2-fbe96b8520ef · outbound

This paper cites What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.083923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.857324Z digest=sha256:9db99e362461ada78f0886f76ed431ada46d1f6cf0f55850bb9a63791850325c

Observation f5639af5-dde6-4f43-9e8d-c87a9a7145d7 · outbound

This paper cites Multiview transformers for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Multiview transformers for video recognition

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.068465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.863125Z digest=sha256:6cdebcb06be53cbfb6d3d679d9e8dbf89550fafc14d92445d03a675f43ff1e56

Observation 253a879a-021b-469c-92db-e516a989b60c · outbound

This paper cites Scaling vision transformers.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Scaling vision transformers

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.053463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.872058Z digest=sha256:4e183cfd43210b031647148c6fc82e170e588b863683947b5afa30855cf2184e

Observation a61480f8-53da-4da3-a8c4-5ff66ee9bd86 · outbound

This paper cites We now describe our choice of the temporal sliding win- dow widths and strides.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification We now describe our choice of the temporal sliding win- dow widths and strides

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.038230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.877590Z digest=sha256:d1455388d223a52ec2aa2a633459838a59de2a5fa2e97b67702a2b3f9c97c7c6

Observation edd693eb-2acf-46f2-9be7-970f10af564f · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:51:57.019889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.884183Z digest=sha256:0ba9f8e8054a5cbb552f7bcdc732970598c26472f07c2a2b605218a7b2238a29

Observation 05b6330b-9b7a-4205-9ffd-8f9f28183805 · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T00:51:57.251878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T00:51:56.781033Z digest=sha256:4986eba514b279e456563fb2d7bfa9159e8eddcd158580b50234f448148671c0

Pith citing papers

No inbound Pith citation observations are available.