Pith. sign in

Paper Citation Record · LEDGER

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification

As of 11 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.12585.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12585 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:51:56.884183Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved9
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b67e59df-72e3-4759-8e5a-2bb81e2eef98 · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.633975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.633975Z digest=sha256:479b62b5d5bcca93dd360074b7a4b4a33d3d496227e87b030378ee3733f5f375

Observation 8ab1f47e-37e1-492d-ba82-f1c039c19797 · outbound

This paper cites Timesformer-base-finetuned-ssv2.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Timesformer-base-finetuned-ssv2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.525211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.642321Z digest=sha256:2ae537d261be8ad683da6e166386ebd7ee8ef90abe8cc09862a4df0ce44e85d8

Observation 658d9a26-9c3b-4d8d-aab3-e2817319650d · outbound

This paper cites Vivit: A video vision transformer.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit: A video vision transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.648106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.648106Z digest=sha256:f1363893d7454f488a4d55c9fa464082cd44572c19807a514aa22ea3ae3106c0

Observation 0e6c9a0d-4b2b-4063-b9ca-347ec6f49e00 · outbound

This paper cites The UEA multivariate time series classification archive, 2018.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The UEA multivariate time series classification archive, 2018

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.654307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.654307Z digest=sha256:946f8e937accdcd342d134ff6320fe39cb47010a47eeea7aca27e8e3392bc394

Observation 5bbfc546-cf26-442b-a620-80a9224747a0 · outbound

This paper cites Using dynamic time warping to find patterns in time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Using dynamic time warping to find patterns in time series

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.499531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.662898Z digest=sha256:873c1c4c9a49f49b3c7f70a5fbe4b7b2ba4953dc9ee151071e1e62f79ff5b614

Observation 1d43921f-e6c8-4bf0-80f9-0808086775e4 · outbound

This paper cites Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.484306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.669434Z digest=sha256:140b5372b092741f4f1e1d011980020ea9f635eb21039ddf98a3af2521b9deb7

Observation cf12174a-9605-4cc6-94ec-25092be3717a · outbound

This paper cites Dtwnet: a dynamic time warping network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dtwnet: a dynamic time warping network

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.468495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.675610Z digest=sha256:2e32b7409f485d3344b11e8e907cd825eff7104fc49d2395941a60f453299b4c

Observation b4090e4c-e43c-4145-bfb2-a4b89ad4c28f · outbound

This paper cites Few-shot video classification via tem- poral alignment.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Few-shot video classification via tem- poral alignment

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.453503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.683460Z digest=sha256:1af00c32d3c8fe31cd177f010a27544507d2434b00b77003dcd8f7889176d7f9

Observation 09578d25-841b-4bdb-bc52-a50e17681f84 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Quo vadis, action recognition? a new model and the kinetics dataset

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.437285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.689715Z digest=sha256:e8999fae02ade98bb32630c06e4b1a38d7017878489f420c959a0d94bf224404

Observation 315ad924-60b3-4429-8890-1adac7358998 · outbound

This paper cites D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.421567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.696692Z digest=sha256:ec7969a8e36832730bd851ff0d5a7c2eb29a204e5786e91802d4a70bec918bc1

Observation 72fb712a-6387-4227-928a-1b4aa814a023 · outbound

This paper cites Soft-dtw: a differen- tiable loss function for time-series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Soft-dtw: a differen- tiable loss function for time-series

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.405107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.702587Z digest=sha256:8613b718c8673468cde855df589c67a115149d0f6653e6e236447b76dd3a25fb

Observation 49b76c07-5636-4dec-89ea-cc453d97ef29 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.708585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.708585Z digest=sha256:f73847f38923bdc2b68d7cb11d9b47d76498bc2338ae967e1582844e2b0ab1b0

Observation 459a09dc-828e-4b8a-b759-e4ff64da2bcb · outbound

This paper cites Slowfast networks for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Slowfast networks for video recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.388333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.714671Z digest=sha256:65958389a30d9d77d692a221ce1b86103d0dfda24cb3b817843cf05404ae8d43

Observation 7f00da72-1183-46ff-a8af-06f48485d2fe · outbound

This paper cites Fine- grained temporal contrastive learning for weakly-supervised temporal action localization.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Fine- grained temporal contrastive learning for weakly-supervised temporal action localization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.372011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.720485Z digest=sha256:6b8effa493a063bb9b2fe9ef443b1b8105a70fea0e94aa030803938b0f3959d3

Observation bfe652ce-9907-4a39-815b-03c21ce5086d · outbound

This paper cites Video action transformer network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Video action transformer network

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.355791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.727998Z digest=sha256:5354375c1d88c36a4cff6cceb201b7a053b227486f14eed65cb2e925069ec377

Observation 5b31ee90-1947-4b84-ad02-bb3848c92ffa · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The” something something” video database for learning and evaluating visual common sense

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.340875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.734194Z digest=sha256:cc05c433785ba3b55c34292a956c9c2ebef83ed9b83f7d4c3f765b08c9ff45e8

Observation d22e2b74-cc59-435e-b28f-a9c623386500 · outbound

This paper cites Large-scale video classification with convolutional neural networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Large-scale video classification with convolutional neural networks

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.325267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.740138Z digest=sha256:6be04b7a47be82e87501d547b18d1cb745756b865b4b964295e9beae4aa13d13

Observation 9f767076-8961-46d2-b541-d5fefcb44829 · outbound

This paper cites The Kinetics Human Action Video Dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The Kinetics Human Action Video Dataset

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.746645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.746645Z digest=sha256:2c11a2487532cc19d81b937a472bba3a0ff05ee2d44202acd393af5e03f3ec0c

Observation 10b678ad-36a8-44d2-9eda-b472ec2b4779 · outbound

This paper cites Imagenet classification with deep convolutional neural net- works.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Imagenet classification with deep convolutional neural net- works

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.753566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.753566Z digest=sha256:828aeed823b690f352e95a657b2c750f3703f88cbd2afcdb7e31145a525341e9

Observation aca566b0-fae4-45de-9ee6-402a446ed581 · outbound

This paper cites Hmdb: a large video database for human motion recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Hmdb: a large video database for human motion recognition

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.299267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.761485Z digest=sha256:5821a18d3c9cb3faaa051cce4704f08e86f89c6bbb6550612d880287f37e3df5

Observation b22a03c0-d157-4d0d-a430-b76faa09ebf0 · outbound

This paper cites Tam: Temporal adaptive module for video recog- nition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Tam: Temporal adaptive module for video recog- nition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.283560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.767773Z digest=sha256:33ee50eeba9b71ca6614706adc6763ad2c11fe189517aede9b0576402c57999e

Observation 7b9f8c0d-00dc-4b12-90ab-d2b9395b7587 · outbound

This paper cites Action recognition on something-something v2 leaderboard.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Action recognition on something-something v2 leaderboard

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.267961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.773604Z digest=sha256:3b61a32473014533e249831777787da2f98232ccac96eb0dc26778bcaa8a531c

Observation b2429d20-ce5a-4c9a-a0f7-4a670b37a1d8 · outbound

This paper cites A global averaging method for dynamic time warping, with ap- plications to clustering.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification A global averaging method for dynamic time warping, with ap- plications to clustering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.235229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.786601Z digest=sha256:d8f03a8118f36aed612364cc632bcc944fa62ea9ad12296149069570fc97968f

Observation 7b01e21c-c369-4531-94f8-37af1feec1cf · outbound

This paper cites Re- thinking video vits: Sparse video tubes for joint image and video learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Re- thinking video vits: Sparse video tubes for joint image and video learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.220040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.793063Z digest=sha256:d6576007ffb5de490298718a9fbcc5c7d86b6c5af417e599b379d77af88bd7df

Observation d893f334-bfab-4f78-a694-2d4300b22767 · outbound

This paper cites Vivit-b-16x2-kinetics400.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit-b-16x2-kinetics400

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.204661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.799549Z digest=sha256:c25a1d7392319fad4372bb20cec53abbe1d2ca88c80511d139db4a132d097e20

Observation 1d9f18c2-6219-4efa-8109-9b06d1675cbd · outbound

This paper cites Dynamic time warping algorithm review.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dynamic time warping algorithm review

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.189675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.805951Z digest=sha256:068e646dceb697842e706b95842bf257068e99a1ffe47757e9db356177f38e32

Observation d83b47c6-d347-44ad-bc5b-bf406b750fb4 · outbound

This paper cites The move-split-merge metric for time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The move-split-merge metric for time series

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.173908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.813505Z digest=sha256:11e373dbbb2c163265b199c115f51423f30d13e4f2384494c3655214c9c12001

Observation ae3fccc3-ab4b-497c-936e-f4bdbf6f40cc · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.820397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.820397Z digest=sha256:af622415748056760112156c64021f2687c277ad967b0e9f2229c19cd4e237f0

Observation 4ca14107-601d-42f1-8128-b6e559cacf50 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Learning spatiotemporal features with 3d convolutional networks

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.148604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.826935Z digest=sha256:d36ef4312f95a798d81e1e6b9d4d5cacb85f97afc83fbfa1078b1239221aff23

Observation a39310a2-bc76-415a-9f7a-b977ed6debc5 · outbound

This paper cites Implicit temporal modeling with learn- able alignment for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Implicit temporal modeling with learn- able alignment for video recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.132257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.833801Z digest=sha256:ca6989fce827b960e3e0676b2bb1690c3b86a949d58e51ec13b30aa85947dc73

Observation 9a368ee8-8e1b-4778-a037-9852857ffb9b · outbound

This paper cites Videomae v2: Scaling video masked autoencoders with dual masking.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae v2: Scaling video masked autoencoders with dual masking

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.116076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.839546Z digest=sha256:f9cc283f2db141a2f224a64d0eafcbd1e2929c063aeb8b149df609f487dcc79b

Observation 5838d60d-65ef-43e8-9612-0def7432d396 · outbound

This paper cites Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.099790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.845472Z digest=sha256:704e7ba20647e74c789214b866b2038729b62910ba5a8a840a0aab3ed829f423

Observation 30471e61-9c6b-4718-a4bb-197fbf6e0fd0 · outbound

This paper cites InternVideo2: Scaling Foundation Models for Multimodal Video Understanding.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.851307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.851307Z digest=sha256:312e1d7bb7962778b503b74cdb33d9deada8c98929a0b3d99bbf0c72d0dd9b0b

Observation cc167c16-f907-46f5-bfd2-fbe96b8520ef · outbound

This paper cites What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.083923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.857324Z digest=sha256:4d757aabbe72c5aa63b444231f5b2b3b7855f8533a5b5406b8e9050d7aa68a8b

Observation f5639af5-dde6-4f43-9e8d-c87a9a7145d7 · outbound

This paper cites Multiview transformers for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Multiview transformers for video recognition

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.068465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.863125Z digest=sha256:2073d0cbe450b2d6812ca3c1cbf069664c3d627a3355be9dd998daf6b4088f8a

Observation 253a879a-021b-469c-92db-e516a989b60c · outbound

This paper cites Scaling vision transformers.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Scaling vision transformers

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.053463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.872058Z digest=sha256:807c231f672bc7a472b0e5ab2a5863c852a34636233a132cb9240c689afdd239

Observation a61480f8-53da-4da3-a8c4-5ff66ee9bd86 · outbound

This paper cites We now describe our choice of the temporal sliding win- dow widths and strides.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification We now describe our choice of the temporal sliding win- dow widths and strides

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.038230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.877590Z digest=sha256:e567b58eeb2344025eb5abb17009a67a8aa876c971119163d99ae067ee10589c

Observation edd693eb-2acf-46f2-9be7-970f10af564f · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:51:57.019889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.884183Z digest=sha256:bb00b0db44f47a931d2c99126bd06158ebddf32f55bb10411ee149f459c5cb91

Observation 05b6330b-9b7a-4205-9ffd-8f9f28183805 · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T00:51:57.251878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-07T00:51:56.781033Z digest=sha256:56131c6204bb3c897b1016bcbf414cb520d4ae68247db40bd904ec8e4228f950

Pith citing papers

No inbound Pith citation observations are available.