Pith. sign in

Paper Citation Record · LEDGER

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.12585.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.12585 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:51:56.884183Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved9
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b67e59df-72e3-4759-8e5a-2bb81e2eef98 · outbound

This paper cites YouTube-8M: A Large-Scale Video Classification Benchmark.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification YouTube-8M: A Large-Scale Video Classification Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.633975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.633975Z digest=sha256:479b62b5d5bcca93dd360074b7a4b4a33d3d496227e87b030378ee3733f5f375

Observation 8ab1f47e-37e1-492d-ba82-f1c039c19797 · outbound

This paper cites Timesformer-base-finetuned-ssv2.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Timesformer-base-finetuned-ssv2

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.525211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.642321Z digest=sha256:b6165dfa94b0f0302c109ee71d8ea95125a54e40d9065aed04e1cdaf2d8eba68

Observation 658d9a26-9c3b-4d8d-aab3-e2817319650d · outbound

This paper cites Vivit: A video vision transformer.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit: A video vision transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.648106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.648106Z digest=sha256:f1363893d7454f488a4d55c9fa464082cd44572c19807a514aa22ea3ae3106c0

Observation 0e6c9a0d-4b2b-4063-b9ca-347ec6f49e00 · outbound

This paper cites The UEA multivariate time series classification archive, 2018.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The UEA multivariate time series classification archive, 2018

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.654307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.654307Z digest=sha256:d3b5ab4cb41703b8b1c1fcaa58b1c53356bef1b06a9664fc2b07efbdacd5173e

Observation 5bbfc546-cf26-442b-a620-80a9224747a0 · outbound

This paper cites Using dynamic time warping to find patterns in time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Using dynamic time warping to find patterns in time series

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.499531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.662898Z digest=sha256:722ed16dce1da8bdadd3ecaaca38d97b18febae1eb37f9ba87ab41e9be8832e8

Observation 1d43921f-e6c8-4bf0-80f9-0808086775e4 · outbound

This paper cites Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Is space-time attention all you need for video understanding? In International Conference on Machine Learning , pages 813–824

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.484306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.669434Z digest=sha256:2de7398389090889a0b6a665647a8f38c706cd45777847d7ee69f2949968691c

Observation cf12174a-9605-4cc6-94ec-25092be3717a · outbound

This paper cites Dtwnet: a dynamic time warping network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dtwnet: a dynamic time warping network

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.468495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.675610Z digest=sha256:c243003071125109deeb80d005fb2eaa6203af7aafb295dbbeec76a7cf2b7067

Observation b4090e4c-e43c-4145-bfb2-a4b89ad4c28f · outbound

This paper cites Few-shot video classification via tem- poral alignment.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Few-shot video classification via tem- poral alignment

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.453503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.683460Z digest=sha256:650d1e5c8528cb0dc802559b8b5132adf6d719ca0453b50f63774906c68541c9

Observation 09578d25-841b-4bdb-bc52-a50e17681f84 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Quo vadis, action recognition? a new model and the kinetics dataset

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.437285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.689715Z digest=sha256:1a05ee4e671a7d82eb8bb40580b73b6629edd3364cd015f692eb7730559013dc

Observation 315ad924-60b3-4429-8890-1adac7358998 · outbound

This paper cites D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification D3tw: Discriminative differentiable dy- namic time warping for weakly supervised action alignment and segmentation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.421567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.696692Z digest=sha256:b83766de7b365d68378cc94d12eb2e1667fb635e112c8b90d02e80b30f994c5b

Observation 72fb712a-6387-4227-928a-1b4aa814a023 · outbound

This paper cites Soft-dtw: a differen- tiable loss function for time-series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Soft-dtw: a differen- tiable loss function for time-series

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.405107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.702587Z digest=sha256:cf5697453b901dc2c96a1aad5c14e3a8ba8994bf8ab577cea237190f048a1cee

Observation 49b76c07-5636-4dec-89ea-cc453d97ef29 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.708585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.708585Z digest=sha256:3973349a55473aa54b0591892ed3573ef849e0d34b871734492a24478a039a3f

Observation 459a09dc-828e-4b8a-b759-e4ff64da2bcb · outbound

This paper cites Slowfast networks for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Slowfast networks for video recognition

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.388333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.714671Z digest=sha256:27b6318247629c8ef7fb7a6f54f1b8464676147bf9589623ac82fee2263384b4

Observation 7f00da72-1183-46ff-a8af-06f48485d2fe · outbound

This paper cites Fine- grained temporal contrastive learning for weakly-supervised temporal action localization.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Fine- grained temporal contrastive learning for weakly-supervised temporal action localization

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.372011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.720485Z digest=sha256:09a28ee0e7e18b5124ea79cdfbe6d4ac076b47aa6620d6b7ff1c4ef83582bf60

Observation bfe652ce-9907-4a39-815b-03c21ce5086d · outbound

This paper cites Video action transformer network.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Video action transformer network

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.355791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.727998Z digest=sha256:a2e0c5f73fc2abf4574f190b75213fef33eb25d8ced43a7c3e170e2821acefeb

Observation 5b31ee90-1947-4b84-ad02-bb3848c92ffa · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The” something something” video database for learning and evaluating visual common sense

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.340875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.734194Z digest=sha256:6bfa197d50c307927738114b61836798a55e261f2626b21b059e937dcbd12b04

Observation d22e2b74-cc59-435e-b28f-a9c623386500 · outbound

This paper cites Large-scale video classification with convolutional neural networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Large-scale video classification with convolutional neural networks

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.325267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.740138Z digest=sha256:4c8bf471038fd869cb5f403b041569f87c637331803ec15e7b2601cf81884d86

Observation 9f767076-8961-46d2-b541-d5fefcb44829 · outbound

This paper cites The Kinetics Human Action Video Dataset.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The Kinetics Human Action Video Dataset

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.746645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.746645Z digest=sha256:4d50a8fab947a757205ce287e3fac36db7a96c616984755087c48856f14bb2a3

Observation 10b678ad-36a8-44d2-9eda-b472ec2b4779 · outbound

This paper cites Imagenet classification with deep convolutional neural net- works.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Imagenet classification with deep convolutional neural net- works

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.753566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.753566Z digest=sha256:828aeed823b690f352e95a657b2c750f3703f88cbd2afcdb7e31145a525341e9

Observation aca566b0-fae4-45de-9ee6-402a446ed581 · outbound

This paper cites Hmdb: a large video database for human motion recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Hmdb: a large video database for human motion recognition

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.299267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.761485Z digest=sha256:d05212ea46564a8d9ab074cc0ff49c2d45d24d866d18ca30cbd4f3d5e982b26e

Observation b22a03c0-d157-4d0d-a430-b76faa09ebf0 · outbound

This paper cites Tam: Temporal adaptive module for video recog- nition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Tam: Temporal adaptive module for video recog- nition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.283560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.767773Z digest=sha256:ecc845a6a7f5919cbca0fbd52a9782c72df2c0c81940acf840e06f391b2cf33e

Observation 7b9f8c0d-00dc-4b12-90ab-d2b9395b7587 · outbound

This paper cites Action recognition on something-something v2 leaderboard.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Action recognition on something-something v2 leaderboard

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.267961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.773604Z digest=sha256:acf7392ff860696186842d5e170f1c3b4953c599999f87f9f5974d86234c04b9

Observation b2429d20-ce5a-4c9a-a0f7-4a670b37a1d8 · outbound

This paper cites A global averaging method for dynamic time warping, with ap- plications to clustering.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification A global averaging method for dynamic time warping, with ap- plications to clustering

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.235229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.786601Z digest=sha256:b37f481965e5e3b89dde95307c212a5cfa69a124c384a47daec58fde12f014be

Observation 7b01e21c-c369-4531-94f8-37af1feec1cf · outbound

This paper cites Re- thinking video vits: Sparse video tubes for joint image and video learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Re- thinking video vits: Sparse video tubes for joint image and video learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.220040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.793063Z digest=sha256:2d88cc3c2305032d704dce940d8c43104904076be7f8e406aa890e83f26a6d80

Observation d893f334-bfab-4f78-a694-2d4300b22767 · outbound

This paper cites Vivit-b-16x2-kinetics400.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Vivit-b-16x2-kinetics400

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.204661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.799549Z digest=sha256:5facf5139f4cfe41aa0ec11544d73df30834d6b0e3b4b1d720a8c5bec40d5268

Observation 1d9f18c2-6219-4efa-8109-9b06d1675cbd · outbound

This paper cites Dynamic time warping algorithm review.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Dynamic time warping algorithm review

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.189675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.805951Z digest=sha256:398a8a7c3c912047495a12119ee4078579b1ff634a071bca037772a1c7b1e7c7

Observation d83b47c6-d347-44ad-bc5b-bf406b750fb4 · outbound

This paper cites The move-split-merge metric for time series.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification The move-split-merge metric for time series

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.173908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.813505Z digest=sha256:0f47645be959b949d93489561b3327aac34184848d83919ecb011e7fd2924787

Observation ae3fccc3-ab4b-497c-936e-f4bdbf6f40cc · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.820397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.820397Z digest=sha256:af622415748056760112156c64021f2687c277ad967b0e9f2229c19cd4e237f0

Observation 4ca14107-601d-42f1-8128-b6e559cacf50 · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Learning spatiotemporal features with 3d convolutional networks

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.148604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.826935Z digest=sha256:81721a989397f30a16400e0e46e855daaf9389b9bc5fbf7b1da28469ff46a5b2

Observation a39310a2-bc76-415a-9f7a-b977ed6debc5 · outbound

This paper cites Implicit temporal modeling with learn- able alignment for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Implicit temporal modeling with learn- able alignment for video recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.132257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.833801Z digest=sha256:1806b373beec9fdf094e5f18d9700187cda5fba20b184227fdb23095495a3c24

Observation 9a368ee8-8e1b-4778-a037-9852857ffb9b · outbound

This paper cites Videomae v2: Scaling video masked autoencoders with dual masking.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Videomae v2: Scaling video masked autoencoders with dual masking

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.116076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.839546Z digest=sha256:493e8138aaea3826c3194c6058aa4d027ee4be5687c6eb31651d72b84b317df0

Observation 5838d60d-65ef-43e8-9612-0def7432d396 · outbound

This paper cites Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Masked video distillation: Rethinking masked feature mod- eling for self-supervised video representation learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.099790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.845472Z digest=sha256:38925f34b643ea33bf8778548bccecb147dfd0f54c0bf266ae207f8e9db8e7c9

Observation 30471e61-9c6b-4718-a4bb-197fbf6e0fd0 · outbound

This paper cites InternVideo2: Scaling Foundation Models for Multimodal Video Understanding.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification InternVideo2: Scaling Foundation Models for Multimodal Video Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:56.851307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:56.851307Z digest=sha256:67e30f9c6fd56ebce5b2072fb08d1c4b1e8075db34e8b18b14e4f4e4c05c1d98

Observation cc167c16-f907-46f5-bfd2-fbe96b8520ef · outbound

This paper cites What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification What can simple arithmetic oper- ations do for temporal modeling? In Proceedings of the IEEE/CVF International Conference on Computer Vision , pages 13712–13722, 2023

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.083923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.857324Z digest=sha256:d9f86968807b5be612e933ebf6a3a784c2a7e95c944d825210ca948d1afb0743

Observation f5639af5-dde6-4f43-9e8d-c87a9a7145d7 · outbound

This paper cites Multiview transformers for video recognition.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Multiview transformers for video recognition

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.068465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.863125Z digest=sha256:32479b9aefe1720c9afb84706b18e23da50cab72c1aad883655264bb017f31b2

Observation 253a879a-021b-469c-92db-e516a989b60c · outbound

This paper cites Scaling vision transformers.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Scaling vision transformers

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.053463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.872058Z digest=sha256:c4c2d69c513e807adab69ef027ffc67cb4dcad36d0868dcd6de5fe92043ddf10

Observation a61480f8-53da-4da3-a8c4-5ff66ee9bd86 · outbound

This paper cites We now describe our choice of the temporal sliding win- dow widths and strides.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification We now describe our choice of the temporal sliding win- dow widths and strides

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:51:57.038230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.877590Z digest=sha256:1a213d80589a019d867292827e2f614b474e7d34420d7a1ad54adad9fa7a3e18

Observation edd693eb-2acf-46f2-9be7-970f10af564f · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:51:57.019889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.884183Z digest=sha256:890c6f04e6fb356cb1ae1a8b2d69d59b40ce6beb4e549302c2968122f0e0c8f4

Observation 05b6330b-9b7a-4205-9ffd-8f9f28183805 · outbound

This paper cites an unresolved cited work.

DejaVid: Encoder-Agnostic Learned Temporal Matching for Video Classification Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-07T00:51:57.251878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:51:56.781033Z digest=sha256:39719acb167fcbac6191d6ee46c507b27d71de37d3431aa8522e837f5dc0ba75

Pith citing papers

No inbound Pith citation observations are available.