Pith. sign in

Paper Citation Record · LEDGER

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition

As of 18 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2411.11335.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.11335 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T18:42:02.699274Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact2
  • verified fuzzy36
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ba030868-4217-49fc-98eb-2c586130c6c7 · outbound

This paper cites Tsm: Temporal shift module for efficient video understanding,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Tsm: Temporal shift module for efficient video understanding,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.424833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.424833Z digest=sha256:0d69b40fb6d40dc50bc2d1368b234042911f2ae7891ecfafc7ab6913bfe83050

Observation 74237eda-b1ca-412d-8e46-d65c3899878c · outbound

This paper cites Learning match kernels on grassmann manifolds for action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Learning match kernels on grassmann manifolds for action recognition,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.768926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.430352Z digest=sha256:46a931dfc3e2b86a450b442ded09050822a28116ebd556b20faa74e6a5a684a2

Observation 028a41fe-2cec-452e-8dec-f15c8881109e · outbound

This paper cites Motion-driven visual tempo learning for video-based action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Motion-driven visual tempo learning for video-based action recognition,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.750692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.435153Z digest=sha256:0cd75f21a603ebacec9d665a0e0eb19b600363d0fd66030c1d265b73fa83dac7

Observation aa604a25-62a9-45d5-9428-68e511aca2c3 · outbound

This paper cites Two-stream convolutional networks for action recognition in videos,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Two-stream convolutional networks for action recognition in videos,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.732839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.440193Z digest=sha256:9aadf9006fcea66b2c0a36be89d1a338f47f06a80cb06bfcf7944eeb9b82f7c3

Observation 16756be5-86fc-4f41-87c7-17109b985630 · outbound

This paper cites Real-time action recognition with deeply transferred motion vector cnns,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Real-time action recognition with deeply transferred motion vector cnns,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.717400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.445192Z digest=sha256:24580103af301c4edb664dc0cea00c4fc3a3593d881d8c2df657e293cca93e12

Observation a8a03092-dd5a-4f4f-b134-36db9925ef49 · outbound

This paper cites Model-agnostic meta-learning for fast adaptation of deep networks,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Model-agnostic meta-learning for fast adaptation of deep networks,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.450119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.450119Z digest=sha256:e5029e0fde5af7e31d600392db97c047a3121ce4d805f2250d09f5a26bc8d50f

Observation 0d73a641-628c-4667-8f26-4ebe17279982 · outbound

This paper cites Learning to compare relation: Semantic align- ment for few-shot learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Learning to compare relation: Semantic align- ment for few-shot learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.690918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.455752Z digest=sha256:f7fdd3890b6d11c6f3675e3b89710e71e5d764b7c9dc69e1b7c3c6d6d8a8aa5a

Observation eeccfa03-5dd3-4add-8d48-42c08c5c5e7b · outbound

This paper cites Hierarchical prototype refine- ment with progressive inter-categorical discrimination maximization for few-shot learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Hierarchical prototype refine- ment with progressive inter-categorical discrimination maximization for few-shot learning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.674636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.461091Z digest=sha256:2877909b019de1f7992c6e2f64d156a5aa94dc3a419e3c82f68d64ff273d1c78

Observation d6916438-1401-4524-95b0-302a5f9b3bac · outbound

This paper cites A two-stage approach to few-shot learning for image recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition A two-stage approach to few-shot learning for image recognition,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.657258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.466606Z digest=sha256:371ac308a8a28a93b81bebdf93ff4094b755bdbefbb558bec3db0eb23e4d621c

Observation 843ead99-96c4-4988-8580-2ee87c025186 · outbound

This paper cites Scformer: Spectral coordinate transformer for cross-domain few-shot hyperspectral image classification,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Scformer: Spectral coordinate transformer for cross-domain few-shot hyperspectral image classification,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.639950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.471212Z digest=sha256:5accfe7cb396bed3c13bf4e4518edbba699877aa96afd89c24549ceb0257e793

Observation dc8009e0-6e57-4fe8-8f3b-273bbfb388e8 · outbound

This paper cites Cross-modal contrastive learning network for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Cross-modal contrastive learning network for few-shot action recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.622104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.475955Z digest=sha256:7cf45a29108aa9561774af60e425b4aa38aba3a9075de6183fcf59a9e270d6f6

Observation fc87d035-6389-4a4d-be34-dd314d5be824 · outbound

This paper cites Few-shot video classification via temporal alignment,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Few-shot video classification via temporal alignment,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.605282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.480588Z digest=sha256:d142ab821e7e404334e847c2f3ce18b53674f7eaa40b5593dff0072e2e76a7e4

Observation ce9a3afd-107b-4d89-8c79-c5f99c7d8b61 · outbound

This paper cites Spatio-temporal relation modeling for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Spatio-temporal relation modeling for few-shot action recognition,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.587842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.485193Z digest=sha256:766a56beb123e110a596402d7c7f20e3f5c3c8450f9c5c08e4b16558cee47ade

Observation 934707c3-4178-4fb8-817a-f30c7f097626 · outbound

This paper cites Few-shot action recognition with permutation-invariant attention,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Few-shot action recognition with permutation-invariant attention,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.571351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.489649Z digest=sha256:aa22ec23afb619037ffa8de82712681dd5919afac27d8b35440555bb0ff6964f

Observation a6c66b2a-5ab4-4c5c-9c91-17d0079d65fa · outbound

This paper cites Compound memory networks for few-shot video classification,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Compound memory networks for few-shot video classification,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.553851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.494972Z digest=sha256:a908d8056abd74c82df8024a306acc7441c81c292df880c8b7739f55d96b8c57

Observation ccd0c47d-176d-4847-b139-aa9b17d68f93 · outbound

This paper cites TARN: Temporal Attentive Relation Network for Few-Shot and Zero-Shot Action Recognition.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition TARN: Temporal Attentive Relation Network for Few-Shot and Zero-Shot Action Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:42:03.040047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.502370Z digest=sha256:33503666228a02a76883eafa799c39e5181638814612358ff6f1df51bc42101a

Observation 20dc7951-52c5-4cb1-81d4-a27a79ef3502 · outbound

This paper cites Ta2n: Two-stage action alignment network for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Ta2n: Two-stage action alignment network for few-shot action recognition,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.535974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.508381Z digest=sha256:30802fd0d8c1eee3cd176879aae8c50d5bb703b4a0e0ef6e996b85aaa9a1dfe0

Observation 0f553773-d8ed-4a5f-a054-7f3d73e602fd · outbound

This paper cites Molo: Motion-augmented long-short contrastive learning for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Molo: Motion-augmented long-short contrastive learning for few-shot action recognition,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.517381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.513802Z digest=sha256:89e7997e73e41ae6663de1dff5dd9788d40b105d0f1e3d79a3e337196c99b145

Observation 65ca4e40-c34a-4d82-8c0a-fe485f66f7f6 · outbound

This paper cites Task discrepancy maximization for fine-grained few-shot classification,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Task discrepancy maximization for fine-grained few-shot classification,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.498572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.519602Z digest=sha256:083ad793d7c58c118ec6694a87f33129339bf35e7a189e11226ca01dd8f783b4

Observation bec247e6-6955-4fea-8fcd-15f42cb5ba55 · outbound

This paper cites Bi-directional feature reconstruction network for fine-grained few-shot image classification,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Bi-directional feature reconstruction network for fine-grained few-shot image classification,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.482110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.524929Z digest=sha256:3ec3e1529db9e9eab5379015abca505ad8a065debc56c1216e41457069bc0ba4

Observation 5a6bac51-61c0-4ac3-9a0c-47db9a93160a · outbound

This paper cites Motion-modulated temporal fragment alignment network for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Motion-modulated temporal fragment alignment network for few-shot action recognition,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.464105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.530753Z digest=sha256:bebc13dc93416ebc3807463b7cd366f6341a5ea3622cdd9f39bc6ad359bd0c20

Observation 7abd7be6-dd60-4f72-8697-3c3b23a53538 · outbound

This paper cites Hybrid relation guided set matching for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Hybrid relation guided set matching for few-shot action recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.445823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.536096Z digest=sha256:5b34ca37d1231454168a88a1a4db0a06726946a46eaab417a7120eb9a97cd12c

Observation 208ce478-b8f9-436e-a865-71a0771b85d4 · outbound

This paper cites Seeing inferences: brain dynamics and oculomotor signatures of non-verbal deduction,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Seeing inferences: brain dynamics and oculomotor signatures of non-verbal deduction,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.427668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.542347Z digest=sha256:04a4db94b022eec639752a6287ab9cf3b05135db87f7e2d9e01f5fe42ecc3c79

Observation c3426b46-dda2-4bfa-89af-4c654d6ff3ef · outbound

This paper cites St-adapter: Parameter- efficient image-to-video transfer learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition St-adapter: Parameter- efficient image-to-video transfer learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.547874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.547874Z digest=sha256:6a07fbb9c2db3851c5eb34d8dcf34755feb80de2ab8effbdb02c2233145ba525

Observation dab1d83e-66f5-4004-959f-ba6c0f701b63 · outbound

This paper cites Temporal-relational crosstransformers for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Temporal-relational crosstransformers for few-shot action recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.399812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.553437Z digest=sha256:e29e2d215434a45ab2a39ed73eb35fdf22287e324d2c08ff263b51364edf4833

Observation e2b4599b-4d25-4353-b51c-fec58c7b54bb · outbound

This paper cites D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition D$^2$ST-Adapter: Disentangled-and-Deformable Spatio-Temporal Adapter for Few-shot Action Recognition

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-12T18:42:03.014767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.558643Z digest=sha256:4462de45a6cb2e10f042d78dd158034ad0e9400a8042969f7c9a39f41e650c3a

Observation 9e7843c7-d288-4823-bbbb-296eae2583c2 · outbound

This paper cites Deep residual learning for image recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Deep residual learning for image recognition,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.564339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.564339Z digest=sha256:33806b5cc2b349f2e72880fd0728fbbe7269aa766d4a8b9322584bb123a8cf42

Observation 4ae4af88-45e4-4b09-a6f3-2dccd9e39d87 · outbound

This paper cites Imagenet: A large-scale hierarchical image database,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Imagenet: A large-scale hierarchical image database,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.372009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.570279Z digest=sha256:2fda13f7e12fad2d28702c9910f170bb0c5b2db8cb7a847cff615a5a11ad52c0

Observation a30b2c89-6cc8-4b9e-9102-38454f58180e · outbound

This paper cites Learning transferable visual models from natural language supervision,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Learning transferable visual models from natural language supervision,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.575592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.575592Z digest=sha256:2408a179a9d975e0efe1dc2b49be943e61e56cbf85129696874ef91f5620c521

Observation d9eee56d-a705-4827-a555-648f446e7e2e · outbound

This paper cites Parameter-efficient fine-tuning for pre-trained vision models: A survey,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Parameter-efficient fine-tuning for pre-trained vision models: A survey,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.580721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.580721Z digest=sha256:b62f9cc20c5a4c293b25eba423d1f2b3a76e48edcea525082f0406f6d48c061a

Observation 1e40d10d-b55f-4d02-af1c-2019e13eab5c · outbound

This paper cites Parameter-efficient transfer learning for nlp,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Parameter-efficient transfer learning for nlp,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.344523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.585567Z digest=sha256:6dcc79a8afaf4cc8a0090749a8ef073e89a9dcd46191e56cfd410e03ea3e7568

Observation 2882a639-0f4c-44f6-b705-4f1e41d18c3b · outbound

This paper cites Vision Transformer Adapter for Dense Predictions.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Vision Transformer Adapter for Dense Predictions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.590238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.590238Z digest=sha256:10a10a70eac07ccdd30a6217cb0118a07812229c873ca30e46b9a4ada4074a9a

Observation bc82d2fb-c967-48fe-8dd4-7ee3b4f06e1b · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.595141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.595141Z digest=sha256:2b2df3a0c03ef2a0efe73568db2dc58eb0cdce7c6585c1029dbdf5b32d69e791

Observation 4bbded91-8141-493a-ba1f-2bcd4bf1883b · outbound

This paper cites ActionCLIP: A New Paradigm for Video Action Recognition.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition ActionCLIP: A New Paradigm for Video Action Recognition

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.599612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.599612Z digest=sha256:852d8fc3f55b9b2bc6081c75b2c37084d994c726813c6e5a7f42322f613bca99

Observation 558e5194-a820-4afc-863b-3c8f064b5c54 · outbound

This paper cites Clip-adapter: Better vision-language models with feature adapters,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Clip-adapter: Better vision-language models with feature adapters,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.604222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.604222Z digest=sha256:2fa2003a610de8c0452e398705b33ad2d2e6d7be927018a3896275e740ce36e5

Observation 92322441-de22-4ab5-89fb-02e10578d85d · outbound

This paper cites Medical SAM Adapter: Adapting Segment Anything Model for Medical Image Segmentation.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Medical SAM Adapter: Adapting Segment Anything Model for Medical Image Segmentation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.608649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.608649Z digest=sha256:3e2eff9f758f81a0cd35074910ff590c850d5fc065110d2df0a42b7f700faa21

Observation 29c8ada3-031b-4be4-935f-5c438082544e · outbound

This paper cites Segment anything,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Segment anything,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.613484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.613484Z digest=sha256:5858897c84e865aed6a4d73ed9bd1996619f4d7f0a1b1137113888b08fe78dc3

Observation c43a917f-17bd-47a6-98c9-7867f08f3700 · outbound

This paper cites Bi-directional motion attention with contrastive learning for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Bi-directional motion attention with contrastive learning for few-shot action recognition,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.307299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.617790Z digest=sha256:14c56722392b0c2b7b1c442b0ea617df2344f7472a61674b9f942fc2b10eb1b2

Observation 29a89695-18f3-4073-a7d6-5aa76b73b911 · outbound

This paper cites Mlp-mixer: An all-mlp architecture for vision,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Mlp-mixer: An all-mlp architecture for vision,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.291130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.622670Z digest=sha256:fd3b51eebdc56199ddf74c3923ad884958485938e09d9c158281196627d2f079

Observation 56008d0a-9967-4db0-87de-6cdbc951657a · outbound

This paper cites The” something something.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition The” something something

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.274656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.628040Z digest=sha256:d39c4f082d5c50fba3e2647c4d67daec910e64d2c57a2fbf9b954ba8cb2cead2

Observation 7b03d5ae-862d-4307-98c5-72d7e7684b39 · outbound

This paper cites Quo vadis, action recognition? a new model and the kinetics dataset,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Quo vadis, action recognition? a new model and the kinetics dataset,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.633365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.633365Z digest=sha256:ac5939ee0d4d3b798646b3c4baa2f1b90d1dc0bf97987f721ca2894312a27642

Observation 0e87b565-9138-46b7-88aa-9315153ff7bf · outbound

This paper cites Hmdb: a large video database for human motion recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Hmdb: a large video database for human motion recognition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.245082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.638914Z digest=sha256:613f041e48ffd00dc7c0cd8db9b8fd98feb887288139b015e3e3d41e07f944d4

Observation 9e05ad2f-9dae-4563-9ecf-60f2c86582e8 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.644143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.644143Z digest=sha256:807e5744768f1fb9189ff1669fde0c01a33eb76942b1e72f92ea615cb618da65

Observation 38700434-ec99-4eb9-ac5e-2dd6481476ef · outbound

This paper cites Temporal segment networks: Towards good practices for deep action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Temporal segment networks: Towards good practices for deep action recognition,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.228569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.649876Z digest=sha256:c4c6a8634af3d8b4d9068b01c12dd0423d201b2205ce7c83aa4ba31f7e698036

Observation 0c45b591-4c8c-4def-80e5-0162aeb2d05f · outbound

This paper cites On the importance of spatial relations for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition On the importance of spatial relations for few-shot action recognition,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.212076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.654940Z digest=sha256:509e18ed16a0c2d2c1f8bb96920d24cf40724b4cae040846e28e0033d6e79c8c

Observation 26035880-a077-469d-b4c6-345585eda808 · outbound

This paper cites Matching networks for one shot learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Matching networks for one shot learning,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.193871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.659829Z digest=sha256:914d18251c11b9f2e8496bb7fdc734fdc69fef96f8c52621b6998c0a8c6ee8ec

Observation da44621a-57d8-40e5-9267-b38f7c556b08 · outbound

This paper cites Prototypical networks for few-shot learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Prototypical networks for few-shot learning,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.175206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.664813Z digest=sha256:75ee84d45685052d2ebc0a1dae7652d4f1c2145c487cc046da35cd30d79f0917

Observation 5424fe99-ed97-4ea8-85e3-e93fa1a9b159 · outbound

This paper cites Clip-guided prototype modulating for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Clip-guided prototype modulating for few-shot action recognition,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.669535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.669535Z digest=sha256:612d1023b82df6c90c9389c1b6f71d203fc17a880a142f05d7c39e77dd1570c1

Observation 55e17378-05c6-4f6a-911c-45beafd184eb · outbound

This paper cites Few-shot action recognition with hierarchical matching and contrastive learning,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Few-shot action recognition with hierarchical matching and contrastive learning,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.141399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.674499Z digest=sha256:dc06d85cce0a20939cfa9339bb967ece1189affb28382229c8656fea441522d4

Observation 38fa62f3-cdeb-4f36-8c85-dc07ddd78a4b · outbound

This paper cites Multidimensional prototype refactor enhanced network for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Multidimensional prototype refactor enhanced network for few-shot action recognition,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.122467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.679452Z digest=sha256:d7df65528c51e6483aa7d5e4af0016dfb68b33beae47c1d78d5358e7c7ff4f56

Observation 8c5c2a09-2650-45d8-841e-bae5b6125e19 · outbound

This paper cites Matching compound prototypes for few-shot action recognition,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Matching compound prototypes for few-shot action recognition,

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.104856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.684385Z digest=sha256:4bf8bf5f5272fa31221dfb2a7420dae16dd3b14b7a431f4691363180175da4bc

Observation e31de9af-4a97-409e-8b60-ceaf184bac43 · outbound

This paper cites Aim: Adapting image models for efficient video understanding,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Aim: Adapting image models for efficient video understanding,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.087593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.689236Z digest=sha256:9b68f8e0fba4c761aa2af2e86eda0ce9899456ab39e3f6dd1eb3cc34c58d5145

Observation 26c541da-5f47-4732-b593-a5e1d05aa5af · outbound

This paper cites Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Blip-2: Bootstrapping language- image pre-training with frozen image encoders and large language models,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T18:42:03.069954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-12T18:42:02.694022Z digest=sha256:a84ebaecfbc6e5a9522526340a20ae0c560db40fb666b213af1c5c3164af12c5

Observation b5bf5290-6431-499c-a927-98b9b97b5806 · outbound

This paper cites Visualizing data using t-sne.

Video-to-Task Learning via Motion-Guided Attention for Few-Shot Action Recognition Visualizing data using t-sne

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-12T18:42:02.699274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:42:02.699274Z digest=sha256:17607ce9331d3638ed1d4ae942416cf25bee7c13e1ad4272343e29886ce954bd

Pith citing papers

No inbound Pith citation observations are available.