Pith. sign in

Paper Citation Record · LEDGER

Punching Bag vs. Punching Person: Motion Transferability in Videos

As of 22 August 2026, this Paper Citation Record lists 72 of 72 outbound references and 0 inbound Pith citation observations for arXiv:2508.00085.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.00085 v1

Coverage vector

measured 72 of 72 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:28:54.659185Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

72 of 72 outbound references displayed

  • verified exact0
  • verified fuzzy69
  • unresolved2
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4115693f-33fc-4219-8181-005da5fc9ce0 · outbound

This paper cites Ez- clip: Efficient zeroshot video action recognition, 2024.

Punching Bag vs. Punching Person: Motion Transferability in Videos Ez- clip: Efficient zeroshot video action recognition, 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.464879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.442094Z digest=sha256:f18a7f40295b1c54c2307623906d740c7aa4a69881660967d35eff0f0227be5f

Observation 405fec8c-1b4a-42ea-b130-d70d1fa5dcc0 · outbound

This paper cites T2l: Efficient zero-shot action recognition with temporal to- ken learning.

Punching Bag vs. Punching Person: Motion Transferability in Videos T2l: Efficient zero-shot action recognition with temporal to- ken learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.454224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.445834Z digest=sha256:6c3e6038c24fafa7c81f39a071795ad9bcb09bdc3f8be2162452ae6956831130

Observation dffcc542-6085-409e-b5ad-c3a9a9ee7130 · outbound

This paper cites Are visual- language models effective in action recognition? a compara- tive study, 2024.

Punching Bag vs. Punching Person: Motion Transferability in Videos Are visual- language models effective in action recognition? a compara- tive study, 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.442867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.449338Z digest=sha256:c21280fae21aaa9831c6f096f8681fef8d323099bab379cae0760c954def1ad9

Observation 555e39a9-946b-4ec2-bd48-630e798d63c3 · outbound

This paper cites Understanding depth and height percep- tion in large visual-language models.

Punching Bag vs. Punching Person: Motion Transferability in Videos Understanding depth and height percep- tion in large visual-language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.432532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.452523Z digest=sha256:6eef0afc5a84977a978392877506bc2295e80c5c917c87f52e388e5a3c363884

Observation 842c855e-97d9-468c-abf2-518e5c1b5ca6 · outbound

This paper cites Hierarq: Task-aware hierarchical q-former for enhanced video understanding.

Punching Bag vs. Punching Person: Motion Transferability in Videos Hierarq: Task-aware hierarchical q-former for enhanced video understanding

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.418886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.456042Z digest=sha256:6dc71bd78364b7ad472633bf432b4c0ff883f3cf959752ce9f1f430019a3a66a

Observation 26c3acdc-df14-4c74-8f83-c0dec426f1ef · outbound

This paper cites Rethinking zero-shot video classi- fication: End-to-end training for realistic applications, 2020.

Punching Bag vs. Punching Person: Motion Transferability in Videos Rethinking zero-shot video classi- fication: End-to-end training for realistic applications, 2020

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.408820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.459258Z digest=sha256:0c980ed5753223f787fa265eaa4f933e22b2aba96ca8696d9c839b5cbda0f207

Observation 40bef970-52e2-428d-9895-065ab544d01e · outbound

This paper cites Temporal attentive alignment for large-scale video domain adaptation.

Punching Bag vs. Punching Person: Motion Transferability in Videos Temporal attentive alignment for large-scale video domain adaptation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.397834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.462416Z digest=sha256:5fac43dd033116dd00a2ae856da8edf70513c39d76cdf612c9f57d4da0fd5f87

Observation a2c48975-54a3-4ac5-8a6f-265a4387d171 · outbound

This paper cites Elaborative rehearsal for zero-shot action recognition, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos Elaborative rehearsal for zero-shot action recognition, 2021

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.387335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.465332Z digest=sha256:c37fc5a5694d92d962f8e873174fc3750ecc8a00263d0d8cb5bd8dd73a59c50a

Observation b48f7bae-281d-483e-9475-52fee87def90 · outbound

This paper cites Unsupervised and semi-supervised domain adaptation for action recognition from drones.

Punching Bag vs. Punching Person: Motion Transferability in Videos Unsupervised and semi-supervised domain adaptation for action recognition from drones

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.376102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.468832Z digest=sha256:326305bee992a5ef2d220445f18e7ac4960297269ec0cdd93f7d91fdfd7d75da

Observation 8f057fac-3d8c-48de-95e5-9b246f5dd664 · outbound

This paper cites Blender: A 3d modelling and rendering package.

Punching Bag vs. Punching Person: Motion Transferability in Videos Blender: A 3d modelling and rendering package

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.363833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.471879Z digest=sha256:ab5fd982401c4ed5d2371148d931a735ef31fe21148a28f29e8f2cbe92582284

Observation ddca8ede-8dbf-4066-8def-58ea508acc06 · outbound

This paper cites Toyota smarthome: Real-world activities of daily living.

Punching Bag vs. Punching Person: Motion Transferability in Videos Toyota smarthome: Real-world activities of daily living

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.352878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.474848Z digest=sha256:7aa8371cb33297906444dc73dd6fdb8e623f51fa48903527a8349051b39452e7

Observation 51ada693-b875-4acd-a6e1-7be1e5956307 · outbound

This paper cites Tinyvirat: Low-resolution video action recognition.

Punching Bag vs. Punching Person: Motion Transferability in Videos Tinyvirat: Low-resolution video action recognition

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.342081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.477665Z digest=sha256:2a12fa1e8399b0632b07aa63a6f3c7d51cc7dfeeee71c69a48e160e7c8fa73b7

Observation f0ff23fe-fe36-4e6f-a321-f3ce0d58b86d · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.330466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.480897Z digest=sha256:169e246fab9c2315e82ec36865d8367ca364acafe27ffef769ed06e0d5fb9e65

Observation 553bceaa-117c-4b9f-aa92-70bc76cb6523 · outbound

This paper cites Video- capsulenet: a simplified network for action detection.

Punching Bag vs. Punching Person: Motion Transferability in Videos Video- capsulenet: a simplified network for action detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.318689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.483836Z digest=sha256:25c7b2bb92ec0c04dc24ce7d113dddc4d35e4609a2e0b17e360ca311786850d0

Observation b87ed6d3-4280-43d1-94aa-4343d13a4bfa · outbound

This paper cites Pyslowfast.

Punching Bag vs. Punching Person: Motion Transferability in Videos Pyslowfast

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.306953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.486818Z digest=sha256:01019d28e33e4fba36207d07a30c8eb3748aed18098bdd2cccd81b8c187c3fa0

Observation 33793a46-2486-4553-a5c9-978cf381f416 · outbound

This paper cites Multiscale vision transformers, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos Multiscale vision transformers, 2021

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.295643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.489800Z digest=sha256:c330348994aadc6d95b8b5ff9f1596402e6aacf11c399d5f60a92666b8a284b0

Observation 9e6c1341-cacb-46f8-b59b-8e298e2138f8 · outbound

This paper cites X3d: Expanding architectures for efficient video recognition, 2020.

Punching Bag vs. Punching Person: Motion Transferability in Videos X3d: Expanding architectures for efficient video recognition, 2020

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.285152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.492568Z digest=sha256:5892f8754eace73da6bd6c8b5ded97a9854175cbe6ddfba3adec8f95250cc0c4

Observation 2aee4bb9-9796-48e8-9f10-9624da7d811d · outbound

This paper cites Slowfast networks for video recognition, 2019.

Punching Bag vs. Punching Person: Motion Transferability in Videos Slowfast networks for video recognition, 2019

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.273232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.495516Z digest=sha256:ed2d7f6fd7b37bd9e1450f9fde3cfa0adbfbad9baf71bd6cedaa03e22b24d0be

Observation c78e5c46-f59a-4d99-9ff2-a5241bd8123d · outbound

This paper cites Telling stories for common sense zero-shot action recognition, 2024.

Punching Bag vs. Punching Person: Motion Transferability in Videos Telling stories for common sense zero-shot action recognition, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.262175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.498488Z digest=sha256:b0c33cae9204dd917938fc963e5131fcd1d195ea3c31c05acde5726dfc3a3147

Observation 18d9ac81-4234-4882-bb4f-714c43027814 · outbound

This paper cites A new split for evaluating true zero-shot action recognition, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos A new split for evaluating true zero-shot action recognition, 2021

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.249222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.501360Z digest=sha256:cf38137fcaa54e6807f6b00f76e1a96268e0ab68997b4c36017b53d9b98b363a

Observation 66a92dda-0919-4f64-b1c7-3f55e16c67c2 · outbound

This paper cites The ”something something” video database for learning and evaluating visual common sense,.

Punching Bag vs. Punching Person: Motion Transferability in Videos The ”something something” video database for learning and evaluating visual common sense,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T10:28:54.504909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:28:54.504909Z digest=sha256:44e882983c495a8478960d42e33220316b32d2c6e9bac9cda8da7a2c45c660ed

Observation 9a60b9a3-745a-4636-b1d3-2804aa873f59 · outbound

This paper cites Hierar- chical explanations for video action recognition, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Hierar- chical explanations for video action recognition, 2023

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.230960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.508185Z digest=sha256:f26b3aa4bf50462223d3e241dfb898441c5e32be7069545c0f461a3773f4153c

Observation 8ed410d2-c1fe-4455-97fd-c8d866d95da5 · outbound

This paper cites Learn- ing spatio-temporal features with 3d residual networks for action recognition, 2017.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learn- ing spatio-temporal features with 3d residual networks for action recognition, 2017

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.218908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.511059Z digest=sha256:ef61beadc12947ce9243a44fe274c5e072144514cd7f81e825491a888a957125

Observation 2ba387f4-591a-4e58-9223-2267ee05f3c1 · outbound

This paper cites Deep residual learning for image recognition, 2015.

Punching Bag vs. Punching Person: Motion Transferability in Videos Deep residual learning for image recognition, 2015

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.208017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.514019Z digest=sha256:06e39beb17e8f522a0e2beeaff5e3a7e6bce311ab2be1391a04f39518f9fb7e9

Observation e38be4fe-896d-47d8-9618-c7d1c05f16d5 · outbound

This paper cites Froster: Frozen clip is a strong teacher for open-vocabulary action recognition, 2024.

Punching Bag vs. Punching Person: Motion Transferability in Videos Froster: Frozen clip is a strong teacher for open-vocabulary action recognition, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.197411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.517116Z digest=sha256:21f4e82f86c923b8d912b418a29feaf4fee5f8d9b0f34f31912d4ce0b29b697f

Observation eafeab26-b4f5-4b32-b2e7-583e9acd9a08 · outbound

This paper cites The kinetics human action video dataset, 2017.

Punching Bag vs. Punching Person: Motion Transferability in Videos The kinetics human action video dataset, 2017

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.185942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.520017Z digest=sha256:cb598e1ef53f3485e06698f4c012c41e1d3cf30e04944f7c91789b69be015dcc

Observation 75c6528c-4b98-47ce-93c0-57a4c8316807 · outbound

This paper cites Reformulating zero-shot action recognition for multi- label actions.

Punching Bag vs. Punching Person: Motion Transferability in Videos Reformulating zero-shot action recognition for multi- label actions

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.174693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.523289Z digest=sha256:cc74348538fe717e7e8252834f3a7e42dc3e6b5b21489cd754b6b2bdd2fbf199

Observation ccfcab97-9d39-4f05-a3b2-fbf7886a5449 · outbound

This paper cites Learning Cross-Modal Contrastive Features for Video Do- main Adaptation.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learning Cross-Modal Contrastive Features for Video Do- main Adaptation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.163962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.526412Z digest=sha256:f06f398738a0a1778fccaf42e85ef927a01f0f4b2dcdf1f9dd18e376d7fd8c76

Observation 708a4cc8-a1d7-4cc9-9875-8f6b1a58a02c · outbound

This paper cites Uniformer: Unified transformer for efficient spatiotemporal representation learning, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Uniformer: Unified transformer for efficient spatiotemporal representation learning, 2022

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.153544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.529621Z digest=sha256:fc4370093ee4967a7f62563f2c6f5f45a47561d043f3636ff9a10e58bc5bbca4

Observation e9110c7b-fb49-4584-9ab5-42bd222e837e · outbound

This paper cites Uniformerv2: Spatiotemporal learning by arming image vits with video uniformer, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Uniformerv2: Spatiotemporal learning by arming image vits with video uniformer, 2022

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.143098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.532725Z digest=sha256:fc8fcbc583992dc0bd65e8f88914298767c2887c4d8c0d9e12a4a31979b374d4

Observation e4486dd0-6061-4174-a596-993a2f0a080b · outbound

This paper cites Videochat: Chat-centric video understanding.

Punching Bag vs. Punching Person: Motion Transferability in Videos Videochat: Chat-centric video understanding

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.131467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.535858Z digest=sha256:1f220147319dfcf22d2e5163c4519d61e7ff4535d7c83c36cc7ad1ffbd7b813c

Observation d144fa42-07e1-4c0c-ac5a-30270d176b33 · outbound

This paper cites Mvitv2: Improved multiscale vision transformers for classification and detection, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Mvitv2: Improved multiscale vision transformers for classification and detection, 2022

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.120320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.539031Z digest=sha256:3ddec410ddeee9c2a1d24c16b7302ac65db9dfcb972ba42e0c35b57f90a19c0e

Observation 57936d54-29a2-4820-abf6-72c0a73ad01f · outbound

This paper cites Cross-modal representation learning for zero- shot action recognition, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Cross-modal representation learning for zero- shot action recognition, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.109500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.542082Z digest=sha256:d7e961c4712517aac950f4598461b85dafc95da05f48e43337768ea519451cc7

Observation b6e8e80e-5344-418f-b62f-76244a5cb4d8 · outbound

This paper cites Diversifying spatial-temporal perception for video domain generalization.

Punching Bag vs. Punching Person: Motion Transferability in Videos Diversifying spatial-temporal perception for video domain generalization

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.097919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.545030Z digest=sha256:67dcd1140c301b3e021214d8488431a4d8ea7f1239663baac42bcb6312c54a9d

Observation b1e6386c-a252-45eb-92d4-43f9adf7725f · outbound

This paper cites CREPE: Can Vision-Language Foundation Models Reason Compositionally?.

Punching Bag vs. Punching Person: Motion Transferability in Videos CREPE: Can Vision-Language Foundation Models Reason Compositionally?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T10:28:54.548143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:28:54.548143Z digest=sha256:8f308a9fa43e6bef6740608b787410c55c7fbdcd35b91e4f1e27aed913364686

Observation 22d44eac-6ed7-4c97-b361-bd56c84f71a2 · outbound

This paper cites Reversible vision transformers, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Reversible vision transformers, 2023

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.085656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.551508Z digest=sha256:608a2256385f432e22b6deb5a04439f9e30bd331798a5cf7e25723e3e219bb18

Observation 2eeb25f0-5cdd-4b83-991a-4435967d8168 · outbound

This paper cites Something-else: Com- positional action recognition with spatial-temporal interac- tion networks, 2020.

Punching Bag vs. Punching Person: Motion Transferability in Videos Something-else: Com- positional action recognition with spatial-temporal interac- tion networks, 2020

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.074646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.554992Z digest=sha256:2fe3e23f36156ca05f481edf8a855915a49261d853d902ed737ffb214fc8cfde

Observation e77c8da6-9f61-4534-a51c-cbc5903eddce · outbound

This paper cites Video action detection: Analysing limitations and chal- lenges.

Punching Bag vs. Punching Person: Motion Transferability in Videos Video action detection: Analysing limitations and chal- lenges

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.063652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.557869Z digest=sha256:33f84b097b3e7964dcf538992a47d1be447a7348525056897275c01db650e283

Observation 892671ef-091a-4b2e-940b-b18ec982919a · outbound

This paper cites Verbs in action: Improving verb understanding in video-language models, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Verbs in action: Improving verb understanding in video-language models, 2023

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.053080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.560757Z digest=sha256:4340ea154346bb1873d86616c9399ff875d6bb5a0f2e74916cbc86b9a5d644cf

Observation 26b927d0-a2b3-4e39-a858-4d3e2c9fd907 · outbound

This paper cites Multi-modal domain adaptation for fine-grained action recognition, 2020.

Punching Bag vs. Punching Person: Motion Transferability in Videos Multi-modal domain adaptation for fine-grained action recognition, 2020

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.040682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.563918Z digest=sha256:ffc463cbc012018f6b883bc5e689eed60ed3cbf227133cc0684ac286873542aa

Observation b6e0b3a6-c9d1-4d8a-a926-3452f9da15e5 · outbound

This paper cites Expanding language-image pretrained models for gen- eral video recognition, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Expanding language-image pretrained models for gen- eral video recognition, 2022

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.029829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.567472Z digest=sha256:0eac65ee3b5ca032514bc29fedf80ecbfe2adb30093eddc59141e23cd7bfaa9d

Observation 787d62e3-cb75-4f10-bf79-827980ed49bd · outbound

This paper cites Object-relation reasoning graph for action recognition.

Punching Bag vs. Punching Person: Motion Transferability in Videos Object-relation reasoning graph for action recognition

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.017675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.570569Z digest=sha256:0a00f6de8355dce5d8e6b9ad00aefde6a8b6139c154c761d2bec4c44957f6b31

Observation af83d9fd-e4c4-4241-87f6-ad5a6a0a9629 · outbound

This paper cites Relative norm align- ment for tackling domain shift in deep multi-modal classifi- cation.

Punching Bag vs. Punching Person: Motion Transferability in Videos Relative norm align- ment for tackling domain shift in deep multi-modal classifi- cation

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:55.005872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.573761Z digest=sha256:07cc11126ca7801b812ae98264ee00d05e777d476f2d264b89a3d3a363680191

Observation 68496f9b-d914-4812-92bb-e028a11de5d2 · outbound

This paper cites What can a cook in italy teach a mechanic in in- dia? action recognition generalisation over scenarios and locations.

Punching Bag vs. Punching Person: Motion Transferability in Videos What can a cook in italy teach a mechanic in in- dia? action recognition generalisation over scenarios and locations

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.994152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.576855Z digest=sha256:8998ea413f46cac8b30de6ac4009d07f5ea8c923e4068e7b89c9f3e6676597bc

Observation 38305df2-6c1c-408d-96c8-cb8443595034 · outbound

This paper cites Haupt- mann.

Punching Bag vs. Punching Person: Motion Transferability in Videos Haupt- mann

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.983155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.580251Z digest=sha256:5cba56d719d133b29df5b315483ce93f5976b712c91ab700efe9334c106a297a

Observation 79955c59-9471-4acf-a6b2-336cc4f9a065 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learning transferable visual models from natural language supervision, 2021

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.972694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.583114Z digest=sha256:2dd231d6a0cb9111f5cb3e088a2bfcefad87e32b7097344822b755b9325c66c5

Observation a7f1584f-51af-4333-bc8e-1fbdaf22e13e · outbound

This paper cites Fine-tuned clip models are efficient video learners, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Fine-tuned clip models are efficient video learners, 2023

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.961749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.586075Z digest=sha256:d4840c680e8666aa81e42a8ca47e222a05c47db2dc8e934d202d438af883b029

Observation f68b5a43-938a-499c-a654-944542e65419 · outbound

This paper cites Towards a fair evaluation of zero-shot action recognition using external data.

Punching Bag vs. Punching Person: Motion Transferability in Videos Towards a fair evaluation of zero-shot action recognition using external data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.949384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.588742Z digest=sha256:0333b40971de1218e6235617e7d6b04ebd150e3db40d6e8dde442e4176240a51

Observation 32fade93-23ff-4d3e-a88c-f4ba88d9696c · outbound

This paper cites Probing conceptual understanding of large visual-language models.

Punching Bag vs. Punching Person: Motion Transferability in Videos Probing conceptual understanding of large visual-language models

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.937728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.591622Z digest=sha256:c6f24a917446fe441a2f602299b76c244c595c863bcf4921e6be77fd32091e14

Observation 504c5f4c-144e-40eb-9581-96c993488552 · outbound

This paper cites A Large-Scale Robustness Analysis of Video Action Recognition Models.

Punching Bag vs. Punching Person: Motion Transferability in Videos A Large-Scale Robustness Analysis of Video Action Recognition Models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.927348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.594758Z digest=sha256:ba7ef6a17113732788c69e1c38882e401e25709d1c29817cc84dfca0826d8c3d

Observation d1362599-b1dc-4c80-8eb9-46476637bc94 · outbound

This paper cites Noisyactions2m: A multimedia dataset for video understanding from noisy la- bels.

Punching Bag vs. Punching Person: Motion Transferability in Videos Noisyactions2m: A multimedia dataset for video understanding from noisy la- bels

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.915584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.597449Z digest=sha256:3489c20cf1d4f1d0545588c3561ae56dd87ffb74ef5fcedcc4d9b020cfb9a610

Observation f1c0c31e-6a0d-41ee-8d07-c595b425f102 · outbound

This paper cites Learning long-term dependencies for action recognition with a biologically-inspired deep network.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learning long-term dependencies for action recognition with a biologically-inspired deep network

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.905442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.600228Z digest=sha256:91a380811d4ecf7f112028db0fdd0d543af00380e2221604abad499904aaf562

Observation 867f2761-c42e-44f3-bd2c-413abf879707 · outbound

This paper cites Dvanet: Disentangling view and action features for multi- view action recognition, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Dvanet: Disentangling view and action features for multi- view action recognition, 2023

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.894525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.603319Z digest=sha256:bd9b33f2a63615993b1968099cc35ab56158475fe3077f3d7492fb17ff414bd0

Observation ad117d2e-3ea7-4eb6-9f30-5bf3734c428e · outbound

This paper cites Spatio-temporal contrastive domain adaptation for action recognition.

Punching Bag vs. Punching Person: Motion Transferability in Videos Spatio-temporal contrastive domain adaptation for action recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.883346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.606637Z digest=sha256:b57cae846229c76b278b3eabc1bc554a2e1e92e11ff008b5765dfd7751975633

Observation 3a8a2d6c-b30e-4d13-b393-c9b9148f1d3f · outbound

This paper cites Learning That Transfers: Designing Curricu- lum for a Changing World.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learning That Transfers: Designing Curricu- lum for a Changing World

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.873357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.609775Z digest=sha256:c2b173c0cbad6b0587f6c7f0f5f7530c5c30ee20824d973e131a0a15b0a460aa

Observation 587e19bd-fdd5-4757-abeb-31ff0b16b62a · outbound

This paper cites Learning spatiotemporal features with 3d convolutional networks, 2015.

Punching Bag vs. Punching Person: Motion Transferability in Videos Learning spatiotemporal features with 3d convolutional networks, 2015

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.863239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.612819Z digest=sha256:4da141cbaf3cd6a06979ff8dd1a307a6bcc9ae17829678d369ee37670f7d1ec0

Observation 9eb3db18-6364-4f5a-ae47-36ee7542ec7a · outbound

This paper cites Actionclip: A new paradigm for video action recognition, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos Actionclip: A new paradigm for video action recognition, 2021

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.854016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.615730Z digest=sha256:ad0589022e671f7d200ac6c34512b82eaf0f13b56042720dce32c0145c828141

Observation bb5607ac-3cb2-424f-82c7-cc0bd58ec9af · outbound

This paper cites An efficient spatio-temporal pyramid transformer for action detection.

Punching Bag vs. Punching Person: Motion Transferability in Videos An efficient spatio-temporal pyramid transformer for action detection

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.844696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.618561Z digest=sha256:2a5ca6c78714e73bdb777285939dbf07ba7fdcbee21a8936fb4ef063571fed52

Observation 11fc7ff5-e0d3-41e4-a802-70a19262aa49 · outbound

This paper cites Action recognition using attention-based spatio-temporal vlad net- works and adaptive video sequences optimization.

Punching Bag vs. Punching Person: Motion Transferability in Videos Action recognition using attention-based spatio-temporal vlad net- works and adaptive video sequences optimization

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.834618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.621464Z digest=sha256:993b043cea4411847405e73560965236baccb272a80489cc24eba62b8eb11a58

Observation 2a226360-e3ad-4405-a521-2893ee49c6d0 · outbound

This paper cites Videoclip: Contrastive pre-training for zero-shot video-text understanding, 2021.

Punching Bag vs. Punching Person: Motion Transferability in Videos Videoclip: Contrastive pre-training for zero-shot video-text understanding, 2021

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.824981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.624228Z digest=sha256:c4ab3bfba7e8a971987e351c464c12c364af4fcc39a205206d044dc32148cb2d

Observation 28f64da9-fbcc-474d-84b5-18a267dadf6d · outbound

This paper cites Seman- tic embedding space for zero-shot action recognition, 2015.

Punching Bag vs. Punching Person: Motion Transferability in Videos Seman- tic embedding space for zero-shot action recognition, 2015

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.816065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.627336Z digest=sha256:2133b495ce344f4562352e6033949f624c72652eff32108d8c9070e2192d7622

Observation 3c6b1119-2439-4be1-9dc6-6a20fa1adf27 · outbound

This paper cites Interact before align: Leveraging cross-modal knowledge for domain adaptive action recognition.

Punching Bag vs. Punching Person: Motion Transferability in Videos Interact before align: Leveraging cross-modal knowledge for domain adaptive action recognition

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.806924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.630326Z digest=sha256:ede9d9523214d26368e7642d886bb96fe4d11293ad82a55dd30c1bfd1f935fb7

Observation 85357b52-8ea7-4850-a168-325123912e01 · outbound

This paper cites Aim: Adapting image models for efficient video action recognition, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos Aim: Adapting image models for efficient video action recognition, 2023

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.797794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.633190Z digest=sha256:ed19f20973ea8eff8c79774cf3996b62d114751e09ce5ec4d539bfa9ec0ab919

Observation 0cef7622-926d-4dc8-af62-601f20884879 · outbound

This paper cites Yu, and Mingsheng Long.

Punching Bag vs. Punching Person: Motion Transferability in Videos Yu, and Mingsheng Long

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.787184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.636128Z digest=sha256:844c8e664b6a14f6ca1877f936d3d33006e4400a975e16e9f858c0e0a4b334c7

Observation 7b86106e-d247-4527-934a-8f81ead14cc9 · outbound

This paper cites Action4d: Online action recognition in the crowd and clutter.

Punching Bag vs. Punching Person: Motion Transferability in Videos Action4d: Online action recognition in the crowd and clutter

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.777616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.638929Z digest=sha256:c0d917737592f46445cfba64395a551a66bbeced6cb3078592589d4aa2c13ef0

Observation e540d4b5-4806-4caf-856a-be6a38c73265 · outbound

This paper cites Eliciting in-context learning in vision-language models for videos through curated data dis- tributional properties.

Punching Bag vs. Punching Person: Motion Transferability in Videos Eliciting in-context learning in vision-language models for videos through curated data dis- tributional properties

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.767562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.641706Z digest=sha256:d407be92e481c12c58fb53eaa733949f36d4209860aab3989f7ddff4f3d1f720

Observation 5cd47e77-f37a-4a85-9fe7-41a933ac0fa3 · outbound

This paper cites Derpa- nis.

Punching Bag vs. Punching Person: Motion Transferability in Videos Derpa- nis

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.756813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.644639Z digest=sha256:b6ecfca5c8ccaf643c96af15fb18ee3b76cd3164554e76f6327e3f07145739f9

Observation bda1ab80-2b39-4446-a4d7-1e173f79c10b · outbound

This paper cites Human-object interaction detection via disentangled transformer, 2022.

Punching Bag vs. Punching Person: Motion Transferability in Videos Human-object interaction detection via disentangled transformer, 2022

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.746720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.647366Z digest=sha256:09d7e48ae9569aa09be4dc841824626b51624035f10a952086ef2360b94d82b0

Observation 3339db00-70e0-474d-9735-90022abe873b · outbound

This paper cites How can objects help action recognition?, 2023.

Punching Bag vs. Punching Person: Motion Transferability in Videos How can objects help action recognition?, 2023

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.735895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.650326Z digest=sha256:27f23b3835fb98f75c64184b4353a0dad62d743b0f324206657bd4953536cbbc

Observation 137d15bd-2b1c-4d35-b8bc-522730ff76cc · outbound

This paper cites Among multi- modal models, we experimented with different variations of CLIP [46] designed for activity recognition.

Punching Bag vs. Punching Person: Motion Transferability in Videos Among multi- modal models, we experimented with different variations of CLIP [46] designed for activity recognition

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.725793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.652926Z digest=sha256:67f9941e3fc7315615696daa04daa071e19cc8ff60c3671fb56654dd721d5702

Observation 85cb6713-3110-4fe6-904a-325892f1c910 · outbound

This paper cites pre-train, prompt, and fine-tune.

Punching Bag vs. Punching Person: Motion Transferability in Videos pre-train, prompt, and fine-tune

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T10:28:54.716403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.656142Z digest=sha256:75925af7a425bc7cabbc127e352ca72264e5d7417ebdace236a0acbd53bfdbb6

Observation ab13211d-4430-44ab-9f90-072ca0fa2e2e · outbound

This paper cites Pushing”) and the other focuses on fine-grained con- text (e.g., “Pushing something from left to right.

Punching Bag vs. Punching Person: Motion Transferability in Videos Pushing”) and the other focuses on fine-grained con- text (e.g., “Pushing something from left to right

Reference 72

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T10:28:54.704808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T10:28:54.659185Z digest=sha256:9c5032e6a0b8c77226334d0bb2ab3583b10343bafd787718e3ae34d526f4d5e3

Pith citing papers

No inbound Pith citation observations are available.