Pith. sign in

Paper Citation Record · LEDGER

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation

As of 24 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.10810.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10810 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:15.859092Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact2
  • verified fuzzy40
  • unresolved10
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6dbd5909-7a62-4a75-93cc-02983ffbd1f0 · outbound

This paper cites Skeleton- aware networks for deep motion retargeting.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Skeleton- aware networks for deep motion retargeting

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.424278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.672560Z digest=sha256:c6cf3ad0becba273cd82accdd449ac9205c49ec567ec08ab4efdab4067070440

Observation cfc22bc0-08ad-412e-84c9-cdeefe734133 · outbound

This paper cites Text2action: Generative adversarial synthesis from language to action, 2017.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Text2action: Generative adversarial synthesis from language to action, 2017

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.414890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.676882Z digest=sha256:3d7f93d0e47289bd2882aed8606d895d867cca44e81b508d910148c8d571fda8

Observation e4545b28-ba0b-4043-8cd5-c9bfd1ed7f1c · outbound

This paper cites Lan- guage2pose: Natural language grounded pose forecasting,.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Lan- guage2pose: Natural language grounded pose forecasting,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.404921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.681533Z digest=sha256:ce605578ce9cf9b4682984850cb0970e25c7b2a6b80890564243decfc6f6fe5b

Observation 072669d4-a2c2-4e36-8c56-bc8f92c283b7 · outbound

This paper cites Make-an-animation: Large-scale text- conditional 3d human motion generation.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Make-an-animation: Large-scale text- conditional 3d human motion generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.395438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.685314Z digest=sha256:2403aa88df9699c2aa0bfbf4d47f040eb4890de362c29e273ce0e2ceced6a358

Observation 5737223f-92a3-4959-83f7-733997754e9c · outbound

This paper cites MoFM: A Large-Scale Human Motion Foundation Model.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation MoFM: A Large-Scale Human Motion Foundation Model

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:07:15.933563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.689404Z digest=sha256:8b664fd25577ac01b31867a268106fba08a7193a6a12f3883ef28cec987d9b07

Observation f4e9aaec-9137-4a03-8f4c-9927d6432e20 · outbound

This paper cites Executing your commands via motion diffusion in latent space.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Executing your commands via motion diffusion in latent space

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.385500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.693407Z digest=sha256:797816c2007b748a349b4093d43873bdc4ceac98b869a96643b808038dfb975d

Observation f70763c8-c69a-4b11-95b5-829d3e12e730 · outbound

This paper cites Livephoto: Real image animation with text-guided motion control.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Livephoto: Real image animation with text-guided motion control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.375261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.696956Z digest=sha256:a0b6b324cb785ed380c00278d6c6a51db4727986f61c13cdf8f442dd8fd2d4a0

Observation 6fe300ae-5bf4-4da2-b6c1-8ae636ced612 · outbound

This paper cites Channel-wise topology refinement graph convolution for skeleton-based action recognition.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Channel-wise topology refinement graph convolution for skeleton-based action recognition

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.365202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.700850Z digest=sha256:637336e888f9ac4ab21f1d8fc083b527c6dc4868a7fb8deeedf3a1d50fb6198a

Observation 261f99c1-c257-48e9-a8e3-22a0b23a535b · outbound

This paper cites Mofusion: A framework for denoising-diffusion-based motion synthesis.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Mofusion: A framework for denoising-diffusion-based motion synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.354790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.704248Z digest=sha256:d48e7387bd7295716b908b7c5329034a030c3dd329a83dea5f54c4e7bc99908f

Observation f763dbc5-9771-4f3e-8217-3300023bb29f · outbound

This paper cites Avatars grow legs: Generating smooth human motion from sparse tracking in- puts with diffusion model.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Avatars grow legs: Generating smooth human motion from sparse tracking in- puts with diffusion model

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.343965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.707609Z digest=sha256:93b97754f6a68fb03fcd1c6f3d5fd6b0aad6178dd40c0c7b1fd5f9c27d411ba3

Observation a1ffecdc-f8b1-4e23-be53-c8d7b9dec5f9 · outbound

This paper cites Transformer- based generative adversarial networks in computer vision: A comprehensive survey.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Transformer- based generative adversarial networks in computer vision: A comprehensive survey

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.333416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.710966Z digest=sha256:d254af0b9736935e0bf09ae67a308cc931ec5db27ac4962275a57ea2f8db28c9

Observation a925de05-4050-4353-9dda-2ceed2f41d07 · outbound

This paper cites Esser, R.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Esser, R

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.323094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.714159Z digest=sha256:da7644b8a3f7dbc60aef66468a5510b705b55511aa1e1d0936b123af0d501de7

Observation 0205ff2e-7475-4fe2-82d1-47296a1fd466 · outbound

This paper cites Ac- tion2motion: Conditioned generation of 3d human motions.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Ac- tion2motion: Conditioned generation of 3d human motions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.312791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.717640Z digest=sha256:dcc8ccb7dc6b85e625d9143ce3be0c2fa393e669256081d0c58adf60f51f06ed

Observation c42b00eb-cc0b-4b5b-924b-ab8308427e18 · outbound

This paper cites Generating diverse and natural 3d human motions from text.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Generating diverse and natural 3d human motions from text

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.302684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.720775Z digest=sha256:5baa1060f5b73b7ba040c1a8a22f5880937155e70ecbbb7fe91c74256f836715

Observation 928de684-9dfb-469d-ad05-473171468e7b · outbound

This paper cites Tm2t: Stochastic and tokenized modeling for the reciprocal genera- tion of 3d human motions and texts.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Tm2t: Stochastic and tokenized modeling for the reciprocal genera- tion of 3d human motions and texts

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.291361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.724668Z digest=sha256:20552bd936c2563126f131f711c876e4d67b8dc6b21fe3a342ef8cd40ac882b5

Observation 54dbd4d6-5ef2-4bfa-a0f0-b59e8e9107b5 · outbound

This paper cites Momask: Generative masked model- ing of 3d human motions.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Momask: Generative masked model- ing of 3d human motions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.280292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.727811Z digest=sha256:8dcf8c7734213bd51922a2d569b297eaaa682a5434049bac3424c123b5723d4c

Observation 04dd317f-e327-4722-8a4b-6acc8278ced4 · outbound

This paper cites BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation BAD: Bidirectional Auto-regressive Diffusion for Text-to-Motion Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.731241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.731241Z digest=sha256:1d0efcde69fb4073b2a6e08465d7163241cab2d308e7be2ae344a1d69584e926

Observation ee9a58a2-4f78-4d29-b091-85733da2c9d8 · outbound

This paper cites A recurrent variational autoen- coder for human motion synthesis.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation A recurrent variational autoen- coder for human motion synthesis

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.269568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.734863Z digest=sha256:85a25277b5210fae2a90d29c858256f9e52b7f8dd99e032c313d15cbbbaf084b

Observation 99603f61-3dba-4277-97e9-87dfa2070fd7 · outbound

This paper cites Omg: Towards open-vocabulary motion generation via mix- ture of controllers.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Omg: Towards open-vocabulary motion generation via mix- ture of controllers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.258073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.738122Z digest=sha256:c76e7ec6d9b128f996c9ba1be283ec510a62f8c3847e520a6d985d35c6cc3a75

Observation 3fb5733c-fbf5-40a0-83d7-0883247a58a8 · outbound

This paper cites Fully Fine-tuned CLIP Models are Efficient Few-Shot Learners.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Fully Fine-tuned CLIP Models are Efficient Few-Shot Learners

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:07:15.907849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.741654Z digest=sha256:87c2334bea67286933e60f14411362151c0856884cb1acd985b5c048b51ed90e

Observation 7f0ff4f3-3bb9-42ac-a1bd-f33b705d2457 · outbound

This paper cites Disentangling and unifying graph con- volutions for skeleton-based action recognition.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Disentangling and unifying graph con- volutions for skeleton-based action recognition

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.247025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.745447Z digest=sha256:2e08c829120e52350ad5a577087662c5560cd69ef642667b4bd6859a09bf5cca

Observation deda2ee9-a0b1-4f03-b078-90a1f59f7d4c · outbound

This paper cites Amass: Archive of motion capture as surface shapes.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Amass: Archive of motion capture as surface shapes

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.749773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.749773Z digest=sha256:8e2fb58b12a72436cd1241e725dcfa86ddaa7926132ff64c458227b2ad6f6c0c

Observation d3d107fc-4dfe-45bc-997f-24b7d70efc24 · outbound

This paper cites Fine-tuning can cripple your foundation model; preserving features may be the solution.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Fine-tuning can cripple your foundation model; preserving features may be the solution

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.753153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.753153Z digest=sha256:cd34bd239b741a320e84e1813006045423703094f1497f82c568944b7df252da

Observation 86571145-d410-4076-b280-dd8a2c3bb643 · outbound

This paper cites An exploratory study on human-centric video anomaly detection through variational autoencoders and trajectory prediction.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation An exploratory study on human-centric video anomaly detection through variational autoencoders and trajectory prediction

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.756948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.756948Z digest=sha256:be14ee4d6bd6844794f0245f0be896afbf037c5a990fa2a2e352b1754ce3ca32

Observation 51b8e465-e521-4ad3-8322-8baf80ab8042 · outbound

This paper cites Ancilia: Scalable intelligent video surveillance for the artificial intelligence of things.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Ancilia: Scalable intelligent video surveillance for the artificial intelligence of things

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.223579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.760495Z digest=sha256:2a6e13c1ae1b193f6fabf56ccf1b2a95969f66d20bc32f23787c864097de4b59

Observation bde71c9b-4a1e-4cde-805c-43cf914b085c · outbound

This paper cites A survey of graph-based deep learning for anomaly detection in distributed systems.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation A survey of graph-based deep learning for anomaly detection in distributed systems

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.212281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.764536Z digest=sha256:5d6f6e7bd9c56b224554fed5e904b34e7a5adc194d5a02063988037512415833

Observation 067cf77d-b3bd-4782-a35b-ece7d9b286e2 · outbound

This paper cites Vt-former: An exploratory study on vehicle trajectory prediction for highway surveil- lance through graph isomorphism and transformer.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Vt-former: An exploratory study on vehicle trajectory prediction for highway surveil- lance through graph isomorphism and transformer

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.201516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.768043Z digest=sha256:763ae467c6d0f1a84588e9935bfc926b7f434b080c276907e70cb0231b655afe

Observation 03ea2d94-68df-40a8-828c-dbeea6c862e9 · outbound

This paper cites Temos: Generating diverse human motions from textual descriptions.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Temos: Generating diverse human motions from textual descriptions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.771371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.771371Z digest=sha256:eaaca54222175749dcdb6b92f9920168b98e10408f90f575ea10ab46e4bbae12

Observation 93277905-7313-4b45-94ed-af181531f350 · outbound

This paper cites Bamm: Bidirectional autoregressive motion model.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Bamm: Bidirectional autoregressive motion model

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.178131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.779002Z digest=sha256:60d5063a6ff4fb9fd385d4b6e1538b9a5b8bfc823e61d19a929043ef4904df72

Observation 315f8342-ba0a-40c8-9dc6-0329332f12db · outbound

This paper cites Mmm: Generative masked motion model.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Mmm: Generative masked motion model

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.167203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.782378Z digest=sha256:761566d5425383ba2c18f048ad4b1675abaf13d3d9a204a97c58f598dd14c7f9

Observation 6a3880da-5556-4286-afc5-0b6f109fd3a0 · outbound

This paper cites The kit motion-language dataset.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation The kit motion-language dataset

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.785755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.785755Z digest=sha256:c187613f8f3b0a35b5273fe296a09833fe5ea0a872ff707065e3010c7669cc6c

Observation f5a11153-d70a-44b6-b8bf-57278ad27f30 · outbound

This paper cites Skeleton-based action recognition via spatial and temporal transformer networks.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Skeleton-based action recognition via spatial and temporal transformer networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.149712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.789637Z digest=sha256:080c4fa0ed352bb7203be51ef3fe8d7b091308ef793b5d694fe2a905231bf13a

Observation a4ef5244-7ddf-4a88-8c50-0d95ddfe9747 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Learning transferable visual models from natural language supervision, 2021

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.137658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.793966Z digest=sha256:bb44a61f7e86ff3ba6f22039ca1ff430265c60f6ef6b202f46b7a114ff619206

Observation 3394c3f3-7150-4006-b62e-bd1b46263c7c · outbound

This paper cites Guided attention for interpretable mo- tion captioning.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Guided attention for interpretable mo- tion captioning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.126577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.798056Z digest=sha256:94fe2b5a81435285536ce4bd168dbbaadcbf1612c0fa1b7a8c9b60674cd9dec9

Observation eccb294a-aa92-43cb-a802-04d3d4cddb3f · outbound

This paper cites Generat- ing diverse high-fidelity images with vq-vae-2.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Generat- ing diverse high-fidelity images with vq-vae-2

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.115451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.801330Z digest=sha256:14ed0c4a46579662fe9439768c102863706ec5024380e8bea7d14a70b5a258d7

Observation ab0af8b5-1384-4337-8e26-b85589b17448 · outbound

This paper cites Skeleton-based action recognition with multi-stream adap- tive graph convolutional networks.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Skeleton-based action recognition with multi-stream adap- tive graph convolutional networks

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.104748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.805295Z digest=sha256:dfc2fccfb4301903cf7e3bbf12268bcafd52a54e7aa8429bcc42dfea34b18103

Observation 747d9abf-31b0-4f72-be8b-2f63767bb9a3 · outbound

This paper cites Opinion unaware image quality assessment via ad- versarial convolutional variational autoencoder.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Opinion unaware image quality assessment via ad- versarial convolutional variational autoencoder

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.093853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.809451Z digest=sha256:cf50e9fec4817cc85978730ce9518892e098f665e281d76f45411de6d16d8737

Observation 0acf5b99-d5dd-411a-be6c-25c41dc0a5ac · outbound

This paper cites Curobo: Parallelized collision-free robot mo- tion generation.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Curobo: Parallelized collision-free robot mo- tion generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.812808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.812808Z digest=sha256:b1cd3301648cc5eadd06acab6dc775714f76d339ca1078b244552f377cde303a

Observation f03457ee-c65b-4fe0-bd54-e11251252657 · outbound

This paper cites Rethinking the inception archi- tecture for computer vision.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Rethinking the inception archi- tecture for computer vision

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.815992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.815992Z digest=sha256:a4d0beebc45e99832ea3c0d18b3145747d9de2298dc6dd1fde7327cf4fe30005

Observation 9ccc6fc6-6bc8-4bf3-9343-bc088b09e645 · outbound

This paper cites Bermano, and Daniel Cohen-Or.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Bermano, and Daniel Cohen-Or

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.070012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.819077Z digest=sha256:5d40490810c0f0cd82e77c882a9506af89554675f7bd127299cbc565adc66a4a

Observation 21643f99-abb5-4d73-b0f1-661edecacc35 · outbound

This paper cites an unresolved cited work.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:07:16.060180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.822462Z digest=sha256:825a375ab10c260bcd585d3eed48630e1c24294dcbe9508f11f3fb0f9e5b4933

Observation b96213ac-0157-40b5-bdb4-3cd61b6b42a0 · outbound

This paper cites Relmogen: Integrat- ing motion generation in reinforcement learning for mobile manipulation.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Relmogen: Integrat- ing motion generation in reinforcement learning for mobile manipulation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.050625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.825600Z digest=sha256:993570e0862148ce733c28921150bb4027a6489159e56855229b38aaf29693e3

Observation aa5c0042-bb31-48df-8cf1-8663d6418974 · outbound

This paper cites Autore- gressive queries for adaptive tracking with spatio-temporal transformers.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Autore- gressive queries for adaptive tracking with spatio-temporal transformers

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.040494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.828754Z digest=sha256:21625b2b4b41c37a11f0efeec73d2a1ad9f6bfa05c3fce1cea71c705639e6f58

Observation eb46b185-cfcc-431b-a344-cdf6d0bb5940 · outbound

This paper cites Improving viewing experiences of first-person shooter gameplays with automatically-generated motion effects.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Improving viewing experiences of first-person shooter gameplays with automatically-generated motion effects

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.029618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.832142Z digest=sha256:0c0a3cdaf5c9940b0eabb8929198e8f8988f57d720a2c57da088225bcecebfb0

Observation 3a32bd5f-41f5-489e-be97-0bc6f4b32b37 · outbound

This paper cites T2m-gpt: Generating human motion from textual de- scriptions with discrete representations.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation T2m-gpt: Generating human motion from textual de- scriptions with discrete representations

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.019139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.835252Z digest=sha256:4d3c97e4057785036b5b750eee35b85e5adac8032c24d70563597e19bb06f2c0

Observation 45102c67-acb5-4d86-bb9b-f2387019354c · outbound

This paper cites Generating human motion from textual descrip- tions with discrete representations.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Generating human motion from textual descrip- tions with discrete representations

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:16.008860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.838218Z digest=sha256:3b735835ecb0a493e4b935d5682afc32566d39072dc4a033a64d276c2acc8e59

Observation 0f1d04e9-780c-47d3-a3b4-15bd9f5d893e · outbound

This paper cites Motiondif- fuse: Text-driven human motion generation with diffusion model, 2022.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Motiondif- fuse: Text-driven human motion generation with diffusion model, 2022

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:15.998196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.841559Z digest=sha256:83888111a34d5657ac79d4571e8665626efbbf952f41da3b59a07acbf157e7d2

Observation f4849938-15ed-445b-b1ce-b90b3c5122d9 · outbound

This paper cites Re- MoDiffuse: Retrieval-augmented motion diffusion model.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Re- MoDiffuse: Retrieval-augmented motion diffusion model

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:15.988146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.844897Z digest=sha256:00ab2be221801fefd7c814bbe0ec335078ad0531df8227d0593b2c4334df8501

Observation 2470cb4c-4bdb-4c92-89bd-64a3eb97ee19 · outbound

This paper cites Large motion model for unified multi-modal motion generation.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Large motion model for unified multi-modal motion generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:15.848309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.848309Z digest=sha256:f112d2a881134c65276bed7ecd59069a8e46b12f3a53897e00affe15c06a1a3b

Observation e0fca982-8711-4f70-8a5b-7d2557399373 · outbound

This paper cites Pose-to-motion: Cross- domain motion retargeting with pose prior.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Pose-to-motion: Cross- domain motion retargeting with pose prior

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:15.971062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.851984Z digest=sha256:cbc5e903442d85675cebf9c38f39c75e3e9078559b23fadb2f14efefa685f3ed

Observation ae4dc78c-b3fb-4f62-9f82-07ac95fa76d4 · outbound

This paper cites Senm-vae: Semi-supervised noise modeling with hierarchical variational autoencoder.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Senm-vae: Semi-supervised noise modeling with hierarchical variational autoencoder

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:15.957336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.855372Z digest=sha256:ac5380eb5cc375cb269f31ae18b6ea3708392cc6e8b8ded526dd51a66fbfff5c

Observation 133ee581-35f6-45f1-b80f-8a9cfedac243 · outbound

This paper cites Human motion generation: A survey.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Human motion generation: A survey

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:07:15.945912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-08-15T21:07:15.859092Z digest=sha256:c7017c9cdfbf5a8abc827feeb5ce68d6b5ac55fb7b6d8dd7273b6cc5267e9eee

Observation 4f6fa141-ba90-49bd-95ad-2cfddc0698e4 · outbound

This paper cites an unresolved cited work.

MoCLIP: Motion-Aware Fine-Tuning and Distillation of CLIP for Human Motion Generation Unresolved cited work

Reference 497

Resolution
parse uncertain
no resolver link, observed 2026-08-15T21:07:15.775142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:15.775142Z digest=sha256:1b50cca9da98fbfbeb73b5d0ccaf444c84875fce15676fc82ebefc256aa7b6b8

Pith citing papers

No inbound Pith citation observations are available.