Pith. sign in

Paper Citation Record · LEDGER

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation

As of 7 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 3 inbound Pith citation observations for arXiv:2512.13840.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.13840 v3

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T16:24:11.789786Z

measured 81 of 81 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-25T02:38:14.542361Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T02:40:14.611495Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved77
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 489cce2e-addd-426c-a9ea-7ba2a434c490 · outbound

This paper cites TMR++: A cross-dataset study for text-based 3d human motion re- trieval.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation TMR++: A cross-dataset study for text-based 3d human motion re- trieval

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:03.735338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:03.735338Z digest=sha256:99d3cf23b058a393401cdf9d5463afcf472afdb639d997d720fd898cd057c1df

Observation d3efa9b9-fea3-4d37-ac04-8de1daab3d6c · outbound

This paper cites Keep it smpl: Automatic estimation of 3d human pose and shape from a single image.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Keep it smpl: Automatic estimation of 3d human pose and shape from a single image

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:03.836592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:03.836592Z digest=sha256:95775de735af3c6514f5eee26e1f4440fb52f20c3f8ac8235acdecdfeed11929

Observation 99d937e4-d9f1-4df4-aea7-fe56ba3e25e5 · outbound

This paper cites The language of motion: Unifying verbal and non- verbal language of 3d human motion.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation The language of motion: Unifying verbal and non- verbal language of 3d human motion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:03.949819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:03.949819Z digest=sha256:d5909dbd58bfda276113c33d9bcef2c83d0d4332cfed96ed206ccbc586351cfb

Observation 57945188-1b59-4058-9170-bb12fe383d7d · outbound

This paper cites Taming diffusion probabilistic models for character control.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Taming diffusion probabilistic models for character control

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.104321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.104321Z digest=sha256:aec053a01d09415da11acd206dfce3f50b86191f2c808e155fcbe4bfc2710629

Observation e5076399-c95a-48ee-9e30-ffbe118e78bf · outbound

This paper cites Free-t2m: Frequency enhanced text-to-motion diffusion model with consistency loss.arXiv preprint arXiv:2501.18232, 2025.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Free-t2m: Frequency enhanced text-to-motion diffusion model with consistency loss.arXiv preprint arXiv:2501.18232, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.254443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.254443Z digest=sha256:8cd12c0a26baa182e093e36c941768ec741cdc398eaa997d6fd96bb156c9e455

Observation a3eac881-5bef-46f9-a13d-ad71d4a635a3 · outbound

This paper cites Executing your commands via motion diffusion in latent space.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Executing your commands via motion diffusion in latent space

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.415455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.415455Z digest=sha256:8c22510d0e0ed297fd2852572354c666c2954c79c88d80eca2b9f6f3d3a4d092

Observation 0fd41aed-3507-4396-bc2c-e5ea6cadd327 · outbound

This paper cites DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation DisCoRD: Discrete Tokens to Continuous Motion via Rectified Flow Decoding

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.619572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.619572Z digest=sha256:88fdbae8450db13f0ff73a45b918b36488962d9bd05c2ff0afb6db83102d8ddb

Observation 6eddafe7-b3f4-4d7e-bf68-d15569393dde · outbound

This paper cites Mofusion: A framework for denoising-diffusion-based motion synthesis.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mofusion: A framework for denoising-diffusion-based motion synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.774401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.774401Z digest=sha256:0661738f1c3c0af4ab08ed5ae4fbab7c8da608ee12acf9b7f98a93f8b5b9038f

Observation 95b2e17e-757a-4020-89d4-9b664dd65e0c · outbound

This paper cites Motionlcm-v2: Improved compression rate for multi-latent-token diffusion,.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Motionlcm-v2: Improved compression rate for multi-latent-token diffusion,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:04.941472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:04.941472Z digest=sha256:e615c7df9638a3bb9e79a93d99729150dd5a72564c3fed3528e63fe70dc920a3

Observation d2560a55-a3ca-43e6-89e8-7ba504f697cd · outbound

This paper cites MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation MotionLCM: Real-time Controllable Motion Generation via Latent Consistency Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.142966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.142966Z digest=sha256:b7c5d7801fda4fa1043c6dc56909a2b8b394e5752fde1f1afe4f8c7374525631

Observation bb1b93a2-bec8-4f22-9543-e8a3e6da7d5b · outbound

This paper cites Sigmoid- weighted linear units for neural network function approx- imation in reinforcement learning.Neural networks, 107: 3–11, 2018.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Sigmoid- weighted linear units for neural network function approx- imation in reinforcement learning.Neural networks, 107: 3–11, 2018

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.305002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.305002Z digest=sha256:a32ce210a6fdd49f48be465d689d754665a5d54e58bff6a3596f853bb73623fd

Observation 9bf1fc10-c7fb-4d52-9cb6-3a634dc38829 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.477812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.477812Z digest=sha256:5382eedd46816666df6e173ee45be99a60f1ee0d7dceef6518f7ec2360276b13

Observation 2bb6e701-839e-45d6-89cc-b3f0e54367f8 · outbound

This paper cites Remos: 3d motion- conditioned reaction synthesis for two-person interactions.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Remos: 3d motion- conditioned reaction synthesis for two-person interactions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.671489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.671489Z digest=sha256:63015e2291c943449603d17a17555121d8ff2ebca3b3ee3b728f4097d148ddf3

Observation d0656859-f318-4303-9712-9517bd1c42d9 · outbound

This paper cites Duetgen: Music driven two-person dance generation via hierarchical masked modeling.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Duetgen: Music driven two-person dance generation via hierarchical masked modeling

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.788186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.788186Z digest=sha256:c1be4e09086e0696b11f8ff8c2e8e72ada965eeb1b66cfbec6698331052fc299

Observation 8ad7661a-6c1e-4505-ac79-3a9fa36653ac · outbound

This paper cites Ac- tion2motion: Conditioned generation of 3d human motions.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Ac- tion2motion: Conditioned generation of 3d human motions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.882970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.882970Z digest=sha256:37c1cd07dcbb26b3beab6353e266fcc8f5410dbbf4f667bc72030334ce390ece

Observation 2a68a7d8-1ced-497c-ab1f-1ce09069a97e · outbound

This paper cites Generating diverse and natural 3d human motions from text.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Generating diverse and natural 3d human motions from text

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.958262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.958262Z digest=sha256:1aee9fb965f60f0e24a1f4bd350ccc3a35d04d0dbc23329dfd774c9763c51014

Observation 500be64b-7553-4e2c-a31d-6231a002e81f · outbound

This paper cites Tm2t: Stochastic and tokenized modeling for the reciprocal genera- tion of 3d human motions and texts.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Tm2t: Stochastic and tokenized modeling for the reciprocal genera- tion of 3d human motions and texts

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.065819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.065819Z digest=sha256:3daad55c875b326ad051b8f0fb49d53fe16164fb0bc38a155c153632b8cb2546

Observation d8dde565-95a5-40ce-947d-fcc4dad2edc5 · outbound

This paper cites Momask: Generative masked modeling of 3d human motions.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Momask: Generative masked modeling of 3d human motions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.129384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.129384Z digest=sha256:64fcb1a1d8bc48f5a18fb46cf8de9153dbc77d054898fa60ba8c05c98a449aef

Observation a4339814-6bea-4c3b-8d20-cde926f6b691 · outbound

This paper cites Snap- mogen: Human motion generation from expressive texts.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Snap- mogen: Human motion generation from expressive texts

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.256735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.256735Z digest=sha256:dc80b2bd856d3b9fbcba0a8c10c01c69b55eade7f0bee0e60abc4259c05db15f

Observation f122d265-da61-4c80-bff0-2f53fc159914 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.354084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.354084Z digest=sha256:13443ebeaa2b14ddbf70629ac3d69ef1c24cb463b18ee4febabfb410d3512a37

Observation f7c17780-9744-4dd2-a530-83f1ffa54ef0 · outbound

This paper cites EgoLM: Multi-Modal Language Model of Egocentric Motions.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation EgoLM: Multi-Modal Language Model of Egocentric Motions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.432236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.432236Z digest=sha256:4f73fd92f0d1c7363a1744941d0294e0aa81ca5611a5c2a7ff826ad8ec8c36a2

Observation 6de5782f-48a2-45c2-a836-74e74da17d26 · outbound

This paper cites Stablemofusion: Towards robust and efficient diffusion-based motion generation framework.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Stablemofusion: Towards robust and efficient diffusion-based motion generation framework

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.512963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.512963Z digest=sha256:8fb19ae438e716c359a6ac12d58494bfd4ce8f4b307737c9bb7f964b188b91fe

Observation ddaf58e9-7c0b-433f-a344-b47e606a44f5 · outbound

This paper cites InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation InterMask: 3D Human Interaction Generation via Collaborative Masked Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.609967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.609967Z digest=sha256:3969653f07cf1ac5dbbc0303817493c3b4ed8e87388e0c3f935770b20a9a8cc1

Observation cde53a3d-066c-407f-8a97-8143423318c6 · outbound

This paper cites Motiongpt: Human motion as a foreign language.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Motiongpt: Human motion as a foreign language

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.672025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.672025Z digest=sha256:0b566174ca8024b410004b278636602e3d24bc395541cc7cf357086ffaae5105

Observation e2200bee-99fe-4219-9f51-c861601a0074 · outbound

This paper cites MotionPCM: Real-Time Motion Synthesis with Phased Consistency Model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation MotionPCM: Real-Time Motion Synthesis with Phased Consistency Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.756769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.756769Z digest=sha256:aea19c6a7938d9750c46d76e9dcc3ea066e0bb544ef13c6d0c7d06de0638b7ef

Observation feee44ab-cf03-46e8-9ba9-8e93b3bc5d04 · outbound

This paper cites Guided motion diffusion for con- trollable human motion synthesis.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Guided motion diffusion for con- trollable human motion synthesis

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.854335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.854335Z digest=sha256:d8247abb52aeddd913b43c4430c8adc6ed08cb07cf57763a87f21ae0829d0eb9

Observation 53e12b1f-158d-4c3f-bd6a-807acd9e4399 · outbound

This paper cites Latent diffusion models with masked autoencoders.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Latent diffusion models with masked autoencoders

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:06.912633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:06.912633Z digest=sha256:4e6e4ad3860eef9702450e03de9393d1071450f9bb0918d1dd243f0bca6b6f03

Observation 8e38ae91-c8b5-4796-b1f1-1596b340fd8f · outbound

This paper cites Repa-e: Unlocking vae for end-to-end tuning with latent diffusion transformers.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Repa-e: Unlocking vae for end-to-end tuning with latent diffusion transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.005968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.005968Z digest=sha256:384a06b084aa2617fcdcbf1037d36b47dbe23df7b1b5bdbcdaf2c845f802976b

Observation 95b75009-b807-4610-b64f-2ec81da94412 · outbound

This paper cites Unimotion: Unifying 3D Human Motion Synthesis and Understanding.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Unimotion: Unifying 3D Human Motion Synthesis and Understanding

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.113970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.113970Z digest=sha256:6b57ae34270c06007e84634bfd10a160a5ec9a898fbf978ffddedb26a2861b43

Observation a6e48cb5-f41c-4369-a8eb-a419b9968de7 · outbound

This paper cites Autoregressive image generation without vector quantization.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Autoregressive image generation without vector quantization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.230200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.230200Z digest=sha256:a9c9a6b365bdc1aa41e75f02ca38cd053898f6a2f12b676d4160880125698c8a

Observation 3522cd83-8a5d-40da-ae3b-8a7f2d01a2db · outbound

This paper cites Intergen: Diffusion-based multi-human motion gener- ation under complex interactions.International Journal of Computer Vision, pages 1–21, 2024.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Intergen: Diffusion-based multi-human motion gener- ation under complex interactions.International Journal of Computer Vision, pages 1–21, 2024

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.355999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.355999Z digest=sha256:7014e81d7eb8e8d7a3513fbf60dce19bf3feacb3891efe1c72d8d72d417e2d44

Observation 44746425-7878-4ece-9a4c-b486e2f69bd3 · outbound

This paper cites Character controllers using motion vaes.ACM Trans.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Character controllers using motion vaes.ACM Trans

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.471224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.471224Z digest=sha256:d2b67c4d4c79e949fb73a56e26210fedd24f3c5c8596fd044f95269a3a980a8c

Observation fd85416b-08b0-41ab-b5ea-7b50414bd5c9 · outbound

This paper cites GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.533226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.533226Z digest=sha256:2dc8e9612220be8434f8a2cbea23a47ef5a8675788fa6203fdf91a0e15738e02

Observation 552b237c-e8f4-4579-8272-052a1a95a16a · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.623874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.623874Z digest=sha256:30b803b058c4bfacc84711500bd300ab282fe529bd1c63b58e4377099202bb3d

Observation ab72fc35-ef71-49b6-840d-5f0e43f29fca · outbound

This paper cites an unresolved cited work.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.696711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.696711Z digest=sha256:d1fe067fd43d46c73f483e967f8063d485dc42b58e2d8d5b2b7781c5f0ee93dc

Observation cf7c807b-b1d2-433f-b7ea-7e2fc8184453 · outbound

This paper cites Perpetual humanoid control for real-time simulated avatars.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Perpetual humanoid control for real-time simulated avatars

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.777266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.777266Z digest=sha256:af8205c1f673652d4361cc178d16be31ab70a3a9ccd8e437f3d64fefd93a7dc0

Observation c9eee30f-b0d2-46d0-a414-e45cd0cb92e9 · outbound

This paper cites Troje, Ger- ard Pons-Moll, and Michael J.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Troje, Ger- ard Pons-Moll, and Michael J

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.846864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.846864Z digest=sha256:4582ac5e98a88ce42f04afa86de608c779471112cb5f36e30e1d4fdce9200881

Observation 74e3c396-5008-4c7b-a42e-5765a4f19a2c · outbound

This paper cites Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:07.941088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:07.941088Z digest=sha256:5cbc8c602eb0c9170f78e2ff3e52dabcbd38db85fa17f4885b8b499d86f9c5c3

Observation 17d8d5b7-6ac0-4414-845d-de473cd02d6c · outbound

This paper cites Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Rethinking Diffusion for Text-Driven Human Motion Generation: Redundant Representations, Evaluation, and Masked Autoregression

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.027036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.027036Z digest=sha256:74fab6fc381d203273fe70fcb6a518e6954a63cef659c354a98b4d39e9fb9535

Observation 796af298-a87d-4a84-8c22-9e9e49074bb6 · outbound

This paper cites Absolute Coordinates Make Motion Generation Easy.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Absolute Coordinates Make Motion Generation Easy

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.124968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.124968Z digest=sha256:a401129ba9fc48c6690a2c772c99e5d25fefaf9c545adb1ada6932abf0b6d265

Observation 5da4460c-2afc-400e-821f-a517d3798f98 · outbound

This paper cites Semantic-vae: Semantic- alignment latent representation for better speech synthesis.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Semantic-vae: Semantic- alignment latent representation for better speech synthesis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.228130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.228130Z digest=sha256:d7997981842aff8e53cab6b8641d86a47e9e9fff7b6662db998e7a72417d94cd

Observation 4c7550f9-5783-497e-a054-e914a4f0927c · outbound

This paper cites Black, and Gül Varol.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Black, and Gül Varol

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.285396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.285396Z digest=sha256:1eda1514d9d0be0dd018d0a45926d5b88045aaeb627c70cd04bad1666173e8fe

Observation 133c7278-3e89-46ec-9773-ac35b35361ae · outbound

This paper cites Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.385589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.385589Z digest=sha256:971c4c154cb4225b87157b9b64f9d5fa5377451557b50782a2c1f055332a7b8e

Observation f72de259-028d-470a-af06-d99a86ed81f6 · outbound

This paper cites Bamm: bidirectional autoregressive motion model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Bamm: bidirectional autoregressive motion model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.513645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.513645Z digest=sha256:e22dd887dc27582d853cea60ab276e69c4fdd4d1d9fa27c96f0e792098c635d2

Observation 9a75b0d5-58a7-4464-8572-92806bf6c917 · outbound

This paper cites Mmm: Generative masked motion model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mmm: Generative masked motion model

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.612633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.612633Z digest=sha256:8dbedbdf6eab73b13b4608906b349d5550c5085b849abbe90f8adc6e67d8e5d0

Observation ebe1bd90-062c-4c69-979e-61c09b980d85 · outbound

This paper cites Maskcon- trol: Spatio-temporal control for masked motion synthesis.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Maskcon- trol: Spatio-temporal control for masked motion synthesis

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.742197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.742197Z digest=sha256:8f618e429bd1051c974ab8dd7a46d5098d69eef96837eab55465ff2b77579d1c

Observation 2606617c-b3ae-46e0-81ca-8e3e3bdf5c32 · outbound

This paper cites The kit motion-language dataset.Big data, 4(4):236–252,.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation The kit motion-language dataset.Big data, 4(4):236–252,

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:08.844155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:08.844155Z digest=sha256:2831ee208b6239c9cb2cb5848a3f496b0eb2d095805ad421c15d4267d69205be

Observation fd5d865c-8258-4930-b214-32c2363e92f6 · outbound

This paper cites Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou, Alejandra Quiros-Ramirez, and Michael J.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou, Alejandra Quiros-Ramirez, and Michael J

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.011842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.011842Z digest=sha256:3b7740f5653ae66b478010e826878e960c6c0593650a031ec42c1e24e5d8d095

Observation 443da5bf-369d-46d8-92ad-ea50d06d5ab3 · outbound

This paper cites Learning transferable visual models from natural language supervision, 2021.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Learning transferable visual models from natural language supervision, 2021

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.104171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.104171Z digest=sha256:3d4972fdceca81064c559b12d183c8a11a114e6ee4ca19c70b9f63f66ce5ab9b

Observation cccad4eb-ed1e-40cd-bf17-405749bfe821 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Exploring the limits of transfer learning with a unified text-to-text transformer.Journal of machine learning research, 21(140):1–67, 2020

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.180075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.180075Z digest=sha256:7c1e3cfa144f83a72127a876e5125f66fdf04f5bfe1b79ca057f69cab7ee1444

Observation 351a0577-5a89-4a14-9a8d-33f7dcc7abc4 · outbound

This paper cites Humor: 3d human motion model for robust pose estimation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Humor: 3d human motion model for robust pose estimation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.254750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.254750Z digest=sha256:0b7418bcf496a1dd1fa9da54e1c8f4b992415b8e34adc946b8ddf64d55bd45c2

Observation ef994afd-b8c7-4189-a9a1-92b80667de12 · outbound

This paper cites Priormdm: Human motion diffusion as a generative prior.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Priormdm: Human motion diffusion as a generative prior

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.316923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.316923Z digest=sha256:bf72b1da9fddda2e200dcfc5fa03b64d64b9fbd6c5682889976b841281ddb4c2

Observation 2119b3ff-3e0d-4577-b969-5410d93d366d · outbound

This paper cites Human motion diffusion model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Human motion diffusion model

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.404757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.404757Z digest=sha256:b947f6995d171bbe4bea68389dc57fa40ed7529d67eca776686a86276301853b

Observation 8148a499-f86a-476f-97b3-4daee30c879a · outbound

This paper cites Mujoco: A physics engine for model-based control.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mujoco: A physics engine for model-based control

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.534994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.534994Z digest=sha256:4caf00bb33a30f23523d0c6355cbdc4aad0e1121a22176c25337799c1b739502

Observation 1315e708-fc83-42d8-b715-6037dbe75803 · outbound

This paper cites Autoregressive motion generation with gaussian mixture-guided latent sampling.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Autoregressive motion generation with gaussian mixture-guided latent sampling

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.600789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.600789Z digest=sha256:64eba20136cff84b96b43be8e23bac42dcbee89568b144d930d2d4ec17a0c01a

Observation e7494b57-da98-4e15-96c8-d00934cc81da · outbound

This paper cites Attention is all you need.Advances in neural information processing systems, 30, 2017.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Attention is all you need.Advances in neural information processing systems, 30, 2017

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.709734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.709734Z digest=sha256:b5a5e959bda2f84782a23970c9d3c71d5c326552b685f236eeb9d2162b2eab7b

Observation c006302d-b6da-4a22-a3f8-d2390c3d8a9f · outbound

This paper cites What is the best automated metric for text to motion generation? InSIGGRAPH Asia 2023 Conference Papers, pages 1–11, 2023.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation What is the best automated metric for text to motion generation? InSIGGRAPH Asia 2023 Conference Papers, pages 1–11, 2023

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.844393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.844393Z digest=sha256:2404ffe2f2037958e3865f219052fdb6e695d019f45c0e2fa4ea78ea4205c6e5

Observation 9b9637f9-3b59-48e0-9140-4da8f0f7a35d · outbound

This paper cites Tlcontrol: Trajectory and language control for human motion synthesis.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Tlcontrol: Trajectory and language control for human motion synthesis

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:09.967630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:09.967630Z digest=sha256:fc756b5020bec6ec8f10308569929d60334c4658d703ee1d7bb09c25800cd044

Observation d3c2e867-fc0f-4ec8-9df0-8c25365c9d6b · outbound

This paper cites Aligning motion generation with human perceptions.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Aligning motion generation with human perceptions

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.066080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.066080Z digest=sha256:3f5950239bd3f21a461abd4cf6c9b462097fa450234f497d39746cb20eee56cc

Observation 1097fc41-7e0e-4758-b198-c2e45bdc75e5 · outbound

This paper cites Mo- tiondreamer: One-to-many motion synthesis with localized generative masked transformer.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mo- tiondreamer: One-to-many motion synthesis with localized generative masked transformer

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.181861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.181861Z digest=sha256:b6e105013e6de4d8acb1ddf7aec2cfb370b0d642055e2bd223b36c183af0ab93

Observation 7b53a770-e00f-4bf4-b084-6ab42ebefae7 · outbound

This paper cites Representation entanglement for generation: Training diffusion transformers is much easier than you think.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Representation entanglement for generation: Training diffusion transformers is much easier than you think

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.278570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.278570Z digest=sha256:de1467b5805fb44b7aaa0e980fe0d858f82f201785ec96a583cb71bb3122c64f

Observation a55b684c-4ee6-4cc2-a8fe-b443a4711d4c · outbound

This paper cites MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.411863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.411863Z digest=sha256:9fc05472836ede06dd59a999f8c4a6a951d27d37290edc8fdf7ebddd66f3ed1f

Observation 8890e7db-9003-46b7-9e25-0db9230f3e9d · outbound

This paper cites Representa- tion alignment for generation: Training diffusion transform- ers is easier than you think.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Representa- tion alignment for generation: Training diffusion transform- ers is easier than you think

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.518066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.518066Z digest=sha256:0f18f8e84e8f1dc89c1289aa1d5cf6a58b22c8b513968efd87fd7c32f0979b34

Observation ba0271bc-a01a-4b74-9025-712dbe5260df · outbound

This paper cites Geometric Neural Distance Fields for Learning Human Motion Priors.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Geometric Neural Distance Fields for Learning Human Motion Priors

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.602170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.602170Z digest=sha256:80d26f874f0b2258436360b80c296afb4147c791a98039d55d90bed7729a77c5

Observation fb651120-b876-4291-ac4c-e1d42fd3fbbd · outbound

This paper cites Mogents: Motion generation based on spatial-temporal joint modeling.Advances in Neural Information Processing Sys- tems, 37:130739–130763, 2024.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mogents: Motion generation based on spatial-temporal joint modeling.Advances in Neural Information Processing Sys- tems, 37:130739–130763, 2024

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.700190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.700190Z digest=sha256:4efc60f6709fca87da800037ab5347cea8c616e876eb2008989ad944859a42af

Observation 9de634a0-f34b-4b62-a183-f32ececbf7e4 · outbound

This paper cites T2m-gpt: Generating human motion from textual descriptions with discrete representations.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation T2m-gpt: Generating human motion from textual descriptions with discrete representations

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:10.886656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:10.886656Z digest=sha256:eb4733b71a104ae314cc1651c757f6930384569d7fe90f6e127ce8a973a963ca

Observation 8fe4e2a9-55b4-4777-8526-eaed33bad137 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.014828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.014828Z digest=sha256:f3485ba1b310bff07973d66c89a562123421e82df98d5eb4d92cd916d577b72e

Observation a97f50a1-a42e-4b24-b107-1085b065e2a1 · outbound

This paper cites ReMoDiffuse: Retrieval-Augmented Motion Diffusion Model.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation ReMoDiffuse: Retrieval-Augmented Motion Diffusion Model

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.054177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.054177Z digest=sha256:938df180044a39bb7bba0c2f9ca57238ba5bfa1a71acee61db131e5a2bf6fd92

Observation 1cfd3a79-cb74-4a3d-81e4-acec8474cf5b · outbound

This paper cites Finemogen: Fine-grained spatio- temporal motion generation and editing.NeurIPS, 36, 2024.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Finemogen: Fine-grained spatio- temporal motion generation and editing.NeurIPS, 36, 2024

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.116136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.116136Z digest=sha256:3fa36c8fdecd81cb05b505fb5647c8dadbecfa7d28793461117168972bf88893

Observation 9ea33396-92b5-485b-b34d-1c241feaf3c4 · outbound

This paper cites Large motion model for unified multi-modal motion generation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Large motion model for unified multi-modal motion generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.172665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.172665Z digest=sha256:56f38ef52e379a85dd3bc9e764c3dd4560c9931d25b2e1879c02cd5633a5ba39

Observation 87a09bf8-8439-4e02-ae09-ea8416f9e804 · outbound

This paper cites Kinmo: Kinematic-aware human motion understanding and generation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Kinmo: Kinematic-aware human motion understanding and generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.237336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.237336Z digest=sha256:815eb55979f849c87e46b8c8ea36e3ee003dae52a22c45aec7150e98b3453d9d

Observation 34014912-a6bf-4c13-95cd-f4183c414b35 · outbound

This paper cites Flashmo: Geometric interpolants and frequency-aware sparsity for scalable efficient motion gen- eration.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Flashmo: Geometric interpolants and frequency-aware sparsity for scalable efficient motion gen- eration

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.352010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.352010Z digest=sha256:6d559b582f172b35afe97bccfbe6e07135b18044f5faa621eee65bb9956e8cb1

Observation bbfa6573-e3f1-4537-9e5f-b6630ff532f4 · outbound

This paper cites Motion Mamba: Efficient and Long Sequence Motion Generation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Motion Mamba: Efficient and Long Sequence Motion Generation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.410248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.410248Z digest=sha256:a662c341d623ad985a01f1bed5a39eedbde80fb409d101d17be4762a9b0561ba

Observation d5d8f2a8-61a7-427d-a978-0f419e11c916 · outbound

This paper cites A diffusion-based autoregressive motion model for real-time text-driven mo- tion control.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation A diffusion-based autoregressive motion model for real-time text-driven mo- tion control

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.516326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.516326Z digest=sha256:61714f6adf9c96d7b67d85f61c6480f54190926dcb597826a1ca272869cd3540

Observation b048a5b9-be58-4796-85b9-c4ed3c0ae776 · outbound

This paper cites EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.562769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.562769Z digest=sha256:5545dc62c9b95f79c826b931fe54685ed4c6d7fcc2adf04a143872e415218f1b

Observation 9129bf4d-fead-4b3a-8a6d-242499ec836c · outbound

This paper cites Mo- tiongpt3: Human motion as a second modality.arXiv preprint arXiv:2506.24086, 2025.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Mo- tiongpt3: Human motion as a second modality.arXiv preprint arXiv:2506.24086, 2025

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.635553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.635553Z digest=sha256:c393599358238af56e146a035cdbfc3b3a9207cbe5c1ea82bb2278b3ae0e604f

Observation 38dd7408-f626-4d77-b9dd-a660c873f322 · outbound

This paper cites an unresolved cited work.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Unresolved cited work

Reference 77

Resolution
malformed identifier
no resolver link, observed 2026-08-03T16:24:11.725291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.725291Z digest=sha256:8cf566f71c33f73e553ba2c5e92af1e45bdd1ce35f1b89fc94c9515df5e5e776

Observation ca55d0ee-b4c0-4293-a769-3d76d65f0b9d · outbound

This paper cites 5 reports an ablation over dif- ferent numbers of text adapter layers.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation 5 reports an ablation over dif- ferent numbers of text adapter layers

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:11.789786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:11.789786Z digest=sha256:9f3b54812f5692395f034b41942a1724e2c923090dd5eb1b0d4a22c2ca441449

Pith citing papers

Observation 6d8e67f4-e507-4e5d-b646-23a855c64a9e · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:25.543486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:6622ac98d868c035350b090479f1c37e5310dacf7156778c91a866093a48f084

Observation 32d94413-4f88-456e-923e-8fff1acbc227 · inbound

Exploring Motion-Language Alignment for Text-driven Motion Generation cites this paper.

Exploring Motion-Language Alignment for Text-driven Motion Generation MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:25.543486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:52:17.036285Z digest=sha256:b108afd6c74d41dc750008e5d792c9f4e81b942cb113736225951d88bf2a5f0b

Observation d62ee051-4ed9-4f31-be05-1e29b96d0eef · inbound

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-based Humanoid Control cites this paper.

SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-based Humanoid Control MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-07-14T03:21:25.543486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-25T02:38:14.542361Z digest=sha256:2beb4d36f3ba97475b36e052f3704949db163e72e02dfed436e2116acc24d2ad