Pith. sign in

Paper Citation Record · LEDGER

Fleximo: Towards Flexible Text-to-Human Motion Video Generation

As of 14 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2411.19459.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19459 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:13:40.746783Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation aa37aba5-55c7-4dd9-a31a-5c6e062c5d5e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.490628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.490628Z digest=sha256:478326af052f97d4b335765d5c33037543c175a15133c13e7c7bf47ad90c5558

Observation 306a3257-29dd-406c-8cd3-31373c77e294 · outbound

This paper cites Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Magicpose: Realistic human poses and facial expressions retargeting with identity-aware diffusion

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.566438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.496525Z digest=sha256:7379abc829b4d579811636c8d37ed0ad1b47144be01cceaff444b43e0c43a4b0

Observation 85175727-acca-46c4-a71b-ac1f02ed6e0c · outbound

This paper cites Videocrafter1: Open diffusion models for high-quality video generation, 2023.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Videocrafter1: Open diffusion models for high-quality video generation, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.548929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.501191Z digest=sha256:d08e007f812f1b35ac4968689c9a827155cabe7f91409cba124b79f23d52834a

Observation 07729790-99de-4d0b-82e3-90bd3ef760d0 · outbound

This paper cites MotionLLM: Understanding Human Behaviors from Human Motions and Videos.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation MotionLLM: Understanding Human Behaviors from Human Motions and Videos

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.506000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.506000Z digest=sha256:0fd3e05515c4a2fbb5bac5bf0fbfaef31992a834bdf2043c3d49259292c47769

Observation a0224a63-38ed-4c7e-b3be-ca156790713c · outbound

This paper cites Executing your commands via motion diffusion in latent space.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Executing your commands via motion diffusion in latent space

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.511123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.511123Z digest=sha256:3898005f67bcdd35accffa6bbfe290efaad71b8be82769f07c8fcce453934bd2

Observation 30dafbfc-6158-4901-a755-c3e07ba38f99 · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.516086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.516086Z digest=sha256:f4da8211507ae166a086dd2106c8491a0c232698bfa28171b1f69e433a0656bf

Observation d5db001f-06ba-4da2-95a7-ed09bb9ec565 · outbound

This paper cites Generating diverse and natural 3d human motions from text.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Generating diverse and natural 3d human motions from text

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.520802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.521478Z digest=sha256:f940f89f7494fe315a87a88f9b61fd469a9ec72a93a1eb52406060eb5ac85714

Observation 35f238d6-8ab1-4ac6-a13b-4728c237111c · outbound

This paper cites Generating diverse and natural 3d human motions from text.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Generating diverse and natural 3d human motions from text

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.504168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.526877Z digest=sha256:eba9d8ac9a59da2929addad5eb8c37c0d43e00e864b46d2edda1acb86c404be9

Observation d1f4abd9-4af8-4eae-98a4-29c44bee6f3e · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilib- rium.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Gans trained by a two time-scale update rule converge to a local nash equilib- rium

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.531947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.531947Z digest=sha256:a61745551a77da2546ce81930575a60e9e66b440ce0dd9d5138e50993338e36f

Observation 4323b666-deb9-408d-96ee-8c9d1e5e353e · outbound

This paper cites Denoising diffu- sion probabilistic models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Denoising diffu- sion probabilistic models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.537129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.537129Z digest=sha256:dfff0caad329ab51ee90993e236c5d149a8243aacceb8f388f931f10c8ebad39

Observation 352d5eae-8fc4-4c5a-98ab-4e28f202ec68 · outbound

This paper cites Image quality metrics: Psnr vs.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Image quality metrics: Psnr vs

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.465326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.542131Z digest=sha256:04601325711903682a1a3953f7a20081b7449449a55e59613d3241522f904d59

Observation 3fcaea52-a3e1-422a-b8ec-8a573cf7fcb8 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Animate anyone: Consistent and controllable image- to-video synthesis for character animation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.547207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.547207Z digest=sha256:9b3ff7ad04d8c878fa8c9e75c31cbc54ff3ff2bda45415ed74680b0bdf78089b

Observation 66a148ff-19f3-4270-a84d-fb56cc3e6f4e · outbound

This paper cites VBench: Com- prehensive benchmark suite for video generative models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation VBench: Com- prehensive benchmark suite for video generative models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.436014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.552596Z digest=sha256:6ad3ca8df5aac4b283de513f1b868061f4faf424d7014fe634071dbb40e4e2a9

Observation 51415097-2d8d-4f04-8d8d-4a1f768a04ff · outbound

This paper cites Learning high fi- delity depths of dressed humans by watching social media dance videos.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Learning high fi- delity depths of dressed humans by watching social media dance videos

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.418841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.558580Z digest=sha256:2f919e587a0f4484a97d98577f2eafb892be93342e88b8afab10202328d7081f

Observation 0a7dcbed-faef-4f16-9566-14794036de09 · outbound

This paper cites Digital image processing.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Digital image processing

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.400503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.563720Z digest=sha256:ff0e9305aac93d57067800eb815d7015ff4b3dcb752013fa0d4723a772fc190e

Observation 1d25c30c-ac2e-487e-bd4d-fe96f12df98a · outbound

This paper cites Motiongpt: Human motion as a foreign lan- guage.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Motiongpt: Human motion as a foreign lan- guage

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.383558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.569750Z digest=sha256:abffa39325c6b606ca897f2febb240d5dfc3ff447d830f181d8576d4b7de883d

Observation c743c967-c44d-4ab0-bce1-0c26cd6deeee · outbound

This paper cites Auto-Encoding Variational Bayes.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Auto-Encoding Variational Bayes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.574807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.574807Z digest=sha256:be5d3cd9f302bbe2fa5857166c83428df23cb3673b8ca5a71e8df4bf68525d5b

Observation 635a1e0d-0bcc-49a2-9b29-6856134bbbd9 · outbound

This paper cites MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation MotionRL: Align Text-to-Motion Generation to Human Preferences with Multi-Reward Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.580329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.580329Z digest=sha256:0c9c9d13fbd359695a3595886bba2da96a9029395be79436272668ff43133127

Observation 8b7b0fe3-09a8-45a0-9c7b-a6c04fa420c9 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.586149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.586149Z digest=sha256:a9094b52d0248e0332a4917f68d347445f64ced44ba665087c860e27e4c6e3d5

Observation e3137eab-588e-410d-8fe4-22a636f1c0c6 · outbound

This paper cites Temos: Generating diverse human motions from textual descriptions.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Temos: Generating diverse human motions from textual descriptions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.591536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.591536Z digest=sha256:14c5c4fc940f300d5396d4767b7e81fc4e6f623a1d6fdc7cd18069e327b15c6e

Observation cbd17b83-ed0c-4bc5-ac1e-51f7ca7be329 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Movie Gen: A Cast of Media Foundation Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.596628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.596628Z digest=sha256:32e9b9d76c9de6ededf2d9ea576fb1a565bcda39fb9bf7e00168c39a6d3c6374

Observation df0364af-f184-47bd-abb4-bfb7d7eb7771 · outbound

This paper cites Improving language understanding by gener- ative pre-training.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Improving language understanding by gener- ative pre-training

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.602806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.602806Z digest=sha256:6ff2a635916341214fce0921c4c836c022fd3fbaa0866fad657b95d716cfdc60

Observation 610265d1-7e15-49f6-8f17-e962004d6f3b · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Learn- ing transferable visual models from natural language super- vision

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.343599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.608126Z digest=sha256:e1013696284ded24f71d1860283e0bce3b101206047fd4e95ab03a749a861975

Observation 099f73fc-c123-481a-8729-e9b283e32030 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation High-resolution image synthesis with latent diffusion models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.613159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.613159Z digest=sha256:abbe4ca418f6ef611f7c32d09a2efd7b1d18435a83f2bf580a91c200769889f7

Observation 8ee44a92-404d-40da-a3ef-885a4514d83e · outbound

This paper cites U- net: Convolutional networks for biomedical image segmen- tation.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation U- net: Convolutional networks for biomedical image segmen- tation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.313799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.617674Z digest=sha256:18da3db8b5d2c96770f8e9f80be545a43bb8d390f91db6ce204a2da1eb09f44a

Observation b4956ba8-fa6f-4f93-9039-bbaa7eedde67 · outbound

This paper cites Denoising Diffusion Implicit Models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Denoising Diffusion Implicit Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.623959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.623959Z digest=sha256:cb9e9dd605bd9cfdd9bfdd1d1e67c0836aa1193d94981dddaca7c963cd0dee30

Observation 8fe6dd56-9acd-417c-b744-5f76035aa2ef · outbound

This paper cites Motionclip: Exposing human motion generation to clip space.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Motionclip: Exposing human motion generation to clip space

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.296313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.629027Z digest=sha256:2f55baf1696dbbf6465bd31c1c531726736c0a634bc90dcae601d65282eca216

Observation dfe17384-989e-4f99-aa52-2f582bd643e2 · outbound

This paper cites Human motion diffu- sion model.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Human motion diffu- sion model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.633810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.633810Z digest=sha256:40250b7a0649ed7443e8f3953bf6b1bef291279ce1074d826cdf0b1046870510

Observation 1eda211a-0435-4395-9a7b-9012e9d5a3e3 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation LLaMA: Open and Efficient Foundation Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.638481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.638481Z digest=sha256:25cc1316e2e4fb98abf4d1542a472dd79e2e3e078ba27464a8a1dc8964fb4b95

Observation 7531c7ab-42c4-4264-bf72-4d052fa1e0d3 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.643288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.643288Z digest=sha256:222be2269ca1e8445fe43f2a6993a5054847aff8ae3323ab44895986b180c363

Observation 71ac6984-57c3-4c31-b2a3-7b57cb4c4fd6 · outbound

This paper cites Neural discrete representation learning.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Neural discrete representation learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.268363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.648803Z digest=sha256:b58f20df5cd46e1384863800ea1d898518b0a6d79806f8ffc1649881019e02f3

Observation 98037765-4d57-4783-b922-55550336bd89 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Disco: Disentangled control for realistic human dance generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.250983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.653923Z digest=sha256:e7d0746d989c326bffd8eed758c61a88efebb9f2878f23318827286e3e856680

Observation fd20ac5c-1ba7-43f4-a49e-379b1d094c9e · outbound

This paper cites MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation MotionGPT-2: A General-Purpose Motion-Language Model for Motion Generation and Understanding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.659063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.659063Z digest=sha256:2a45440a73b911cde84b41f94d2f017461b748a15309ce8f4d0b702c0c3ac17b

Observation 33df348a-e6a5-4dcb-b727-039626c45b6b · outbound

This paper cites HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation HumanVid: Demystifying Training Data for Camera-controllable Human Image Animation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.664845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.664845Z digest=sha256:32268b6930d838b438fd2702008a6fdd5d574e4ceef7880b203cb6a1f6eb4707

Observation 8bb9181f-870d-4545-9828-2e672ae51df0 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.670394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.670394Z digest=sha256:f588a30b99f5514e3504b7a52a42ed8586d1158809eb8d8ea1847a092d88bb56

Observation f3572c70-a08b-4924-98ca-d62dd77bfcd8 · outbound

This paper cites Magicanimate: Temporally consistent human im- age animation using diffusion model.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Magicanimate: Temporally consistent human im- age animation using diffusion model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.676128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.676128Z digest=sha256:7bb9ceb394b07bc5c3763d9ff3eab98525f422ea43baf29163719b0decc2bec2

Observation b6f9a81c-d404-49f7-908b-0989387c194f · outbound

This paper cites Mastering text-to-image dif- fusion: Recaptioning, planning, and generating with multi- modal llms.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Mastering text-to-image dif- fusion: Recaptioning, planning, and generating with multi- modal llms

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.681243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.681243Z digest=sha256:810306077d6015e22fa1e2be7bcb06c69ae4ecf131fda1dbca269257ec8511e1

Observation 78dc247a-d061-4453-a51a-ca5796d4e07f · outbound

This paper cites Effec- tive whole-body pose estimation with two-stages distillation.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Effec- tive whole-body pose estimation with two-stages distillation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.686289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.686289Z digest=sha256:74d3abb5ca78ef78c91a3c8953c02e55d93157dd66a54948fa02f2faba9b3a14

Observation b6c0336b-a910-4c1b-b797-23bc570f50e2 · outbound

This paper cites Generating human motion from textual descrip- tions with discrete representations.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Generating human motion from textual descrip- tions with discrete representations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.188838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.691285Z digest=sha256:460535df310f7213590d798615bd1ae5e4bebd46db4ba961a9a8e6a365f21942

Observation fd0a3502-9bb2-4c97-b219-6ee916ee7cf3 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.696670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.696670Z digest=sha256:f5b38e7f5d8465d57025dc7a1646d9c3bfe35fc126a82ef019629bf2e0659688

Observation cbe6ab5c-e06a-4d22-a0da-2082230bd188 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation The unreasonable effectiveness of deep features as a perceptual metric

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.171000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.702359Z digest=sha256:cc57e56383132430d01fe4738c7f5bb837aeecfbea510bd571265f3d5c881823

Observation e685905d-b4ac-4548-b98d-2af5ae8f7a7b · outbound

This paper cites I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation I2VGen-XL: High-Quality Image-to-Video Synthesis via Cascaded Diffusion Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.708407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.708407Z digest=sha256:b853834ccbfdbc12876d1125fb61873f103cedf79bf2882d478a306d83c4b5d0

Observation ac604c50-b044-487e-8a94-667f7703c171 · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.714663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.714663Z digest=sha256:79ad6bc550c88c1f80e67fe719f1c7932d7d39d7b27d09a4c8401db67102fd4d

Observation 751af8ba-2a72-4a5c-a9cb-cd6a7cb971b5 · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.720090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.720090Z digest=sha256:c116d346332db5d03f3b248b58794c7312f35e16fca50cf13dd9ebf8adb1bdb8

Observation 6a15dd41-6bbb-4c42-ad05-0eef4427a5b5 · outbound

This paper cites Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Champ: Controllable and Consistent Human Image Animation with 3D Parametric Guidance

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T10:13:40.725281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:13:40.725281Z digest=sha256:776e112beb7bf2527c9302251772dea74224ec553bda9e81e9ab34d8eed8fb4d

Observation c12d70c2-b26d-4e82-b9e9-36160d1dd5c4 · outbound

This paper cites an unresolved cited work.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:13:41.154430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.730645Z digest=sha256:0cc99b0197784aa73240333d422d7c3c2392371e781fdf5003a24a793ba27620

Observation 059d3f59-9fa6-4b2b-8c68-2424071bec42 · outbound

This paper cites an unresolved cited work.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:13:41.137854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.736302Z digest=sha256:61b8aa137b133fe62478744e7bdc1012c727a2d3998635a4ab5628200ce3b8b9

Observation 6570606a-56ba-4320-a462-a4119f302c96 · outbound

This paper cites Eight participants evaluate each video across three criteria: video quality, identity preservation, and motion-text alignment.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Eight participants evaluate each video across three criteria: video quality, identity preservation, and motion-text alignment

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:13:41.120246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.741528Z digest=sha256:968690931da7f186847b8d8a7868a90d17b1d5c30b2b52579ed544fb6c5b4fd8

Observation d8e6b706-78f6-4a4c-b1fa-2c6eb3f5b2fd · outbound

This paper cites an unresolved cited work.

Fleximo: Towards Flexible Text-to-Human Motion Video Generation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:13:41.101458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:13:40.746783Z digest=sha256:62e3a7685a0a724247679c8052e9205fc069172c217b27c71754d3f3116511a4

Pith citing papers

No inbound Pith citation observations are available.