Pith. sign in

Paper Citation Record · LEDGER

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning

As of 6 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 3 inbound Pith citation observations for arXiv:2511.18209.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.18209 v3

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T19:45:11.576808Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T19:39:05.211169Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T13:38:19.013877Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact3
  • verified fuzzy34
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 127df5a5-95e4-4cb7-a511-634946184d36 · outbound

This paper cites Motiondiffuse: Text-driven human motion generation with diffusion model.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Motiondiffuse: Text-driven human motion generation with diffusion model

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.305262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:33d40c39a49cda4994e3b32cf49206f33f5427eb7e92780a48b8565b1bbe6889

Observation bd746331-99bb-41ec-8b1d-50bd3518caf2 · outbound

This paper cites Revideo: Remake a video with motion and content control.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Revideo: Remake a video with motion and content control

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.312966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:3b61c6f809d5e57c909ad4307dc75e8723abe8535e37f4ef38a5b76d6a344daa

Observation fa7c81a6-ece7-4192-bad2-17e362a097ac · outbound

This paper cites Continuous, subject-specific attribute control in t2i models by identifying semantic directions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Continuous, subject-specific attribute control in t2i models by identifying semantic directions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.307708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:cc5279e666b461f08cbb80bb362597fdf4263ccf562aa4247749a503148da37b

Observation 847ef798-0c6d-45e4-bf0a-13f1e318b94b · outbound

This paper cites Fg-t2m++: Llms-augmented fine-grained text driven human motion generation.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Fg-t2m++: Llms-augmented fine-grained text driven human motion generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.302594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:5ebb68980dfd74f4fc9b22bc209b07519f814e5fdc2e8baffd134b47ef4ade79

Observation 2a03cc19-5955-485c-a1bf-afadb0028650 · outbound

This paper cites Vmc: Video motion customization using temporal attention adaption for text-to-video diffusion models.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Vmc: Video motion customization using temporal attention adaption for text-to-video diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.230843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:972f1ed223e84cc39f230feca04f1ae75e01d4a6661880b8220f5c4a39e07528

Observation bd68c92a-b5c1-4bef-94db-ceb38dbf39ed · outbound

This paper cites Intergen: Diffusion- based multi-human motion generation under complex interactions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Intergen: Diffusion- based multi-human motion generation under complex interactions

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.258410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:5f7dc023adc0a18937ac058849d8095bedb757233ba5ef0a4f4a9b69fb141f84

Observation acfd8d6a-966e-48d7-a268-710eef2c6a73 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning DINOv2: Learning Robust Visual Features without Supervision

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-21T19:45:33.190595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:069c6442e0ffb01bd07a6c7983d88918aa6c9831456f2332d43be84032f34d10

Observation 81864ecb-e436-4803-b771-7dd4f2ae94e4 · outbound

This paper cites Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-21T19:45:33.203525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:2c2b25e536a80a2a3a240c61e16ea298ade7428214e829449edb0fc991f7b619

Observation c197634f-2ff6-47ff-a06e-f10db333adc6 · outbound

This paper cites Parco: Part-coordinating text-to-motion synthesis.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Parco: Part-coordinating text-to-motion synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.237147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:032de0c727b35e730cac0e831419fefa67a43cc11bc497f389b70e3d8d632a67

Observation b140d07d-8aca-424f-86eb-af84b6064d30 · outbound

This paper cites Exploring text-to-motion generation with human preference.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Exploring text-to-motion generation with human preference

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.269666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:a9552392e7b122f4ab6b17a45865842f159f1b4c1a915257b8add8b02c842f83

Observation 2e2b77bd-80a8-4587-b708-86e2c332f0e1 · outbound

This paper cites Learning Variational Motion Prior for Video-based Motion Capture.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Learning Variational Motion Prior for Video-based Motion Capture

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:45:33.209225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:e04e6eea57007a6b2a31cf83c9c0703826f93c109d45ecab18679c71dc8895a6

Observation 57c3f223-258a-491d-bf96-e07cbbaefa28 · outbound

This paper cites Dancecamera3d: 3d camera movement synthesis with music and dance.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Dancecamera3d: 3d camera movement synthesis with music and dance

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.233889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:1ef966c370e9c452c1a95785f0b80e3298ac06c6beb532c34750c276babbe7d1

Observation 83dfe628-0237-4fd2-9b69-19a45350be06 · outbound

This paper cites Bidirectional autoregessive diffusion model for dance generation.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Bidirectional autoregessive diffusion model for dance generation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.275061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:babf5070a030757b01e3b4eeb9d6cc163108519060f509dec1f75de3ad9e6c8d

Observation a9b0aacd-2c76-42e2-989b-828a43808f2a · outbound

This paper cites Modi: Unconditional motion synthesis from diverse data.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Modi: Unconditional motion synthesis from diverse data

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.290058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:0e9f320e72b112c708169f7a8b2a7ca09333881176dbfddd1d243f018bcbdcc7

Observation 97a60cd3-5873-456d-b97a-a33164c062c2 · outbound

This paper cites Fg-t2m: Fine-grained text-driven human motion generation via diffusion model.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Fg-t2m: Fine-grained text-driven human motion generation via diffusion model

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.297643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:d4371fa3d835df7a28ee3ef14aa73dbf9eb4d8fc904ab5c532062052167e3cf0

Observation 6c424c4e-deb0-48bc-9024-c06ea45603f8 · outbound

This paper cites Synthesis of compositional animations from textual descriptions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Synthesis of compositional animations from textual descriptions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.255921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:404fa5ed669f2013ef0d51c47c554d7c859d07e22616fc7a25a57d5f589f105a

Observation 96560a3b-594a-4b3a-8744-a9f34d9ba21e · outbound

This paper cites Motiongpt: Human motion as a foreign language.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Motiongpt: Human motion as a foreign language

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.220724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:15e58c9a5334c4ada8e8d0d209f00149be1cf75bf9c04f7a09bda82451c36eb1

Observation 0b4e5988-df41-4c86-97c0-350b1ba56409 · outbound

This paper cites Momask: Generative masked modeling of 3d human motions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Momask: Generative masked modeling of 3d human motions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.272684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:b59c9b0f67a6d434bf49fd92c46a15314a64c21f954f5c56b05323c5306c3658

Observation ed74cbea-2f2a-484b-a4cb-42d5224a902e · outbound

This paper cites Smpler-x: Scaling up expressive human pose and shape estimation.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Smpler-x: Scaling up expressive human pose and shape estimation

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.295163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:b23b30299e7e2bc11198ee8bb767b3d765fb5a790a049c36621e3e14f21fffbb

Observation 3493ad64-093d-4f46-a9cc-be7177b83acd · outbound

This paper cites Smpl: A skinned multi-person linear model.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Smpl: A skinned multi-person linear model

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.280970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:410117e393ad4797004e8bdda43a3e5d2cf30f2f1c45b5e8faf402cce107dd64

Observation 8b77342f-04f1-4ffa-884f-62f3a97e063e · outbound

This paper cites Disentangled clothed avatar generation from text descriptions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Disentangled clothed avatar generation from text descriptions

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.252584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:e30746453050ab79fa4502a6db3e52a1e94797ffa5162e7563fb1f850069d71c

Observation 78d1e966-45ea-4374-8f93-7a094bbe041a · outbound

This paper cites Sesdf: Self-evolved signed distance field for implicit 3d clothed human reconstruction.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Sesdf: Self-evolved signed distance field for implicit 3d clothed human reconstruction

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.263974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:521d99a69a6a64f0cefbc47ee24f26d2229f35f6a318423279c3c7b7c777d990

Observation 5725e924-77bd-4be4-b7bd-9208159f83de · outbound

This paper cites Generating diverse and natural 3d human motions from text.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Generating diverse and natural 3d human motions from text

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.292581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:705a71ce3491b8e428376ea695792a504900fad6adb11b39d94cf3d68eabcea0

Observation 333031e6-ddc0-42ae-beb9-99c317ab1da2 · outbound

This paper cites Deepphase: Periodic autoencoders for learning motion phase manifolds.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Deepphase: Periodic autoencoders for learning motion phase manifolds

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.250000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:c7e053f956891e0ba0ab49084fe358791b7a57a8b73c235f76c7290c7f7382aa

Observation 7a52f8ae-9b87-4d74-b91b-934ad1e581b8 · outbound

This paper cites Executing your commands via motion diffusion in latent space.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Executing your commands via motion diffusion in latent space

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.261467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:7dcab8b6f38204446cc30e7b4bb45e40923e7951bd9badb1c225ea1848b59707

Observation 052605d9-74d3-4699-8a39-5f170ce20efe · outbound

This paper cites Mcre: Multimodal conditional representation and editing for text-motion generation.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Mcre: Multimodal conditional representation and editing for text-motion generation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.243611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:8addf49996f54b04a94a1999431063434253b6ecfced6d3f6aa76d4a3dfac236

Observation 02e25b98-399c-4902-9b68-5d0719d99a6a · outbound

This paper cites Tl- control: Trajectory and language control for human motion synthesis.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Tl- control: Trajectory and language control for human motion synthesis

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.246661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:bcbfddef4587eebd15dcc6edc8acad540f5a5c968fbad1f9a521bd401107a393

Observation 98bc8131-42cc-41ec-8f5a-e6fcda54e3ef · outbound

This paper cites Rethinking the spa- tial inconsistency in classifier-free diffusion guidance.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Rethinking the spa- tial inconsistency in classifier-free diffusion guidance

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.217777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:73e5a41ac9fb2f86311de5be94a9feaf348cd3cda692677d219aaac25fa02a42

Observation fa1378a6-5c1f-40e5-ba6f-23b77bec60b7 · outbound

This paper cites Tcfg: Tangential damping classifier-free guidance.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Tcfg: Tangential damping classifier-free guidance

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.284601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:43f392364ab47c05e9604acb288b6614f2220a16eb35ac31e45c5b42c0823169

Observation 3ed1c6ec-e0b5-4721-b10c-8cb545e59618 · outbound

This paper cites Videomae v2: Scaling video masked autoencoders with dual masking.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Videomae v2: Scaling video masked autoencoders with dual masking

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.310520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:87a2f901670e5902af17aec2b8e7febb4f9f50fb311191bb8c2343e04b27ae3b

Observation 58c02c2b-e497-482f-86ee-dfe46a181f9b · outbound

This paper cites Guiding a diffusion model with a bad version of itself.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Guiding a diffusion model with a bad version of itself

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.214217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:fc2afca9dc660e36da68bac44f5113f305301cfcffcd5cec1bcb85079c452cd6

Observation 334ac9d2-b5c0-434c-9e47-8f918699360d · outbound

This paper cites Learning transferable visual models from natural language supervision.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Learning transferable visual models from natural language supervision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.277782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:fe33b082ca22f79d79743238957f63c73b8160db4aed933b64e86a9698d68339

Observation 39a03f6b-1f60-4922-b64b-a5f53d7d782a · outbound

This paper cites Facenet: A unified embed- ding for face recognition and clustering.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Facenet: A unified embed- ding for face recognition and clustering

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.300239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:e8be468bd12d609e231c52362b9210db06a1d934056bf05133cc52c3f07dd820

Observation 211ec898-13a4-4c22-ae08-016d597f5d7c · outbound

This paper cites Computational optimal transport: With applications to data science.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Computational optimal transport: With applications to data science

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.240251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:05bb53057695fadae41e892603b9fcaaf1a25c2909bfcc5e21782881c08b0e87

Observation 7ddddd2d-3492-4b37-9760-1b77dff33a5a · outbound

This paper cites Generating diverse and natural 3d human motions from text.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Generating diverse and natural 3d human motions from text

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.287273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:9285dfccdded87de9398aa7823cf18dc319d20a40c943148bd6c15bb992a90a4

Observation 70eb6b22-71c8-4427-83f4-4510a6ecbb19 · outbound

This paper cites Action2motion: Conditioned generation of 3d human motions.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Action2motion: Conditioned generation of 3d human motions

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.224089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:a46ebf6683a3ff2799f68065904439b0c461fe9abae9586b7130ef04b4df0b95

Observation 9e1aa53a-f362-47f2-b009-8374792142fd · outbound

This paper cites Amass: Archive of motion capture as surface shapes.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Amass: Archive of motion capture as surface shapes

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T19:45:33.227118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:fa9fe05259734186275e416a3e5dfef3310dc0c661f2d76024b4437cddd24807

Observation 741bd8c7-cd46-4ba1-9508-aafd36729c46 · outbound

This paper cites Motion capture from internet videos.

MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning Motion capture from internet videos

Reference 38

Resolution
malformed identifier
raw_fallback, observed 2026-05-21T19:45:33.266984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T19:45:11.576808Z digest=sha256:44ae54048ac1a56d32937c1d27bad700979f5faf0dc00a09b0895fc5b69c397f

Pith citing papers

Observation 734c08a4-3362-4298-b2c7-c783446adb18 · inbound

AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro cites this paper.

AnchorRoute: Human Motion Synthesis with Interval-Routed Sparse Contro MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:05:37.711805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T16:14:13.795423Z digest=sha256:106b099931ab95cd26dcd50aabb9b75f50f9da6ae1164577f7962acd8313bd61

Observation e18a68c4-5eb1-4fca-b887-0435bd83f35c · inbound

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars cites this paper.

UMo: Unified Sparse Motion Modeling for Real-Time Co-Speech Avatars MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-06-30T19:45:00.897818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-30T19:39:05.211169Z digest=sha256:cd2847d7a5a0f247da8c89bfb7c71877597adb56905fe316e9e66531e3884121

Observation 27528227-68ec-42de-a421-9c21366ae4d2 · inbound

MoGeFlow: Flowing Through Motion Codebook Geometry for Text-to-Motion Generation cites this paper.

MoGeFlow: Flowing Through Motion Codebook Geometry for Text-to-Motion Generation MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T13:38:19.015130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T07:48:24.383009Z digest=sha256:cc44a3511a989a29cb9f53c366adb6cf6c7a06a2355e02de1e3d475bb8af3a2d