Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:00:10.571415Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 5 inbound Pith citation observations for arXiv:2412.07320.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:00:10.571415Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:19:20.685328Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T11:51:20.243250Z
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 61c33c22-48d6-46cb-a7d1-a16b9a71e2da · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edfc546a-ef57-4ca7-a15f-6bd52badd79e · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Lan- guage2pose: Natural language grounded pose forecasting
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d8cfe53a-c757-4f2d-9681-ce2c0d44af8c · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Teach: Temporal action composition for 3d hu- mans
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c148009a-93c8-436d-a8bf-c6ce7ae2830f · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Black, and Gül Varol
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 39c2907e-97ea-4dcd-a6af-45961dfbb8c9 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Is space-time attention all you need for video understanding? In Proceedings of the International Conference on Machine Learning (ICML), 2021
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a28ee2ca-3954-4bf7-8446-21d104855ea0 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Executing your commands via motion diffusion in latent space
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 90ebcefb-4bc7-4367-b8d9-dee6fb004806 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Synthesis of compositional animations from textual descriptions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 49ff24c6-e9fa-438e-bc9f-b2d596e77a9b · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Generating diverse and natural 3d human motions from text
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8a5d14a6-dfe0-438c-b474-2a6e28343603 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Generating diverse and natural 3d human motions from text
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 333f7d69-1c8d-4870-84da-04d665e241a6 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Tm2t: Stochastic and tokenized modeling for the reciprocal gener- ation of 3d human motions and texts
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fac05a42-127d-45b3-80aa-d334ecc73e0f · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Momask: Generative masked modeling of 3d human motions
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b369e979-c33e-4fa0-a6cf-66853cc81893 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Como: Controllable motion generation through language guided pose code editing, 2024
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e3d6f5a9-4afa-4b98-b874-45a64ad1f235 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Motiongpt: Human motion as a foreign language
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2f6ec7da-fb1f-48a1-8e56-621faeb1751e · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Motionchain: Conversational motion controllers via multimodal prompts
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c4e06e09-2702-4b59-98ba-a70402e57f1b · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Optimizing Diffusion Noise Can Serve As Universal Motion Priors
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1c06346-6654-44ce-9dd7-9c2d71bec522 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Guided motion diffusion for con- trollable human motion synthesis
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1da931b3-a7ab-4d06-b902-71d2798a631e · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Mvbench: A comprehensive multi-modal video understand- ing benchmark
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6f8dac3-1ef8-4191-bb81-cd1b73a6eb75 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Generating animated videos of human activities from natural language descriptions
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a03850cc-81e5-41b0-9eaa-38e7a74a8078 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Rouge: A package for automatic evaluation of summaries
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 141843eb-6f4f-4ee5-a6e8-9b539e19c3b9 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Generation of Complex 3D Human Motion by Temporal and Spatial Composition of Diffusion Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 87b3f4a5-71d4-4044-8bad-15e66f26c1d9 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Bleu: a method for automatic evaluation of machine translation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 955ab591-3b6f-49ba-9c9f-40cae5f4e8dd · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Temos: Generating diverse human motions from textual descriptions
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0be91046-a35e-442e-b02a-bc918f5739ad · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Mmm: Generative masked motion model
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7331ff43-667a-44d7-91eb-75745c59761f · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Bamm: bidirectional autoregressive motion model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2ebef19e-6f6b-4134-86f2-465e88160b9a · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Learning a bidirectional mapping between human whole-body motion and natural language using deep recurrent neural net- works
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation cd27aeae-af2b-471f-afd2-69677ccf29a5 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Generat- ing diverse high-fidelity images with vq-vae-2
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 03d927a9-4c31-4afc-b252-c7c2d4717aa6 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Human motion diffusion as a generative prior
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a8ae3c8b-7fed-48ea-ac4b-681386c292f8 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents MotionCLIP: Exposing Human Motion Generation to CLIP Space
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9978507-2cc8-4a14-a4c4-e7bb11544156 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Human motion diffusion model
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6be87e95-b4ba-42c0-b601-d452ede9848a · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Nvae: A deep hierarchical variational autoencoder
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2ed343e7-210d-4431-ae9e-eb3e99d7cba9 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Neural discrete representation learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation afd2a95d-a97a-4d31-8e10-572485c0f37a · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Neural discrete representation learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ac324ba9-a799-4086-bdbc-b85c0fae5c2f · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Cider: Consensus-based image description evalu- ation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 239b5771-0720-455a-b6c0-132e85953186 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Fg-t2m: Fine-grained text-driven human motion generation via diffusion model
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 27900b70-bcb0-4693-8446-9af75264d895 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents InternVideo2: Scaling Foundation Models for Multimodal Video Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58033f15-761b-4826-b236-ee34b50c189d · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Chain-of- thought prompting elicits reasoning in large language models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8f2c2547-68c1-4278-b4ce-95f4de278728 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Motion-agent: A conversational framework for human motion generation with llms, 2024
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dacbca06-3457-4f17-8320-4535ba01713f · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Omnicontrol: Control any joint at any time for human motion generation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 03658c26-6736-4b8a-b14d-4e197e7c89e5 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Attribute2image: Conditional image generation from visual attributes
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d527f119-4457-4763-8c83-52fa0b28fced · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Soundstream: An end-to- end neural audio codec
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 88aa3ae7-21fc-48a9-938d-84332972f880 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents T2m-gpt: Generating human motion from textual descriptions with discrete representations
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8b2cf08d-c217-4932-98ed-0c751abd0167 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85e193bd-00fb-4747-8761-a01976b33fb2 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents ReMoDiffuse: Retrieval-Augmented Motion Diffusion Model
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b02d67c-90d1-485f-bd03-ee19956c6023 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents Finemogen: Fine-grained spatio- temporal motion generation and editing
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7f6a5162-2dac-4a3e-8aaf-e61805966484 · outbound
CoMA: Compositional Human Motion Generation with Multi-modal Agents BERTScore: Evaluating Text Generation with BERT
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38885931-82a8-49b5-9d59-d0ff84385ed7 · inbound
Absolute Coordinates Make Motion Generation Easy CoMA: Compositional Human Motion Generation with Multi-modal Agents
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d162210-87f4-4338-b20a-c50fb13eafb4 · inbound
Multi-Modal Manipulation via Multi-Modal Policy Consensus CoMA: Compositional Human Motion Generation with Multi-modal Agents
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 728fbaf6-d3bd-44f2-a166-9c0415fd05ea · inbound
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning CoMA: Compositional Human Motion Generation with Multi-modal Agents
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bd5556f-4023-411f-9219-28f0c3fa9a5c · inbound
LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents CoMA: Compositional Human Motion Generation with Multi-modal Agents
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c6481434-0f9e-4afa-9cc2-cdf53114e0fb · inbound
FineMoLA: Towards Fine-Grained Motion-Language Alignment from Clip-Level Supervision CoMA: Compositional Human Motion Generation with Multi-modal Agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.