Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:07:20.190660Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 0 inbound Pith citation observations for arXiv:2507.23188.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T11:07:20.190660Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c3d0d4c0-abe6-40c1-825e-40047a55204f · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Dual stream relation learning network for image-text retrieval,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d2d2a13-d0a9-4b2e-b8d1-c2da8f8b414c · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space One-shot human motion transfer via occlusion-robust flow prediction and neural texturing,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ceb4654-f82d-4be2-9590-087d0ae274dc · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Ta2v: Text-audio guided video generation,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ac0208f8-bb09-4692-a9a2-d32d9ba81178 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Cross-modal quantization for co-speech gesture generation,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d7739b76-5e27-4f66-bc15-a47e715def06 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Generative adversarial graph convolutional networks for human action synthesis,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 522a19a5-af03-4a68-9383-3fd6bc0db884 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Action-conditioned 3d human motion synthesis with transformer vae,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a865e48a-ab43-4969-8c7d-b0be2125ced4 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Multiact: Long-term 3d human motion generation from multiple action labels,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 49b08ef6-91b7-48c6-ad88-c94983e8c4f1 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Executing your commands via motion diffusion in latent space,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 01d81e7a-032c-4844-8915-e6e7ae299068 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space The kit motion-language dataset,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 095dfe23-4e48-47f8-9e90-2b6f8acdda8e · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Generating diverse and natural 3d human motions from text,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 227eaee2-97e9-449c-a7c2-4ab8269e7e2f · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Human motion diffusion model,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e0e44298-732d-4cc5-b0ba-bed362a7e0a1 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Generating human motion from textual descriptions with discrete representations,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36c58672-c52c-481e-a62c-dc7cb2ad78ac · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Groupdancer: Music to multi-people dance synthesis with style collaboration,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 743646df-9965-40b7-9484-63b14329913c · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Music- driven group choreography,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c150045f-27fa-4875-a74a-f66c9598bc86 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Edge: Editable dance generation from music,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d56ec5b9-7cc4-4c18-b93d-4e9ab369e531 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Pc-dance: Posture- controllable music-driven dance synthesis,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0abead2e-3c61-4f25-aff1-009e29565489 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Couch: Towards controllable human-chair interactions,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a509508f-d436-49c5-983f-9de104d75d41 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Goal: Generating 4d whole-body motion for hand-object grasping,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86592244-b176-4903-be80-4943ac117cde · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Human motion generation: A survey,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e0bc659-6b8e-48e9-878f-c6b549566402 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Phase-functioned neural networks for character control,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0b5a84df-0c83-4c98-bc86-5774b7ba02bb · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Learned motion matching,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d0429251-a731-407a-bc5d-e9b12eb55dce · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Tmr: Text-to-motion retrieval using contrastive 3d human motion synthesis,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46e64a69-2467-4b76-8040-8d643eca10fd · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Tri-modal motion retrieval by learning a joint embedding space,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation af304b00-2e5b-41c4-a3b6-7ff024a44696 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space GPT-4 Technical Report
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbaf0f20-03b5-457b-ac6b-7415c8d21989 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Wavlm: Large-scale self-supervised pre- training for full stack speech processing,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 980f1468-0f6e-48d7-bd95-59f812a4ea20 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Better speech synthesis through scaling
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e4cc8ae-f314-40d2-8d38-0f7fa9a0b957 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Motionclip: Exposing human motion generation to clip space,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d148ccc-0ace-447d-a638-ecab485f4080 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Auto-encoding variational bayes,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9f74b422-15a8-4075-b93f-212138c3eea6 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Temos: Generating diverse human motions from textual descriptions,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1d28c1c-9ff6-4dae-8c36-96221fdfd196 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Tm2t: Stochastic and tokenized modeling for the reciprocal generation of 3d human motions and texts,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cc001b1f-5078-4bf3-8b27-61411de93679 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Neural discrete representation learning,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7d1a1a47-e705-484a-aa53-034314937976 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f22fbf2-7921-43ad-98df-5f01da09ceb6 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Coca: Contrastive captioners are image-text foundation models,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 801e400f-3c78-4cc2-8d11-48671b98f329 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Global meets local: Dual activation hashing network for large-scale fine-grained image retrieval,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 085ce84c-cda7-4bf3-a71f-3ed7e673b97b · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Dvf: Advancing robust and accurate fine-grained image retrieval with retrieval guidelines,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ecb2ba1-991b-4672-ae3d-88e04a36b54b · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Bert: Pre-training of deep bidirectional transformers for language understanding,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d3d0b5d-d03c-4756-a1dc-c42c77076b7a · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Distilbert, a distilled version of bert: smaller, faster, cheaper and lighter,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 982ee88c-1a82-46ff-b814-f8c92ab31047 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Exploring the limits of transfer learning with a unified text-to-text transformer,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0871aeb-5f63-4cab-af47-1f51f2579950 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space LLaMA: Open and Efficient Foundation Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b59dbc53-86c7-4293-8276-ac490fdaf30c · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Audiolm: a language modeling approach to audio generation,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4f80dfef-9939-46f6-9a24-a32d0ec6e2e1 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Robust speech recognition via large-scale weak super- vision,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4660fb3d-fdf7-40e6-89ae-fe74d8854c43 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space SUPERB: Speech processing Universal PERformance Benchmark
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc30a259-434d-4211-a537-8b4b5270b804 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d449bd73-c834-4a0d-80f1-de423635cec1 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Learning transferable visual models from natural language supervision,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae742115-187d-4977-aa0a-de9c99ed3517 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Videopoet: A large language model for zero-shot video generation,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b317a51-829d-4de2-b82a-963fb5ab2e10 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Label independent memory for semi-supervised few-shot video classification,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 801c818e-be16-4958-8cb8-7cc96a45afc4 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Memory-enhanced transformer for representation learning on temporal heterogeneous graphs,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 392e1a5d-4b2b-4ed8-b666-15d909101a2c · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space An efficient memory module for graph few-shot class-incremental learning,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b0982a6-d2e9-414b-bb65-99388e515b22 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Imagebind: One embedding space to bind them all,
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4b8a03c-d931-481a-88e3-c217fb40e66d · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Grounded language-image pre- training,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0e899e0e-25bf-4e3e-ab1b-d6b312477551 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space ActionCLIP: A New Paradigm for Video Action Recognition
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 026a9884-f216-416a-b701-5d015e57fc77 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Delving into multimodal prompting for fine-grained visual classifica- tion,
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 426b2e26-f8d3-4a76-b4a3-3f56eee89156 · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Amass: Archive of motion capture as surface shapes,
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ef95020-8e13-4f9c-9011-b7434957e4bd · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Action2motion: Conditioned generation of 3d human motions,
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0f038371-bae1-4e37-9541-80b9b3c4ff3d · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Representation Learning with Contrastive Predictive Coding
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1a739eb-9df3-4685-bf1f-f5100fa96c4b · outbound
Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space Randaugment: Practical automated data augmentation with a reduced search space,
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.