Pith. sign in

Paper Citation Record · LEDGER

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning

As of 21 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2507.19359.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.19359 v1

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:58:10.803277Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy51
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2a369914-0187-4864-98ee-d518b843b712 · outbound

This paper cites Low-resource adaptation for personalized co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Low-resource adaptation for personalized co-speech gesture generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.978522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.464014Z digest=sha256:4d086f12083b95f5b2f23a9de15c5951042b9a2ce91a323c006202b14416633c

Observation 5457dfed-6409-4d3d-9c77-85298f2dc914 · outbound

This paper cites Continual learning for personalized co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Continual learning for personalized co-speech gesture generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.960158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.470750Z digest=sha256:076488fa4d7a6eb992515b75106d11a8ee18ce658df9aec326ff1bf449fd17a2

Observation e7646036-b0ba-452e-85a7-73f58a081788 · outbound

This paper cites Style-controllable speech-driven gesture synthesis using normalising flows.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Style-controllable speech-driven gesture synthesis using normalising flows

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.942564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.476122Z digest=sha256:d6472f57b0cec3f73ec21e0010fdeacbda033e9f1447804d29038c2a720bba82

Observation c1fdb6c8-8a30-45a9-a9c4-01bcf7c97132 · outbound

This paper cites Listen, denoise, action! audio- driven motion synthesis with diffusion models.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Listen, denoise, action! audio- driven motion synthesis with diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.922392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.482307Z digest=sha256:a74ec5315d4f017c16cfaca27833db584586b76c8f35f32e1451613ad3452f65

Observation 8bfc8bb7-ec6b-410b-9053-672dc46a4cea · outbound

This paper cites Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural em- beddings.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural em- beddings

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.899289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.487888Z digest=sha256:e5c1b23bd77b9444fed003e3c15171479ad4c99e8e5bae334c70f8b3c9135d10

Observation ab987e37-8eff-4508-aa7e-e20bae2e95f7 · outbound

This paper cites Gesturedif- fuclip: Gesture diffusion model with clip latents.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Gesturedif- fuclip: Gesture diffusion model with clip latents

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.875091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.493705Z digest=sha256:be5dcbda9dffc48eca751bcd2224842cb36c1e0ad33d82b895aa187c1daefd19

Observation 94e9c36c-e4f0-41ff-8a54-ac23ea392734 · outbound

This paper cites Probabilistic FastText for Multi-Sense Word Embeddings.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Probabilistic FastText for Multi-Sense Word Embeddings

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.500414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.500414Z digest=sha256:184096029f782ec880f874b9ca0ccb0c1c94a688571d13ec891bd97a28cdb3f4

Observation 88566ee7-fe0f-43ca-a0de-3fd8913cd186 · outbound

This paper cites Enriching word vectors with subword information.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Enriching word vectors with subword information

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.853752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.506359Z digest=sha256:a26c1e2a00ec11991cdcb8ff049a3c44c801c0a671ef2dd88d9ad945d69df063

Observation 5cf0e808-1768-47e1-8c7d-4b3d7c6dca23 · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Realtime multi-person 2d pose estimation using part affinity fields

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.828785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.512123Z digest=sha256:937656d8e2beca310a37054f091474adaeadcbbe4220356d1bdc91be0dd438f0

Observation fd74e80b-f75e-48c0-aabd-82e88a65ae08 · outbound

This paper cites Diffsheg: A diffusion-based approach for real-time speech-driven holistic 3d expres- sion and gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Diffsheg: A diffusion-based approach for real-time speech-driven holistic 3d expres- sion and gesture generation

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.811894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.518555Z digest=sha256:60ea6c73ffbcb38dffa4a0281e9a8c0bab22d7c2fdf5a293c1bb3c98029f092c

Observation 80f96e32-7842-4549-96fe-32f8a6f0c444 · outbound

This paper cites Emotional speech-driven 3d body animation via disen- tangled latent diffusion.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Emotional speech-driven 3d body animation via disen- tangled latent diffusion

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.795182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.524635Z digest=sha256:5a435379c01bbe7800707f1f93332a35886112a2b20cc17db8c8988f5557ed51

Observation 5cbc59ad-8620-4aaf-8b6d-1a57ab31e08b · outbound

This paper cites Ad- versarial gesture generation with realistic gesture phas- ing.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Ad- versarial gesture generation with realistic gesture phas- ing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.774747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.530964Z digest=sha256:3f78fab8cb13447f0a21c7d6a6cee5168337b547535070d3464ae43d28ec6f7d

Observation 6ec8beb7-3664-4251-862f-8fa7967d9515 · outbound

This paper cites Learning co-speech gesture representations in dialogue through contrastive learning: An intrinsic evaluation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Learning co-speech gesture representations in dialogue through contrastive learning: An intrinsic evaluation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.751168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.536833Z digest=sha256:7a7141b6ae063c968c76fa0f0fae73dc2da72036e239ec0fb92d6ff2958e9896

Observation 0462ef3e-370e-430f-8cd0-27163ed3b5be · outbound

This paper cites I see what you mean: Co-speech ges- tures for reference resolution in multimodal dialogue.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning I see what you mean: Co-speech ges- tures for reference resolution in multimodal dialogue

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.730723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.542416Z digest=sha256:6233bd1c4b199697b635e21f7e1215c43355d704fceae8bffdf02fedb6aaf21f

Observation b01c32f6-47b7-4b71-a9b0-d6c28032f172 · outbound

This paper cites Generating diverse and natu- ral 3d human motions from text.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Generating diverse and natu- ral 3d human motions from text

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.712533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.547421Z digest=sha256:17586b9fcbe967956e5c8bac02b4accb6a235afda6d41dae2634feaaddf378a2

Observation fb4b91c1-ab5d-4b05-b117-5b4ca7acdf3b · outbound

This paper cites Tm2t: Stochastic and tokenized modeling for the recip- rocal generation of 3d human motions and texts.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Tm2t: Stochastic and tokenized modeling for the recip- rocal generation of 3d human motions and texts

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.693687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.555658Z digest=sha256:1538e91aaf9bc03ca89a95c93b4182e369cbf8676a9d6e14bfad6d0b36d258ad

Observation addd1381-b6b9-4a1e-9cc8-7053f404b510 · outbound

This paper cites A motion matching-based framework for controllable gesture synthesis from speech.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning A motion matching-based framework for controllable gesture synthesis from speech

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.676393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.561784Z digest=sha256:cbe2d841eef1ddbb6fdfc7cdfea1ebcea60c17245ad87ebbcd93403d0cd08eae

Observation 2f6af49f-cfe9-4110-94fb-b6128bde25a0 · outbound

This paper cites Moglow: Probabilistic and controllable motion synthesis using normalising flows.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Moglow: Probabilistic and controllable motion synthesis using normalising flows

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.567367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.567367Z digest=sha256:5a3858cf9103774f16288ad1c02be760b06514b4dc9e14916f967bc88451792e

Observation e918b296-5546-4274-9935-04ae64c7514c · outbound

This paper cites Multimodal lan- guage processing in human communication.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Multimodal lan- guage processing in human communication

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.644964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.573762Z digest=sha256:c4fe0080a8ee95b95993b3a6ce6a1fc288e6118185c3d1fb7833bde3497e7a3e

Observation e4146ade-7975-44de-aa2f-31bbb7fc22dc · outbound

This paper cites Hubert: Self-supervised speech repre- sentation learning by masked prediction of hidden units.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Hubert: Self-supervised speech repre- sentation learning by masked prediction of hidden units

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.627735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.579069Z digest=sha256:d4c0474530ec2f4e9d24de8e157b27bd019c79a8cc986e4147d161c50f1b219a

Observation 6c52d0c9-5e73-436c-b98c-407a1c1b034d · outbound

This paper cites Gesture units, gesture phrases and speech.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Gesture units, gesture phrases and speech

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.608962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.584471Z digest=sha256:d88262e6ac77eb3d4683f92702f554413725d52578eff4e17942d2b9ce9d95d6

Observation c026dc28-04f2-4622-b924-19220748c1b3 · outbound

This paper cites Gesticulator: A framework for semantically-aware speech-driven gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Gesticulator: A framework for semantically-aware speech-driven gesture generation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.592239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.589553Z digest=sha256:2c6606581d5d259743d9f4b6fc31a58e34195e8ccca1839874201cb58da92114

Observation 97c4ef38-f28b-4b04-a39c-1a96f9e9373d · outbound

This paper cites Danceformer: Music conditioned 3d dance generation with parametric motion transformer.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Danceformer: Music conditioned 3d dance generation with parametric motion transformer

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.575266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.595042Z digest=sha256:a2af45d5bbe1e7cbec0ceccf15bb701437bf3b037a64e973f3176b781da3a9ae

Observation f2f3dc36-2673-4da8-a0cd-c47ec2a4ea88 · outbound

This paper cites Audio2gestures: Gen- erating diverse gestures from speech audio with condi- tional variational autoencoders.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Audio2gestures: Gen- erating diverse gestures from speech audio with condi- tional variational autoencoders

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.555279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.601122Z digest=sha256:924bfc31982808ee0a265782bb9d251465deec3666eb942e3e6032eddb063a2b

Observation e5d98716-a61d-4a6a-92c2-402b399f3a29 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.537429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.606397Z digest=sha256:c8fcc41d97ad0ff0a5fcb8d7015c10428c0846bef58d73f18f1c357b9c6711aa

Observation 1859d201-ee9b-42c0-8442-6871cacac9e4 · outbound

This paper cites Seeg: Semantic energized co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Seeg: Semantic energized co-speech gesture generation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.514193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.612059Z digest=sha256:a612f9565600cee37c0c327cadeca0aa7717c72763cbbd318cbfa0887061b48e

Observation 63723e77-5e9b-4643-9449-9c8f39cae7b7 · outbound

This paper cites Beat: A large-scale semantic and emotional multi-modal dataset for conversational gestures synthesis.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Beat: A large-scale semantic and emotional multi-modal dataset for conversational gestures synthesis

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.493945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.617841Z digest=sha256:f41d473f67b99adaff0d7a6fde335ee913ba639dac0e473652939869052eb6ea

Observation e1af1a41-b178-4235-bf36-c50901ac28a7 · outbound

This paper cites Emage: To- wards unified holistic co-speech gesture generation via expressive masked audio gesture modeling.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Emage: To- wards unified holistic co-speech gesture generation via expressive masked audio gesture modeling

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.471095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.623866Z digest=sha256:3adc94ea6177f9010cbdcefd5617ed2ce29424725314c85f9e31e933a382e5d0

Observation 3e79e1f4-e4b6-4c70-96ca-70fef823aebe · outbound

This paper cites A Survey on Deep Multi-modal Learning for Body Language Recognition and Generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning A Survey on Deep Multi-modal Learning for Body Language Recognition and Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.629664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.629664Z digest=sha256:48857a57bda5f35f1bf5b4bd97738c21a0604a5d268e394c2e11df368e028169

Observation b1a42bd8-8968-4654-b16d-612e8948511d · outbound

This paper cites Human gesture recognition with a flow- based model for human robot interaction.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Human gesture recognition with a flow- based model for human robot interaction

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.452611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.637240Z digest=sha256:5c0b9473e76919912234d303bd198bea1f85cb1a7968c4a7cff2fe486b902074

Observation 5e91daf3-6612-44c1-9665-b28a190c8489 · outbound

This paper cites GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.642697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.642697Z digest=sha256:c59c3f690261a70eb8efd768801b861184f8ba196a26d6504bd1233812b59783

Observation d5d1c9b4-d8e5-4d3b-b904-a876b3f87a33 · outbound

This paper cites Audio-driven co-speech gesture video generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Audio-driven co-speech gesture video generation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.431276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.649300Z digest=sha256:496331da0da556526585ade84e33a277aaa47a4f6913888e062b2060e0af75df

Observation 41f809e5-4bb7-4437-b1d7-653c7fd6735f · outbound

This paper cites Learning hierarchical cross-modal associa- tion for co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Learning hierarchical cross-modal associa- tion for co-speech gesture generation

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.412218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.654514Z digest=sha256:e725bf705b1a7b4e25578187299959ff363a47ca55a30d68530491ca4398df2e

Observation 87005c43-eed6-4748-bd01-3a20151ea9b9 · outbound

This paper cites Towards variable and coordinated holistic co-speech motion generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Towards variable and coordinated holistic co-speech motion generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.392979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.659509Z digest=sha256:78a563d3ff5a9574cbf6765448aaa8e9748e5b458d45cb8a043d53432e9b805d

Observation 4f06c41b-699f-4fc2-955a-ade00a863d27 · outbound

This paper cites Hand and mind.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Hand and mind

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.373974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.664810Z digest=sha256:73b80ba800fc4dae456b9d2f1bfe9daa450761c23b69b32f2073eb3d7f168a79

Observation 487beccf-5f39-4888-b5c1-1dc4b8f6fad0 · outbound

This paper cites Convofusion: Multi-modal conversational diffusion for co-speech gesture synthesis.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Convofusion: Multi-modal conversational diffusion for co-speech gesture synthesis

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.353990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.670321Z digest=sha256:b7d167489256217d2bc7ee415070bdf0c5fe3019a2b1dd5e836d0abfa9bebfb8

Observation b411826f-470f-4d9f-8202-386eced30e1d · outbound

This paper cites Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.675454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.675454Z digest=sha256:b402bf6f0b51c3e0e0105c40b542a17ed40e9f4b278c06abfd388a4fb2941373

Observation c18250c2-8866-497d-a9bd-95386bdc026f · outbound

This paper cites From audio to photoreal embodiment: Synthe- sizing humans in conversations.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning From audio to photoreal embodiment: Synthe- sizing humans in conversations

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.334874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.680733Z digest=sha256:4ab9746d1a4cb771ccafdc2d2e1117812e7c09363160bf87decb622c88421d4b

Observation e88d1ab4-2573-4245-a5ab-c4fce6e7b9b3 · outbound

This paper cites DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning DCTdiff: Intriguing Properties of Image Generative Modeling in the DCT Space

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.685667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.685667Z digest=sha256:7d4df668516a278a3eec900671d061d9754aef61e14321cdef99c7897ab3cf66

Observation 7383a860-25bc-4815-a8e6-bafee2ce15f5 · outbound

This paper cites A com- prehensive review of data-driven co-speech gesture gen- eration.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning A com- prehensive review of data-driven co-speech gesture gen- eration

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.310383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.691493Z digest=sha256:8cac3be94c387872c669966741ce5139c19b59463d9981c9938229c4dded4b4c

Observation f8aa5d44-0b04-4fc0-8fc6-c0ddf206c1a6 · outbound

This paper cites Hearing and seeing meaning in speech and gesture: Insights from brain and behaviour.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Hearing and seeing meaning in speech and gesture: Insights from brain and behaviour

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.284370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.698400Z digest=sha256:04af6738aef4f66c99e327dc8d425579116a16ff6426062b8905318652500f0c

Observation 94b68f22-5e8d-4e95-b43f-68146d7a46f0 · outbound

This paper cites Bodyformer: Semantics-guided 3d body gesture synthe- sis with transformer.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Bodyformer: Semantics-guided 3d body gesture synthe- sis with transformer

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.266701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.703991Z digest=sha256:8149505c533d376ba492451ee39617a51274bbfa29ead6fc23ffb640dec5f2d2

Observation 35202fb1-4ff7-49b8-b3b5-4faa959ec217 · outbound

This paper cites Expressive body capture: 3d hands, face, and body from a single image.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Expressive body capture: 3d hands, face, and body from a single image

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.247742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.709402Z digest=sha256:514314abafe202a0c5357263014a10b2b2d538570487979d5b7fd50d81ebde8b

Observation 2fbc9fa5-6cde-4070-83f9-666baf205acc · outbound

This paper cites Weakly-supervised emotion transition learning for diverse 3d co-speech gesture gen- eration.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Weakly-supervised emotion transition learning for diverse 3d co-speech gesture gen- eration

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.231582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.714964Z digest=sha256:9bb8f8556aff0b89fa53c4a3169214c5caa07a9e482125dcdd3b8daf3e8c13ed

Observation 3cfa6d47-4c23-4b02-9a70-37a219192fa8 · outbound

This paper cites Co- speech gesture synthesis by reinforcement learning with contrastive pre-trained rewards.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Co- speech gesture synthesis by reinforcement learning with contrastive pre-trained rewards

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.214404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.720089Z digest=sha256:0638dcc20ae218bd864123426068df5d2d290ba6c214ef9da2b6951b0eaf0c9b

Observation 26f59969-8eb0-4db0-8d0e-85b107a50ba1 · outbound

This paper cites Motionclip: Exposing human mo- tion generation to clip space.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Motionclip: Exposing human mo- tion generation to clip space

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.198056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.725430Z digest=sha256:3ea207e58c722f9ea99e1909218dcc965f9cf7dbd19d34d18d062ffe81356fa0

Observation 5a4778a1-63df-4c90-956d-0724c13c0dfb · outbound

This paper cites Human mo- tion diffusion model.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Human mo- tion diffusion model

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.181326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.731077Z digest=sha256:d90fad923c0d747cabeab4206b694ce697c73b0a7906e939940979ac2fb64837

Observation 5ce1873c-8004-4224-86f5-1001e2dacb45 · outbound

This paper cites Neural dis- crete representation learning.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Neural dis- crete representation learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.163899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.736803Z digest=sha256:04c255fcf97a658088b9a4e312d580a25f834bc334252676538b885dea6f6f01

Observation 3577dac7-e6c7-4f0a-b6d3-fc49f6dd8cd1 · outbound

This paper cites Augmented co-speech gesture generation: Including form and meaning features to guide learning-based gesture synthesis.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Augmented co-speech gesture generation: Including form and meaning features to guide learning-based gesture synthesis

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.146181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.744751Z digest=sha256:56276f41f9fb6f69982bef3ab67d1ddb90b6e20d5003d89919f60c5213cccec8

Observation c6bd66ce-285e-4c4b-93c4-2ac292c4543f · outbound

This paper cites DiffuseStyleGesture: Stylized Audio-Driven Co-Speech Gesture Generation with Diffusion Models.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning DiffuseStyleGesture: Stylized Audio-Driven Co-Speech Gesture Generation with Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.750171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.750171Z digest=sha256:dd1a2404764a16d391a663367f87c7d53e387838a5b00f01df1fd88bddd03676

Observation 96139976-0213-4646-8765-753212772a96 · outbound

This paper cites Qpgesture: Quantization-based and phase-guided mo- tion matching for natural speech-driven gesture gener- ation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Qpgesture: Quantization-based and phase-guided mo- tion matching for natural speech-driven gesture gener- ation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.124883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.756455Z digest=sha256:43c4f2497356ea9b2897dd3506e8e7acccd62c3458ad7357af547e8d1e865ac2

Observation 1e9ac3a9-2807-4c75-a82a-837aa08e7fc5 · outbound

This paper cites Audio-driven stylized gesture generation with flow- based model.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Audio-driven stylized gesture generation with flow- based model

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.106645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.763496Z digest=sha256:efa553918bec6bcd41f21d7d072e66179aab1ed54a3f71cab6a4fab354b5b545

Observation e93a48d9-221e-4878-8c88-4bac1dfcb41a · outbound

This paper cites Generating holistic 3d human motion from speech.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Generating holistic 3d human motion from speech

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.080714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.771499Z digest=sha256:793590fe11b874d56372c54528e8b7e2540ac5ed32fe62863af17b2f19bbc808

Observation 063dd7b3-3c6d-4fe4-857b-b9aaae4ee39c · outbound

This paper cites Generating holistic 3d human motion from speech.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Generating holistic 3d human motion from speech

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.060915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.776489Z digest=sha256:120c503f34363433afb05100af45af9e1a64434ec4549e00a61bb227a87e38fa

Observation ed1d6621-7d32-4833-95de-1a93bec97f47 · outbound

This paper cites Speech gesture generation from the trimodal context of text, au- dio, and speaker identity.ACM Transactions on Graphics (TOG), 39(6):1–16, 2020.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Speech gesture generation from the trimodal context of text, au- dio, and speaker identity.ACM Transactions on Graphics (TOG), 39(6):1–16, 2020

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.038719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.782378Z digest=sha256:3f674fe2cf4ba1f146a6b7fb3052129d9314d8375c6ff6cae9364260dfc95333

Observation cf017540-659c-45d3-a98a-072f73df7a45 · outbound

This paper cites KinMo: Kinematic-aware Human Motion Understanding and Generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning KinMo: Kinematic-aware Human Motion Understanding and Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T17:58:10.787476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:58:10.787476Z digest=sha256:f1bcd8333069a03ba34a31e42c48b91447f0c4d9f66dd548a85037e85c3d5dc4

Observation e239708d-84e9-4dd9-91d2-d2b33f01222c · outbound

This paper cites Semantic gesticulator: Semantics-aware co-speech gesture synthe- sis.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Semantic gesticulator: Semantics-aware co-speech gesture synthe- sis

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:11.017920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.793178Z digest=sha256:ee84566325d57ea4a45e1ea2be30ff1621c1c979fc4419fcf36da7f1dc8b9b61

Observation e2d67e0e-60e4-421e-9917-c72c6c669777 · outbound

This paper cites Livelyspeaker: Towards semantic-aware co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Livelyspeaker: Towards semantic-aware co-speech gesture generation

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:10.999963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.798218Z digest=sha256:6c546b4ac87837ccf2d2c177ae044f3ae10109eca993a55e5ea02861e7e96381

Observation 61770172-1db9-4db1-b6de-c788e0b5e8f9 · outbound

This paper cites Taming diffusion models for audio- driven co-speech gesture generation.

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning Taming diffusion models for audio- driven co-speech gesture generation

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:58:10.981260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T17:58:10.803277Z digest=sha256:80ff93e0958e06bad81a7eb963a001800c2c0d12696574ef9ed5e5ca22660bf0

Pith citing papers

No inbound Pith citation observations are available.