Pith. sign in

Paper Citation Record · LEDGER

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

As of 13 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2412.02419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02419 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:34:10.895977Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:16:44.323653Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:16:44.447784Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy60
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ca2d14f-68e6-416e-8beb-072713cd51e4 · outbound

This paper cites Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.790910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.605320Z digest=sha256:7fe58fc4c970e21968d664d0dd530ae53a13250bcbe5b2907184ef7e88f43f8b

Observation 5915c4ce-a3a8-4a78-9186-f1c5c2c6e941 · outbound

This paper cites Style-controllable speech-driven gesture synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style-controllable speech-driven gesture synthesis using normalising flows

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.780191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.609753Z digest=sha256:45e08aa26f274fc317cd9197f8523ad216db871746d17ecf2b6d4663ee0bcb0a

Observation 23fa6aa1-13a1-4474-81cc-37bca8c711a9 · outbound

This paper cites Listen, denoise, action! audio-driven motion synthesis with diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Listen, denoise, action! audio-driven motion synthesis with diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.769829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.613798Z digest=sha256:0c76424cd29ea1a1a391f1255d6443982e037783d28f09c3fbb18a90a1be3bdb

Observation 56da8f1c-79f6-436a-be1c-d6fb598e5c18 · outbound

This paper cites Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.758838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.618075Z digest=sha256:24662fb875a9d235086de890262adead0b22aa522bb4d43158f711e85f2428b2

Observation 20fae4cf-0e19-4bf7-a8e5-5b3370613036 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.747711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.622171Z digest=sha256:8a9dd697dcfa35232be4a4dd4965a6968b313a9c8c1bfc8a2ef72dd46958ec5d

Observation 8cb414e2-3ad3-4685-80cf-eba842005ca4 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.736630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.626457Z digest=sha256:addc3dfdb75055f3cba8d8407a4747891d6f9b2844059b6f02a3299c01a9db02

Observation b397dc9e-a53d-4c1d-9d7f-a9140d0dd09e · outbound

This paper cites Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.725245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.630449Z digest=sha256:a047cf1330363b33ffd323ae73576eb668ef470e04d1fef0dae38c9792b42039

Observation 6431bad8-6118-4472-876a-11ca2642cacd · outbound

This paper cites Digital life project: Autonomous 3d characters with social intelligence.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Digital life project: Autonomous 3d characters with social intelligence

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.711947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.634808Z digest=sha256:4583816c634bee6ca950b9292520fea84e62e6ab82cd6fbf9d69229c29cd2920

Observation 7e103cd7-d7ae-428b-84aa-57f58c0f276f · outbound

This paper cites Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.700250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.638849Z digest=sha256:f3e3c28dbabdda1042ccef3a5902206a267286ea53f06aead92878e782714717

Observation cfb72b07-3375-4e42-8bd1-870df924531b · outbound

This paper cites Beat: the behavior expression animation toolkit.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Beat: the behavior expression animation toolkit

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.687482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.642619Z digest=sha256:16850c3ed4034efbdcca95952e7e30f38dc1d5a85bb69f9e4453bf8d56b2da0d

Observation 85d61874-e5c6-4922-96d3-b600001ea7ed · outbound

This paper cites Taming diffusion probabilistic mod- els for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Taming diffusion probabilistic mod- els for character control

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.674539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.646002Z digest=sha256:4cba62901802d296405d8baf81fb02aa85bc57d16ed361b63a466d352669cd17

Observation 02990db2-beda-4498-86b3-5acf292fe63c · outbound

This paper cites Executing your commands via motion diffusion in latent space.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Executing your commands via motion diffusion in latent space

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.662865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.649538Z digest=sha256:663b655001e679c86f4a88955b70df5641989f938201cc83fffb12c9faf78678

Observation 3999d157-0ad4-48c8-8792-b022fbee905e · outbound

This paper cites Black, and Timo Bolkart.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.651114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.652577Z digest=sha256:416b720c0dc71b5c10aed1c2751a23bf9222e284827921092eef236af043538f

Observation d942fe81-7f19-4eeb-8422-60fc06bfc535 · outbound

This paper cites The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.639164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.655939Z digest=sha256:9b92cb8b9be55985570ca8e8e9c94d8723d3e6c09098ab677e78fed5c65370f0

Observation 8d99c66d-a64f-4890-9654-6d40300cbef8 · outbound

This paper cites Diffusion-based co-speech gesture genera- tion using joint text and audio representation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Diffusion-based co-speech gesture genera- tion using joint text and audio representation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.658853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.658853Z digest=sha256:c77c35dadd0e54b21e231176637de1275f9be5ec45f61503c52b76b92879e3f0

Observation 3abea151-4294-4c25-b01d-f7eaf2b334c2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.662380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.662380Z digest=sha256:b622e79eedc7d7227dc20c81902dd924e6afb29d9008d27341946643503b19a8

Observation d96f74c4-816d-45e0-8b73-21673e297df5 · outbound

This paper cites Freemotion: A unified framework for number- free text-to-motion synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Freemotion: A unified framework for number- free text-to-motion synthesis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.618609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.666246Z digest=sha256:c68f07b78c512afdb8688df134a8158a53ec2264169f752fc5e6b06d732f730c

Observation 1a8b866c-d5e1-4e38-bcd5-f656c1de7739 · outbound

This paper cites Troje, and Marc-Andr ´e Carbonneau.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Troje, and Marc-Andr ´e Carbonneau

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.606790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.669724Z digest=sha256:d2822a913110ffc27d72eb0d9285fad0cdc74979a958ff2d9797389adbd36e78

Observation 41df565f-9d18-483f-93ee-3c04dfccf4db · outbound

This paper cites Remos: 3d motion- conditioned reaction synthesis for two-person interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Remos: 3d motion- conditioned reaction synthesis for two-person interactions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.595162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.673292Z digest=sha256:74e2e0447ce5e9e88caf06043b8b23c9d12fd67760c4da0e29d43e2739168346

Observation be8d6683-299f-4191-ad28-cc5dbb6023c7 · outbound

This paper cites Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.582801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.677265Z digest=sha256:b36046eaea8ed433dcac2bdd15cc1e075a5bf13e80342fb7431fc60ee9cc171f

Observation f65e4349-e081-4312-bacb-6d623a6315ab · outbound

This paper cites Learning speech-driven 3d conversational gestures from video.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning speech-driven 3d conversational gestures from video

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.571181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.680789Z digest=sha256:2b7359237c4175258a3371344596a44409ed34c639d4b3ce11eacdeeeebb44fc

Observation 439a4759-23c4-4b32-8f36-fc848410ef8a · outbound

This paper cites Evaluation of speech-to-gesture generation using bi-directional lstm network.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluation of speech-to-gesture generation using bi-directional lstm network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.558985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.684444Z digest=sha256:036492bea50dc38ba810514b9d2d739fc50a44e6546f5dd395b6acadd16bc28d

Observation 6316e41f-a6d5-4520-957f-96fb48617ede · outbound

This paper cites Moglow: Probabilistic and controllable motion synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Moglow: Probabilistic and controllable motion synthesis using normalising flows

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.547685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.687870Z digest=sha256:12653d6ee45575579fc114c837522644fe57fe7700eab1e966707d2bb9b8b2d4

Observation 6880d816-5fdb-4531-a03d-3c4ebe8d07d4 · outbound

This paper cites Denoising dif- fusion probabilistic models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Denoising dif- fusion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.536260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.691524Z digest=sha256:a824d147d3fa785c137db74b511360df7720de0671aa360f6ddd9f438f8c384e

Observation 3d47a550-f56b-4174-abb6-ac35337d1631 · outbound

This paper cites Dead blending, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Dead blending, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.522980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.695463Z digest=sha256:0ed750788128a35d7b525f61906f18fecfce01fdc5bbce020684d1671a8489a8

Observation cc08d7be-1e85-419d-9115-5f77e7e8c519 · outbound

This paper cites Phase- functioned neural networks for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Phase- functioned neural networks for character control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.511947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.698968Z digest=sha256:6ee1e5a4124352e30206e048ce5cae62e7fcf48940f20f45655e4dd162b19fa7

Observation 2b0e30cf-8f77-4624-a3e6-e99e4ea8d345 · outbound

This paper cites Example- based control of human motion.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Example- based control of human motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.500817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.702423Z digest=sha256:951a3dcaab00758e5b2c0ace6bde2be1ddd7aecaa0c7787f383afa07ca3817e8

Observation d1972587-0b70-413a-a0b4-cd858bc3d2e3 · outbound

This paper cites Robot behavior toolkit: generating effective social behaviors for robots.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Robot behavior toolkit: generating effective social behaviors for robots

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.489364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.706354Z digest=sha256:716b91bcb0de6c556467ae568aa90fcdd6980f7b2f7593e913072878d4019d18

Observation e76a4e53-20ee-4230-b662-17e640e00c88 · outbound

This paper cites Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.478397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.710164Z digest=sha256:ec4698c6583d0459b70f612e97b59433d6b33ae6d07d0080ec839dcf1cdd6127

Observation 5975b0a1-b9a4-4922-b180-6af8b50f6144 · outbound

This paper cites Intermask: 3d human interaction generation via collaborative masked modelling, 2024.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intermask: 3d human interaction generation via collaborative masked modelling, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.467350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.713926Z digest=sha256:111f6e048eaeee4ac6f5350329451269985b4a594bca24f1eb1d1f272679c97d

Observation 6a1026dd-d786-468d-ba8d-5105a70365b1 · outbound

This paper cites Synchronized multi-character motion editing.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Synchronized multi-character motion editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.456925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.717512Z digest=sha256:6d7aeda4531f9834cafaeb9a2268393f057f2abb2f2abad2a3e502084c29d9e3

Observation b151996e-cce2-4bae-b3c5-f427d7b187d9 · outbound

This paper cites Tiling motion patches.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Tiling motion patches

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.446327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.721459Z digest=sha256:3578c00174d6a4fdd4bb41c620e83ee51e5ee59cfd99efc8f0a32f6510aaaeed

Observation f4c80f5a-1a47-4127-ae36-802a4891f356 · outbound

This paper cites Towards a common framework for multimodal generation: The behavior markup language.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Towards a common framework for multimodal generation: The behavior markup language

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.435790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.725418Z digest=sha256:1b59bb263218e4b297abfde44e9f3b86936a26e0a3c3604166cbc3e1ae636b59

Observation 3ee368d6-8506-447a-8370-2ed9a34fc80c · outbound

This paper cites Gesticulator: A framework for semantically-aware speech-driven gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesticulator: A framework for semantically-aware speech-driven gesture generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.424954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.729119Z digest=sha256:aa560d885f61f08072cdd1c2a04bf12c711cec372066e8733d509d2e6dd4f53e

Observation fe930589-3c4b-4eac-8d7d-633d2cfbad82 · outbound

This paper cites The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.413976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.733295Z digest=sha256:11022c1602d29e4d73c8e6c90118041264933a198da545cb8f1573dd9ab3acbd

Observation db9a7417-ebc6-4442-8197-e2d8f5ffe4ba · outbound

This paper cites Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.402783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.737923Z digest=sha256:6c067044e9c986d9a06ca0ffb16127337ce3b1074d401e09092350625fe739ba

Observation 07d536e3-ea5a-4b05-81ed-b0a8cc5e7334 · outbound

This paper cites Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.391368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.742001Z digest=sha256:1d4bdab2ccd967f494e3d7e9ff41889ec986a349a21d30dcbd89f70c040f4bca

Observation 34fd5287-c8b0-4620-ad91-b95e190c0122 · outbound

This paper cites Two-character motion analysis and synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Two-character motion analysis and synthesis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.380112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.745937Z digest=sha256:95b01409bd5f22b42f3cc83aa618e5ca8c9a123fe12fb079f9f6f5132baff25d

Observation 88e21c0f-5319-46d7-a528-85150a04f76f · outbound

This paper cites Motion patches: building blocks for virtual environments annotated with motion data.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Motion patches: building blocks for virtual environments annotated with motion data

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.369173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.749738Z digest=sha256:437cfc0ec47445ff32cae0f6a47b9e62fcc06f54b01a68fa59af7c00f1f7b3a2

Observation eeb58929-60ac-4dcb-bb5f-cb6f417baeec · outbound

This paper cites Gesture controllers.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture controllers

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.358230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.753522Z digest=sha256:e600c4b189b2140ac073bf5a3cc881fbbf15afce1799c7b13846fc47238bc2d1

Observation ce76ea1f-9bcb-4a30-83b7-e9f25ee0032a · outbound

This paper cites Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.346673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.757026Z digest=sha256:43f771ad4509390524c18c679a711266b2e9441dea55c4c1dcca2c50cbf5fede

Observation 33e9020d-0cf5-4ca8-b632-16a7e2524a53 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.335671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.760840Z digest=sha256:0fe2f7a02dd8147ea05b74014a3c35bf9fb3ea67a5cd568a69fec92fc4a646f9

Observation 66f12494-a44e-4e61-aecc-0c913f74421f · outbound

This paper cites Intergen: Diffusion-based multi-human motion genera- tion under complex interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intergen: Diffusion-based multi-human motion genera- tion under complex interactions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.323941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.764298Z digest=sha256:54ccfeb135e27db15448d5e4de74030b79f33c0955f2dc887dc0114743200083

Observation b7ec29bf-bd6c-4ff3-87d4-1650789ade19 · outbound

This paper cites Com- position of complex optimal multi-character motions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Com- position of complex optimal multi-character motions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.311617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.768474Z digest=sha256:b8974c3473b85332bc68142602525b5f0aab2f391804be8766edb9de8d7eb225

Observation ec01dcc6-bd0c-4b0d-a6d3-caf1a887e39e · outbound

This paper cites BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.771875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.771875Z digest=sha256:5e77e9a4452a088b23293ddfea50e06834cbab6f077f7a39d2373b6fb4a72772

Observation 2ed6c7a3-56e8-4729-b2b3-af2a888a2c7c · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.300739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.775706Z digest=sha256:a38b4425bcd30b7755a54688249867f31937e991964b4e4052cd5492ea379f6e

Observation 896ea1c6-8479-480e-8f25-f37af429c758 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.289500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.778805Z digest=sha256:2d963ea054649efdce3f546ea70293f7a6e1144b56f1b6e985363c81dbbcea7f

Observation 492e9b10-72f3-4075-9c87-15620a75b76e · outbound

This paper cites Learning hierarchical cross-modal association for co- speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning hierarchical cross-modal association for co- speech gesture generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.278004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.782159Z digest=sha256:52a73981729a59f826a188bdd44e9542ddc482a403997de2439146aadf61c707

Observation 87d76ea7-3213-4a20-9f4d-760622c75bc4 · outbound

This paper cites librosa: Audio and music signal analysis in python.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model librosa: Audio and music signal analysis in python

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.265161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.785431Z digest=sha256:4c65538de3ff97549f7a70be423978e7b1959486f6ab2ba4c94bbf331d29fdac

Observation 841a0d39-cf2a-4aa3-b837-0cdeb2fc7b2f · outbound

This paper cites Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.788440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.788440Z digest=sha256:9038a561467a4e16922b5f333acd3866d48f63112bdd2b6857db6d50314ba35f

Observation 3d7bf52f-16fe-4278-9957-e82fc9ed806d · outbound

This paper cites Gesture modeling and animation based on a proba- bilistic re-creation of speaker style.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture modeling and animation based on a proba- bilistic re-creation of speaker style

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.246159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.793221Z digest=sha256:fd2ca7a9c1f3f16a9e67b17afd380ac6ba3f5b387c7d2fbd23d520c6e60d8597

Observation dbacaf9f-79b6-4c46-8db8-274cf3022c59 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.234149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.797880Z digest=sha256:28aabebe0decbfd995bb66528ccdf17fa7dd45f903f6e2dcb42ec1678b687227

Observation 6b842028-841c-4509-abbd-437d226e06e3 · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.221478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.801777Z digest=sha256:69201e9594ba3124ce56e3b1709b1d55964d16e63838b2d7228d8b311da1cbe8

Observation 2cb3c7b4-ee91-49dd-a6fa-dd0cbacbb2e0 · outbound

This paper cites A comprehensive re- view of data-driven co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model A comprehensive re- view of data-driven co-speech gesture generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.209841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.805434Z digest=sha256:bd2f0e1ecdbc21abd69752fecf7d04f03befae8baadc840229b0e048fd2005aa

Observation a6cdc883-08ad-40df-8e89-2e719926fbe7 · outbound

This paper cites Librispeech: An asr corpus based on public do- main audio books.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Librispeech: An asr corpus based on public do- main audio books

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.197919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.809497Z digest=sha256:7b6403e60a44bdb2fc2216ecda7b0459c1682016f072e621b60f01a4f43e3f2f

Observation 96373b2b-9d16-4261-be07-a9b8a5389ead · outbound

This paper cites Bodyformer: Semantics-guided 3d body gesture synthesis with transformer.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Bodyformer: Semantics-guided 3d body gesture synthesis with transformer

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.186332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.813195Z digest=sha256:f6393ffa6fbc7dd693cb81ee560e2f03cb6608d30b711fcebcae0c5997edc21b

Observation 06947c17-cb57-4ad1-a79e-9497e1f20f13 · outbound

This paper cites Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.175473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.817013Z digest=sha256:c42ba95f30aae1429c8dd6835c3f381cff4c38944ab7b96d21fea85cfa709ac6

Observation 9608dfc9-9723-4269-bd87-dc81c22c86ef · outbound

This paper cites Hierarchical text-conditional image gener- ation with clip latents, 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Hierarchical text-conditional image gener- ation with clip latents, 2022

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.820676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.820676Z digest=sha256:74d17bb8b7163c2366edc1591246cd566d478ff24b925241d1d3a1ad66160b39

Observation 320c13ba-d1ba-42b0-9e1f-905dc283375d · outbound

This paper cites Human motion diffusion as a generative prior.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffusion as a generative prior

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.824254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.824254Z digest=sha256:b0098527f186e4e2ca321b369bc6d4c496e0e5ee428720e065bb83404f173c66

Observation f186f572-c4df-4ef5-9c3c-59cd0b75dff6 · outbound

This paper cites Interaction patches for multi-character animation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction patches for multi-character animation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.150778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.828174Z digest=sha256:70bcbd8c5c8134400d5c68f8531cb361ee3979304d990162de08c28b3a3ff3dc

Observation 0730844b-9f8c-4d86-bf71-2a7a2509860e · outbound

This paper cites Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.140373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.831698Z digest=sha256:eb7784288f5cc66a48e409f556e918153bdbb436a86c4d9aa09450380bcb101c

Observation 025536bc-6071-4b24-b77e-2d7cd30f4927 · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.130124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.835159Z digest=sha256:eaddf349f7be4751e35ed4feb934c86e186c16bd29fb534f55b454b47ba7c4ac

Observation 032f199d-881a-410c-a726-83cede472aae · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.118915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.838922Z digest=sha256:9a34f7f51c0630069f7ec32169acbf48c0b76b975c4e9e21876ab4b4c536a144

Observation 1f82c318-2fb5-4e16-8299-1414a0979751 · outbound

This paper cites Human motion diffu- sion model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffu- sion model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.842737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.842737Z digest=sha256:d4e35280459523e042390b5a8acf1a85ea67a468aefddd2865ff895566ad442c

Observation 8480fd60-50da-43cc-9e08-ab52fbfc4716 · outbound

This paper cites Attention is all you need.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Attention is all you need

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.846546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.846546Z digest=sha256:9b4c1455a1d62d722e8948d4c830d00266b48a1b3e32efe7671b06823a622820

Observation 9904a46b-a906-44db-bb81-bda659acea85 · outbound

This paper cites Gesture and speech in interaction: An overview, 2014.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture and speech in interaction: An overview, 2014

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.093731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.850351Z digest=sha256:ba0be85e038909c0725141fabcf13768c71c4f856f6421afbeeab07829deb972

Observation 0cfc6845-f777-4876-9fc4-187caeb14324 · outbound

This paper cites Generating and ranking diverse multi-character interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Generating and ranking diverse multi-character interactions

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.082699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.854232Z digest=sha256:fcf54c18c199d87bbf6b35bde01510cab1014e95b33b3418611b4159bcb0ffb1

Observation 04ebfbf7-9c3e-4e2b-9f7b-4cad06788d4e · outbound

This paper cites Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.070716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.858168Z digest=sha256:c7214458c87cf6b9219c7d36d18956094b413577017eb146eb6f1c148454f0dc

Observation 1857373f-ab58-4012-b90b-9c575ef60273 · outbound

This paper cites Speech ges- ture generation from the trimodal context of text, audio, and speaker identity.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech ges- ture generation from the trimodal context of text, audio, and speaker identity

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.058176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.861661Z digest=sha256:ec330d68301e575db1e308018abe471f0e51bae42a5b5c9ba9aa57a96313ea34

Observation 6c8e1412-677c-4b5e-a378-08917eb488c0 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.865334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.865334Z digest=sha256:1432bb3ac61b310a306ffbb9ba905ad9260755940424e42531162ab9e2212305

Observation b08a51bb-227b-4d0a-a766-a8bb1b413897 · outbound

This paper cites Speechtokenizer: Unified speech tokenizer for speech language models, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speechtokenizer: Unified speech tokenizer for speech language models, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.046810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.869270Z digest=sha256:7c1bab29bd3500db619157c2b0726482f7326ac101525320c32dae36c0c3675c

Observation e9b5eec0-f331-46e2-bc64-4fc7626ba157 · outbound

This paper cites EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.872691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.872691Z digest=sha256:369a7011790947a95bbc719c4c70e3073c3652e0474af06a03efc1d949ed0a53

Observation bed212dd-405f-48ee-b3e1-43f298d8d68e · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.034625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.877279Z digest=sha256:7535d73d1a30e67993c1b84fb2a2dbc52d98dd53a2235a3cc03beb7976300475

Observation 537878f9-1104-4be0-ac14-303447a336c2 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.021563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.880990Z digest=sha256:675ed0e69038331af89475bafef7ebfc60d5e2aa5e0ce2751eb651fb8aca647c

Observation adbe8dce-c156-489a-9358-81100f6ab823 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.005552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.884500Z digest=sha256:5d9ba77020a5158c79e501ae90a71a20d4035daa577563175642a41fb42f4ac6

Observation 180a1b52-9edf-4be0-af4c-74965d9361a1 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.993214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.888908Z digest=sha256:58d2f200eeea46e8093295b3e77a0f69c704a679647b0d18915fd5b99eb8425a

Observation 53797734-b0a1-43f3-9f72-245a532149a3 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.982324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.892402Z digest=sha256:cc9e911c4df9b9f06208612713d36403fd41752cf03c65507e63c4d43c065da2

Observation a563ca93-52e9-4dc4-90a4-6e4250b80133 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.970852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.895977Z digest=sha256:79560a8465a069b3b000156746631468bc26500418d4f6d23587c1f50b6ef50f

Pith citing papers

Observation 72db793a-a47a-4f04-b638-14853299fecc · inbound

MotionPersona: Characteristics-aware Locomotion Control cites this paper.

MotionPersona: Characteristics-aware Locomotion Control It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T12:16:44.518506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:16:44.323653Z digest=sha256:58a0e3da52227c48f2df64e345cccf75836d67f64dd8b52c768cefddadd5e031