Pith. sign in

Paper Citation Record · LEDGER

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

As of 13 August 2026, this Paper Citation Record lists 78 of 78 outbound references and 1 inbound Pith citation observation for arXiv:2412.02419.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02419 v1

Coverage vector

measured 78 of 78 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:34:10.895977Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:16:44.323653Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:16:44.447784Z

Reference resolution

78 of 78 outbound references displayed

  • verified exact0
  • verified fuzzy60
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ca2d14f-68e6-416e-8beb-072713cd51e4 · outbound

This paper cites Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style transfer for co-speech gesture animation: A multi-speaker conditional-mixture approach

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.790910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.605320Z digest=sha256:fbe621da13922e6dc14badebafd0354f8013751812f5eb476bf1ca13f1eb7d48

Observation 5915c4ce-a3a8-4a78-9186-f1c5c2c6e941 · outbound

This paper cites Style-controllable speech-driven gesture synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Style-controllable speech-driven gesture synthesis using normalising flows

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.780191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.609753Z digest=sha256:a288b1a759f68d60cee808ab1a3c0432b3d1cc8346bce32c507aedeee0861044

Observation 23fa6aa1-13a1-4474-81cc-37bca8c711a9 · outbound

This paper cites Listen, denoise, action! audio-driven motion synthesis with diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Listen, denoise, action! audio-driven motion synthesis with diffusion models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.769829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.613798Z digest=sha256:85a3f08e6f8a914a7e9050a3776a23f0a84574551608ef54d528926742af3703

Observation 56da8f1c-79f6-436a-be1c-d6fb598e5c18 · outbound

This paper cites Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Rhythmic gesticulator: Rhythm-aware co-speech gesture synthesis with hierarchical neural embeddings

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.758838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.618075Z digest=sha256:f0cf1d4e3b5d9c08ea534f2a587b2b9db08d7c4b40d27d26fd25fe4b73e8962a

Observation 20fae4cf-0e19-4bf7-a8e5-5b3370613036 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.747711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.622171Z digest=sha256:9105d21e4c1e767f6752fe06b72733e2c8e67758cf33b3e2575ff53fa8b2418b

Observation 8cb414e2-3ad3-4685-80cf-eba842005ca4 · outbound

This paper cites Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesturediffuclip: Gesture diffusion model with clip latents.ACM Transactions on Graphics (TOG), 42(4):1–18, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.736630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.626457Z digest=sha256:f07dce54b45694077da87e2be0ab6518a8ba40601da5d55b7bf0529bcc0460ba

Observation b397dc9e-a53d-4c1d-9d7f-a9140d0dd09e · outbound

This paper cites Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech2affectivegestures: Synthesizing co-speech ges- tures with generative adversarial affective expression learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.725245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.630449Z digest=sha256:eca1214612e2670a0bb5533306ec9573500df8b5102a4222fbf02736d46d3dec

Observation 6431bad8-6118-4472-876a-11ca2642cacd · outbound

This paper cites Digital life project: Autonomous 3d characters with social intelligence.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Digital life project: Autonomous 3d characters with social intelligence

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.711947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.634808Z digest=sha256:7dcc27a22d40f8bd4588175c137ef2888c0d7ab058ebd033047c3b85235ed4cc

Observation 7e103cd7-d7ae-428b-84aa-57f58c0f276f · outbound

This paper cites Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech-gesture mismatches: Evidence for one underlying representation of linguistic and nonlinguistic information

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.700250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.638849Z digest=sha256:4d8d15d38dc78dec8f1f312ead0f3b7381e9753d27bd0f8a415b8cacb43870fd

Observation cfb72b07-3375-4e42-8bd1-870df924531b · outbound

This paper cites Beat: the behavior expression animation toolkit.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Beat: the behavior expression animation toolkit

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.687482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.642619Z digest=sha256:c99d5970c72607743b72bf317624e06a213ba3081d51d72dfb1ee0da0b06eec5

Observation 85d61874-e5c6-4922-96d3-b600001ea7ed · outbound

This paper cites Taming diffusion probabilistic mod- els for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Taming diffusion probabilistic mod- els for character control

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.674539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.646002Z digest=sha256:44157c782a6800803dc6e14b5b7e607a4152da99afc7008209f99717f15b6f52

Observation 02990db2-beda-4498-86b3-5acf292fe63c · outbound

This paper cites Executing your commands via motion diffusion in latent space.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Executing your commands via motion diffusion in latent space

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.662865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.649538Z digest=sha256:dc1246e80206e2d3e0f317da027654c16fe06e3466e27fa1a4ce379da9f13cc2

Observation 3999d157-0ad4-48c8-8792-b022fbee905e · outbound

This paper cites Black, and Timo Bolkart.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Black, and Timo Bolkart

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.651114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.652577Z digest=sha256:1db9052253862b86a97a85ffae1b3f6aab83470def23eac65d02af45eae5e8d7

Observation d942fe81-7f19-4eeb-8422-60fc06bfc535 · outbound

This paper cites The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The interplay between gesture and speech in the production of referring expressions: Investigating the tradeoff hypothesis

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.639164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.655939Z digest=sha256:e494fbb32bc8c318cb8c4b563720e2758ea21ebda2ff73c64f2155dea65fedca

Observation 8d99c66d-a64f-4890-9654-6d40300cbef8 · outbound

This paper cites Diffusion-based co-speech gesture genera- tion using joint text and audio representation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Diffusion-based co-speech gesture genera- tion using joint text and audio representation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.658853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.658853Z digest=sha256:08032fef5a661ff388ef57f2420bafac2a0af380e191c55d024eb0a5e9740a4d

Observation 3abea151-4294-4c25-b01d-f7eaf2b334c2 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.662380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.662380Z digest=sha256:ab14f1fa2cc0909e9e56e6331534f0f30d5fd12b8d5b4e7f7c2985820da932d7

Observation d96f74c4-816d-45e0-8b73-21673e297df5 · outbound

This paper cites Freemotion: A unified framework for number- free text-to-motion synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Freemotion: A unified framework for number- free text-to-motion synthesis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.618609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.666246Z digest=sha256:a1f148504ef1b80e9ae7dba0d021cea85cd2e1be2b6bb394d14c0648993b1bcd

Observation 1a8b866c-d5e1-4e38-bcd5-f656c1de7739 · outbound

This paper cites Troje, and Marc-Andr ´e Carbonneau.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Troje, and Marc-Andr ´e Carbonneau

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.606790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.669724Z digest=sha256:f45a410b697a6c2eb14e8391b788739a77617d4140e8dd850fb760782b9cf761

Observation 41df565f-9d18-483f-93ee-3c04dfccf4db · outbound

This paper cites Remos: 3d motion- conditioned reaction synthesis for two-person interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Remos: 3d motion- conditioned reaction synthesis for two-person interactions

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.595162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.673292Z digest=sha256:8ce185bf43f975dd492bab602a64b4582f56eb56cffe77030b614df9540344dd

Observation be8d6683-299f-4191-ad28-cc5dbb6023c7 · outbound

This paper cites Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction mix and match: Synthesizing close interaction using condi- tional hierarchical gan with multi-hot class embedding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.582801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.677265Z digest=sha256:7c1e545cc4b99ac985697a30744dfbb9d3cac2457205fc13358a302dc80a4439

Observation f65e4349-e081-4312-bacb-6d623a6315ab · outbound

This paper cites Learning speech-driven 3d conversational gestures from video.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning speech-driven 3d conversational gestures from video

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.571181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.680789Z digest=sha256:7a9bcdb5d4214e63caab407b4f1f8e0c350a6add229a2a1b3913c364610601f4

Observation 439a4759-23c4-4b32-8f36-fc848410ef8a · outbound

This paper cites Evaluation of speech-to-gesture generation using bi-directional lstm network.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluation of speech-to-gesture generation using bi-directional lstm network

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.558985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.684444Z digest=sha256:5ff769a52c9cc6776178808d9ebab28fd2b25e8cdcf98dd70f3449dc6ea2c89d

Observation 6316e41f-a6d5-4520-957f-96fb48617ede · outbound

This paper cites Moglow: Probabilistic and controllable motion synthesis using normalising flows.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Moglow: Probabilistic and controllable motion synthesis using normalising flows

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.547685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.687870Z digest=sha256:8c98f82ea96d442f641c565f565a29df1971e0e4858e746d03bccf620ef57fe4

Observation 6880d816-5fdb-4531-a03d-3c4ebe8d07d4 · outbound

This paper cites Denoising dif- fusion probabilistic models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Denoising dif- fusion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.536260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.691524Z digest=sha256:c87cf1d8681674e0d9ee469b2ea1a31252af27c692f0fafd4f6f0794200accde

Observation 3d47a550-f56b-4174-abb6-ac35337d1631 · outbound

This paper cites Dead blending, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Dead blending, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.522980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.695463Z digest=sha256:e6355d2ee56418a38aac83c914fb0026f4c1e684d09394fcaa06d0bf18924ac7

Observation cc08d7be-1e85-419d-9115-5f77e7e8c519 · outbound

This paper cites Phase- functioned neural networks for character control.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Phase- functioned neural networks for character control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.511947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.698968Z digest=sha256:67179e98cb4d914673d45b5ddd48030756683efd4e4695cdb50fd798ec005178

Observation 2b0e30cf-8f77-4624-a3e6-e99e4ea8d345 · outbound

This paper cites Example- based control of human motion.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Example- based control of human motion

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.500817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.702423Z digest=sha256:097268a6a5bf922975bc10fac8206a2de0d638e78a3186eaaf77cb647224df83

Observation d1972587-0b70-413a-a0b4-cd858bc3d2e3 · outbound

This paper cites Robot behavior toolkit: generating effective social behaviors for robots.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Robot behavior toolkit: generating effective social behaviors for robots

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.489364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.706354Z digest=sha256:789bb1fe06303f3bd3cfc98915909c79c6e147ea8013413d47102771108b5b72

Observation e76a4e53-20ee-4230-b662-17e640e00c88 · outbound

This paper cites Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interact: Capture and modelling of realistic, ex- pressive and interactive activities between two persons in daily scenarios

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.478397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.710164Z digest=sha256:6d29934a2fa2204c4bda5438b1536912c04b94e9daea5f03124cc681edcebafe

Observation 5975b0a1-b9a4-4922-b180-6af8b50f6144 · outbound

This paper cites Intermask: 3d human interaction generation via collaborative masked modelling, 2024.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intermask: 3d human interaction generation via collaborative masked modelling, 2024

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.467350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.713926Z digest=sha256:3fe3607dc85f680e4a1b2ef19caa84e32a6f1916ef01e5ac917017b4656e05cb

Observation 6a1026dd-d786-468d-ba8d-5105a70365b1 · outbound

This paper cites Synchronized multi-character motion editing.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Synchronized multi-character motion editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.456925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.717512Z digest=sha256:b343f82302d298fb0a9bba03806bc49cd0b1429c5f8cbaeec19755e706d04c62

Observation b151996e-cce2-4bae-b3c5-f427d7b187d9 · outbound

This paper cites Tiling motion patches.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Tiling motion patches

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.446327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.721459Z digest=sha256:63b0dd916fa13875fb23334efff50e604f9cbc2ef859dc932fb24e2c41ce60ea

Observation f4c80f5a-1a47-4127-ae36-802a4891f356 · outbound

This paper cites Towards a common framework for multimodal generation: The behavior markup language.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Towards a common framework for multimodal generation: The behavior markup language

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.435790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.725418Z digest=sha256:23ec2985534f8aa6d8cbce0b925392e2599f0c7a16ead791fd099ad393ff68d8

Observation 3ee368d6-8506-447a-8370-2ed9a34fc80c · outbound

This paper cites Gesticulator: A framework for semantically-aware speech-driven gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesticulator: A framework for semantically-aware speech-driven gesture generation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.424954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.729119Z digest=sha256:3967799b109d62efbc6b2f8bb2aa299a461a56f6a9ffb9edcebd2f48f61e29a7

Observation fe930589-3c4b-4eac-8d7d-633d2cfbad82 · outbound

This paper cites The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model The genea challenge 2023: A large- scale evaluation of gesture generation models in monadic and dyadic settings

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.413976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.733295Z digest=sha256:70e38009ca22c60a500631e20f8f2cbba4283b04e50d2190116ae31d27f3a4ec

Observation db9a7417-ebc6-4442-8197-e2d8f5ffe4ba · outbound

This paper cites Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Evaluating gesture generation in a large-scale open challenge: The genea challenge 2022

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.402783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.737923Z digest=sha256:e7f5cb5279737d00cea8f74f3549a0fb301de7724ddf232e69f58de5bfe61117

Observation 07d536e3-ea5a-4b05-81ed-b0a8cc5e7334 · outbound

This paper cites Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Cross-conditioned recurrent networks for long- term synthesis of inter-person human motion interactions

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.391368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.742001Z digest=sha256:52839efa1d8a76be1f26feef9b5248ed60fa940bfdd138e90e91b368967a9a61

Observation 34fd5287-c8b0-4620-ad91-b95e190c0122 · outbound

This paper cites Two-character motion analysis and synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Two-character motion analysis and synthesis

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.380112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.745937Z digest=sha256:b9906a3b84afef2e45027e745f81289218d68f47af881a7613f59caabaa9e01e

Observation 88e21c0f-5319-46d7-a528-85150a04f76f · outbound

This paper cites Motion patches: building blocks for virtual environments annotated with motion data.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Motion patches: building blocks for virtual environments annotated with motion data

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.369173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.749738Z digest=sha256:5fc17e9a1d818bfd705c331573520128180a1ec1de0ff95356049b3750787be7

Observation eeb58929-60ac-4dcb-bb5f-cb6f417baeec · outbound

This paper cites Gesture controllers.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture controllers

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.358230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.753522Z digest=sha256:4b329273d72df1e31a6ccefc476ec0589891824cb1e1578bc8beed0320ca8765

Observation ce76ea1f-9bcb-4a30-83b7-e9f25ee0032a · outbound

This paper cites Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Audio2gestures: Generating diverse gestures from speech audio with conditional varia- tional autoencoders

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.346673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.757026Z digest=sha256:42905859c32d054892feee57fd2b39de2eb4d84c2759315a0d9240253442b2c5

Observation 33e9020d-0cf5-4ca8-b632-16a7e2524a53 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.335671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.760840Z digest=sha256:9397055b6121b2827a3c8161da6db03629953852536221b9d2ce31d5a50c4849

Observation 66f12494-a44e-4e61-aecc-0c913f74421f · outbound

This paper cites Intergen: Diffusion-based multi-human motion genera- tion under complex interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Intergen: Diffusion-based multi-human motion genera- tion under complex interactions

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.323941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.764298Z digest=sha256:516bdc1df61dd39ed63ef3f6be77d31f7022a1c225c01c1f7a6c234d22631cbd

Observation b7ec29bf-bd6c-4ff3-87d4-1650789ade19 · outbound

This paper cites Com- position of complex optimal multi-character motions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Com- position of complex optimal multi-character motions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.311617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.768474Z digest=sha256:e60d5c1a38bb4a1b07560d196ad7e7820d601bd4d83d6195e47abdf851eaef8e

Observation ec01dcc6-bd0c-4b0d-a6d3-caf1a887e39e · outbound

This paper cites BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model BEAT: A Large-Scale Semantic and Emotional Multi-Modal Dataset for Conversational Gestures Synthesis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.771875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.771875Z digest=sha256:637b2c917d0a241e72ee3d2060bd002b93bd8703af763058a5e6b335a046de76

Observation 2ed6c7a3-56e8-4729-b2b3-af2a888a2c7c · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.300739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.775706Z digest=sha256:27d2fb95ed8c168da9b718d7994c99ac8a40c5d64d32ae97816ffceb016c96c1

Observation 896ea1c6-8479-480e-8f25-f37af429c758 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.289500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.778805Z digest=sha256:92ff4bcc1ac5909da0d17cb4467655f0628b001020d1ffc7ba4906051e5e7dde

Observation 492e9b10-72f3-4075-9c87-15620a75b76e · outbound

This paper cites Learning hierarchical cross-modal association for co- speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Learning hierarchical cross-modal association for co- speech gesture generation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.278004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.782159Z digest=sha256:b579e95c64179041c3ff0d0658392cda91301c97a070f315c65c8c4cc568af0b

Observation 87d76ea7-3213-4a20-9f4d-760622c75bc4 · outbound

This paper cites librosa: Audio and music signal analysis in python.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model librosa: Audio and music signal analysis in python

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.265161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.785431Z digest=sha256:3511a6363000ef3f13c104299cedab5d0779aa8c2a3d85e60ac257847cca2a2b

Observation 841a0d39-cf2a-4aa3-b837-0cdeb2fc7b2f · outbound

This paper cites Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gan-based reactive motion synthesis with class-aware discriminators for human–human interaction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.788440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.788440Z digest=sha256:654f5c81f1028d5d0317879d7e91b4d0a9f2b6f654a832d09d95bda238156dfe

Observation 3d7bf52f-16fe-4278-9957-e82fc9ed806d · outbound

This paper cites Gesture modeling and animation based on a proba- bilistic re-creation of speaker style.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture modeling and animation based on a proba- bilistic re-creation of speaker style

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.246159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.793221Z digest=sha256:5d8d5185d6234d8618b198a68b9cd175acc2d67200843c4ed74c0a9c554a4c6b

Observation dbacaf9f-79b6-4c46-8db8-274cf3022c59 · outbound

This paper cites From audio to photoreal embodiment: Synthesizing humans in conversations.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model From audio to photoreal embodiment: Synthesizing humans in conversations

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.234149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.797880Z digest=sha256:8c418a1952bfe3576908ed1b0cc71d8397ed7f66763dac4fb96e2f5bc65210a7

Observation 6b842028-841c-4509-abbd-437d226e06e3 · outbound

This paper cites Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Glide: Towards photorealis- tic image generation and editing with text-guided diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.221478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.801777Z digest=sha256:cff39deedc4894f1a4bd828bc12ebd3263e3bbdf65f9b008181c6b4b2842cdbb

Observation 2cb3c7b4-ee91-49dd-a6fa-dd0cbacbb2e0 · outbound

This paper cites A comprehensive re- view of data-driven co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model A comprehensive re- view of data-driven co-speech gesture generation

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.209841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.805434Z digest=sha256:183c0e7716fa3642ee6d6c270f4b439e0d1f4c59efa7722b645962699865a316

Observation a6cdc883-08ad-40df-8e89-2e719926fbe7 · outbound

This paper cites Librispeech: An asr corpus based on public do- main audio books.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Librispeech: An asr corpus based on public do- main audio books

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.197919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.809497Z digest=sha256:d3f7c6c459b760c0652f89882afd009f9ca082575f5efa5a2acdcf886aedf053

Observation 96373b2b-9d16-4261-be07-a9b8a5389ead · outbound

This paper cites Bodyformer: Semantics-guided 3d body gesture synthesis with transformer.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Bodyformer: Semantics-guided 3d body gesture synthesis with transformer

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.186332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.813195Z digest=sha256:83b004fb8497b785fcb21821b9d9d9904cce1e6ab8f568db48b9779a5c5e222f

Observation 06947c17-cb57-4ad1-a79e-9497e1f20f13 · outbound

This paper cites Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Do people use lan- guage production to make predictions during comprehen- sion? Trends in cognitive sciences , 11(3):105–110, 2007

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.175473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.817013Z digest=sha256:2130ac466e479f01a03bb56fdb23a748423913e898a87efcf238ba9571683cc3

Observation 9608dfc9-9723-4269-bd87-dc81c22c86ef · outbound

This paper cites Hierarchical text-conditional image gener- ation with clip latents, 2022.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Hierarchical text-conditional image gener- ation with clip latents, 2022

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.820676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.820676Z digest=sha256:b3dad0b40c47752a1dbff2e8e86df5b9e7dc5e6bd3e7180e5c098508b8209d0c

Observation 320c13ba-d1ba-42b0-9e1f-905dc283375d · outbound

This paper cites Human motion diffusion as a generative prior.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffusion as a generative prior

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.824254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.824254Z digest=sha256:3745ab9b81db75217e1a0097bbc9a5507c0c24ba3a845badb110a795edb78ed9

Observation f186f572-c4df-4ef5-9c3c-59cd0b75dff6 · outbound

This paper cites Interaction patches for multi-character animation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Interaction patches for multi-character animation

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.150778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.828174Z digest=sha256:3fb77804807d5fd8384fc5e74cf2fb3da1182e20da01fa23eb0e3fc07d1fcc00

Observation 0730844b-9f8c-4d86-bf71-2a7a2509860e · outbound

This paper cites Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Duolando: Follower gpt with off-policy reinforcement learn- ing for dance accompaniment

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.140373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.831698Z digest=sha256:425c4d75a15dd5f27ced9d6c9540f9895229c7b28aaaaf456c51be23f80393f4

Observation 025536bc-6071-4b24-b77e-2d7cd30f4927 · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.130124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.835159Z digest=sha256:341ac767d08760a252da360977df96dd4f9c0b0803a4b6cfe23571e9ba417920

Observation 032f199d-881a-410c-a726-83cede472aae · outbound

This paper cites Local motion phases for learning multi-contact charac- ter movements.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Local motion phases for learning multi-contact charac- ter movements

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.118915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.838922Z digest=sha256:5827c80a3a532ba28876865a18fe89ca418ca1cdd8940aead7ddada3d0ecd49e

Observation 1f82c318-2fb5-4e16-8299-1414a0979751 · outbound

This paper cites Human motion diffu- sion model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Human motion diffu- sion model

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.842737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.842737Z digest=sha256:cc34950223652915a406a955b8ee85f5773be186c84d460fce51026e4ebef90b

Observation 8480fd60-50da-43cc-9e08-ab52fbfc4716 · outbound

This paper cites Attention is all you need.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Attention is all you need

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.846546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.846546Z digest=sha256:22498cbb2c6c90a8d8fbe2908b264743515e9cb262f9a9209a1d009066f87216

Observation 9904a46b-a906-44db-bb81-bda659acea85 · outbound

This paper cites Gesture and speech in interaction: An overview, 2014.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Gesture and speech in interaction: An overview, 2014

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.093731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.850351Z digest=sha256:76c11b9d9218a6d22e845c8147fba2d007a386f67cf9f681d59ed83a843f1033

Observation 0cfc6845-f777-4876-9fc4-187caeb14324 · outbound

This paper cites Generating and ranking diverse multi-character interactions.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Generating and ranking diverse multi-character interactions

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.082699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.854232Z digest=sha256:3e5ce5cb7e757f3c8a65722038e6200085b9734b253dd76c5c11700981fbd19e

Observation 04ebfbf7-9c3e-4e2b-9f7b-4cad06788d4e · outbound

This paper cites Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Eggesture: Entropy-guided vector quantized variational autoencoder for co-speech gesture generation

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.070716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.858168Z digest=sha256:df9cc668725104b720657c3b8079fc337af1ba039cc12fc52760c91de23bb10f

Observation 1857373f-ab58-4012-b90b-9c575ef60273 · outbound

This paper cites Speech ges- ture generation from the trimodal context of text, audio, and speaker identity.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speech ges- ture generation from the trimodal context of text, audio, and speaker identity

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.058176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.861661Z digest=sha256:b7a551cc35ce8d5f20647cd91c0cc5b57cfd367c0373e196fa830c0db63ad61f

Observation 6c8e1412-677c-4b5e-a378-08917eb488c0 · outbound

This paper cites MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model MotionDiffuse: Text-Driven Human Motion Generation with Diffusion Model

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.865334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.865334Z digest=sha256:d09b751b06f53158a036340ecd81a03b6ae8a62d8045168e82a5ceb423443c4d

Observation b08a51bb-227b-4d0a-a766-a8bb1b413897 · outbound

This paper cites Speechtokenizer: Unified speech tokenizer for speech language models, 2023.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Speechtokenizer: Unified speech tokenizer for speech language models, 2023

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:34:11.046810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.869270Z digest=sha256:d728923a8478e16697cd4f0255e8c966d9a897dca7254c00193558abb76965eb

Observation e9b5eec0-f331-46e2-bc64-4fc7626ba157 · outbound

This paper cites EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T23:34:10.872691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:34:10.872691Z digest=sha256:081a3c82aec0628a063a5a73fff1bd098d28c754fd1fc1f0057825238b8a8d26

Observation bed212dd-405f-48ee-b3e1-43f298d8d68e · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.034625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.877279Z digest=sha256:8df93c88d42e31b1e13a011cf2136315891a3d6544c95f2307a6e247828cf70d

Observation 537878f9-1104-4be0-ac14-303447a336c2 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.021563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.880990Z digest=sha256:40fe98dd12303359eeca6fc39671d703d2e9efc30c569bec8c6e7fe39facc7e0

Observation adbe8dce-c156-489a-9358-81100f6ab823 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:11.005552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.884500Z digest=sha256:b9df745c17e9f8050bf8d3c86d463824c3111a6c5461a26c5cd74178faf6c436

Observation 180a1b52-9edf-4be0-af4c-74965d9361a1 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.993214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.888908Z digest=sha256:4b45886c7233a91a9a95467a3eeba301515c4ea5b7478f9b159399a5828c215b

Observation 53797734-b0a1-43f3-9f72-245a532149a3 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.982324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.892402Z digest=sha256:e604ffec7b235ccabc4328d25ab2b0e5cc3d5f6cbb87d093dcaf8f66b3bdf4fa

Observation a563ca93-52e9-4dc4-90a4-6e4250b80133 · outbound

This paper cites an unresolved cited work.

It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:34:10.970852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-11T23:34:10.895977Z digest=sha256:4971e3075dd1cf21841f71d379b3b4fa2e8e908978495bd2dfee9a9566db8903

Pith citing papers

Observation 72db793a-a47a-4f04-b638-14853299fecc · inbound

MotionPersona: Characteristics-aware Locomotion Control cites this paper.

MotionPersona: Characteristics-aware Locomotion Control It Takes Two: Real-time Co-Speech Two-person's Interaction Generation via Reactive Auto-regressive Diffusion Model

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T12:16:44.518506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T12:16:44.323653Z digest=sha256:7572e3e65280dbfaf01d83de5fb7970eaf35db5c0015fe4eab2c08d86abbdd7c