Pith. sign in

Paper Citation Record · LEDGER

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

As of 7 August 2026, this Paper Citation Record lists 100 of 122 outbound references and 4 inbound Pith citation observations for arXiv:2505.18078.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18078 v1

Coverage vector

measured 100 of 122 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:40.471781Z

measured 104 of 104 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:53:27.675148Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.170733Z

Reference resolution

100 of 122 outbound references displayed

  • verified exact2
  • verified fuzzy17
  • unresolved81
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2900a376-f8b4-4b1b-99e2-9c70f8521af7 · outbound

This paper cites Kling ai: Next-generation ai creative studio.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Kling ai: Next-generation ai creative studio

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:27.854249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:27.854249Z digest=sha256:22add4c9d8ae2c24d9406491197c79893644d9889f0efe8208edc877747c1bf0

Observation 1168006f-b201-49e8-9707-34b34f9389b7 · outbound

This paper cites NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:27.962933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:27.962933Z digest=sha256:0e47ff5377bf0d3fb2f63210b848bbd14804f38f9001ef6a51e95e1d04a427d4

Observation 600e7f26-f3fa-4eeb-b067-daa256352da9 · outbound

This paper cites Evaluating multiple object tracking performance: the clear mot metrics.EURASIP Journal on Image and Video Processing, 2008:1–10, 2008.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evaluating multiple object tracking performance: the clear mot metrics.EURASIP Journal on Image and Video Processing, 2008:1–10, 2008

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.058069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.058069Z digest=sha256:072a6c603addd9a319c2cd7ae1823f0f3d80f6ba807f4c5d62b4a786d0437518

Observation e47415ba-09ac-47d7-9104-881ef902dfbf · outbound

This paper cites Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.179162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.179162Z digest=sha256:674a4822c2a60e5294f83f023b66cac20794747281c269fa3920dd020e79defa

Observation df3c0f70-42af-4cd7-8b4c-70e5deb96fb6 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.316171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.316171Z digest=sha256:5f148744b7b03e247b2dbfee829879b17bb83ef58bed50a04ee13e13707ee846

Observation 957d04fc-d08e-4f4d-8e24-11439fcae46a · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Align your latents: High-resolution video synthesis with latent diffusion models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.432698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.432698Z digest=sha256:ab5c0169990a87ea8a6d9853cb7952116b535e0d34a370cb52279831d297f0a6

Observation cdf151f3-4e91-4d33-902e-62499e8bd14a · outbound

This paper cites What Are You Doing? A Closer Look at Controllable Human Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation What Are You Doing? A Closer Look at Controllable Human Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.537052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.537052Z digest=sha256:734883252b368bb519ed2503a0ed581bf9efe602cfbff997a174b087c3c71760

Observation b2f98aa9-0a88-4e9a-9f9a-d44c9902f925 · outbound

This paper cites Deep video generation, prediction and completion of human action sequences.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Deep video generation, prediction and completion of human action sequences

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.702062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.702062Z digest=sha256:1b8fd4ad646298703bd4fcab6a2b69fd032a044438cad0d53aa92593a0cd18fd

Observation c227ce5b-e088-42c4-b987-b010f71ce6e2 · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Realtime multi-person 2d pose estimation using part affinity fields

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.910813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.910813Z digest=sha256:68f2bfb76cc912f46f58ea75a6c0322eeaff213b2e5f595ff622dbd436118aab

Observation fe4bec42-cecc-420a-8afe-80132b8a15a3 · outbound

This paper cites Everybody dance now.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Everybody dance now

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.012468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.012468Z digest=sha256:f89511ba4013abda948859a14959fac6e39dbe8d21c618990f96bd7c2cefbc5b

Observation a11af6a5-570e-4f38-841b-c9cc3c9f17a0 · outbound

This paper cites A survey on evaluation of large language models.ACM transactions on intelligent systems and technology, 15(3):1–45, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation A survey on evaluation of large language models.ACM transactions on intelligent systems and technology, 15(3):1–45, 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.121664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.121664Z digest=sha256:adf6c363c91f79606628a00b3d523c07dc7673f4321b299e1003a3543995e856

Observation 44b5296a-29b9-49a1-b2ca-279347e69f5a · outbound

This paper cites Idea23d: Collaborative lmm agents enable 3d model generation from interleaved multimodal inputs.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Idea23d: Collaborative lmm agents enable 3d model generation from interleaved multimodal inputs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.276838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.276838Z digest=sha256:d2fa269e1c024873e8c10bb993b935eecbed28fa1c8288faad90b809f6281707

Observation 5ee3d724-3094-49fa-ae27-139edab772e8 · outbound

This paper cites Ultraman: Ultra-fast and high-resolution texture generation for 3d human reconstruction from a single image.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Ultraman: Ultra-fast and high-resolution texture generation for 3d human reconstruction from a single image

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.431038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.431038Z digest=sha256:5fc437e5823313e85bd3616b417ae52abefcf60054408073816af58fcdfcd058

Observation 0787fcbb-8a67-45a3-bacf-9a7be07ce9da · outbound

This paper cites Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.576107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.576107Z digest=sha256:a47f4e3ec6db630e8f2d7d2b201c5259efd980fcc52824a83c611362cbe2a840

Observation c6267999-d86e-4c8f-be8a-773bfa80472c · outbound

This paper cites Control3d: Towards controllable text-to-3d generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Control3d: Towards controllable text-to-3d generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.716317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.716317Z digest=sha256:6812f46bc58c451b7451a821e365177d10d79c9512741218714ea0c5311a9845

Observation 02698d6c-1116-4086-82a6-f5fe84d2b928 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object understanding.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Abo: Dataset and benchmarks for real-world 3d object understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.878018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.878018Z digest=sha256:901d930bdc159b07a1427943ada74039b85ea3388e73f90339d6984c891d79c4

Observation ef74ef35-882e-452e-8674-9c4390b03984 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Arcface: Additive angular margin loss for deep face recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.971259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.971259Z digest=sha256:3559fd4600fd490e1e270491a9b5be107b0c8264b28f626972a93415e8caa74b

Observation e60f2730-3902-4bdd-b390-6833cb5fff33 · outbound

This paper cites MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.078011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.078011Z digest=sha256:4c9aecdc2f9ad8ffa9879004355a41c606451ddbf59d32ea1456f1a0fb13fa52

Observation 267dda78-9c90-4817-b5f6-b6137d6b873f · outbound

This paper cites Image quality assessment: Unifying structure and texture similarity.IEEE transactions on pattern analysis and machine intelligence, 44(5):2567–2581, 2020.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Image quality assessment: Unifying structure and texture similarity.IEEE transactions on pattern analysis and machine intelligence, 44(5):2567–2581, 2020

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.203373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.203373Z digest=sha256:42d03438969561c3c3d17bf4e19ece258d3d7bb3b61f53d22f1aa0037d3a52a4

Observation 3a8d5778-848c-46ab-8ac9-064bf19ec487 · outbound

This paper cites A survey of embodied ai: From simulators to research tasks.IEEE Transactions on Emerging Topics in Computational Intelligence, 6(2):230–244, 2022.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation A survey of embodied ai: From simulators to research tasks.IEEE Transactions on Emerging Topics in Computational Intelligence, 6(2):230–244, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.357678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.357678Z digest=sha256:4a40b83796d4da5b36ae90ddf3cadd16caffff4c85b4d762b573aa7e02a9109f

Observation 7e3136c8-8dd5-40fd-abf9-f1c597ca576e · outbound

This paper cites DreaMoving: A Human Video Generation Framework based on Diffusion Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreaMoving: A Human Video Generation Framework based on Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.469507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.469507Z digest=sha256:23f12296a34d9e3eb5832b1427eb77ae01b8e533256d526425960489233414e2

Observation a156db48-51d7-49dd-973d-9dd3e69652f7 · outbound

This paper cites Reconstructing Three-Dimensional Models of Interacting Humans.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Reconstructing Three-Dimensional Models of Interacting Humans

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.578026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.578026Z digest=sha256:fcfc9adc37fb72926fc2b9309cf5f7b7d8f32431783d05074309eb9976bc68f0

Observation 24d1dbb2-82b1-4de1-a26f-aa50f20d428f · outbound

This paper cites Iw-bench: Evaluating large multimodal models for converting image-to-web.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Iw-bench: Evaluating large multimodal models for converting image-to-web

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.706731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.706731Z digest=sha256:582d50c6384aee2625c3c756803db072bab34fb49d7990325466f7a214020eed

Observation a5f23d6e-e533-4821-bd2f-20bd4b8923be · outbound

This paper cites REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.852424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.852424Z digest=sha256:fd86788f8e9cae33cfa8add0b821a318a43ac655fc92980a3f46f45599b2f7de

Observation 3a2db74a-06c7-4702-a933-436efdfd1d1a · outbound

This paper cites Controllable video generation with sparse trajectories.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Controllable video generation with sparse trajectories

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.000337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.000337Z digest=sha256:bf27dedb0748237d64c1a2b2a6f724ad18f121e30a778041ec85764a723fe7bd

Observation 966235b4-d245-45a2-adbe-c8cce4d0a6cd · outbound

This paper cites Clipscore: A reference-free evaluation metric for image captioning.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Clipscore: A reference-free evaluation metric for image captioning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.077064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.077064Z digest=sha256:a81439f447374b02a673b4547d95821a229e53226919fc341dd4ea84c487cd82

Observation 2ac61a93-afe7-4117-a40d-0a7215b9d292 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.195892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.195892Z digest=sha256:100f62599b8097883926b3925c23359bd1f6f445d5469a9a59f8ef3ef2be4cb5

Observation 19b36b3b-9e7e-4807-b977-68aadc3a9386 · outbound

This paper cites Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.330830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.330830Z digest=sha256:4a6cdce2f2d51a1c375c2b4f1af82a16b3c0556e524725f355d0451825d5d0a1

Observation 3a338dd0-f641-4126-ad46-10370e14664c · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.448865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.448865Z digest=sha256:fe69689ec70479edbc01c814ef72abb59b07c9874e89a2792ba43e1c166a1bab

Observation 4bdc7f19-0ce8-4b8d-a1cd-666550d4a03d · outbound

This paper cites Make it move: controllable image-to-video generation with text descriptions.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Make it move: controllable image-to-video generation with text descriptions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.551390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.551390Z digest=sha256:e814f241cff9279504b85f52a5af47139d35e79de7cc6a2043d25dc3941e15a6

Observation 148131a0-b0cb-4168-8428-04eda8d52060 · outbound

This paper cites T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation.Advances in Neural Information Processing Systems, 36:78723–78747, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation.Advances in Neural Information Processing Systems, 36:78723–78747, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.697935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.697935Z digest=sha256:dcf8f62fc4ffee40083e8951b98d3a45bcf03d4df510756e4f287349e5cc5c50

Observation 7c81c201-5c26-4245-a1ca-67ffbeecd471 · outbound

This paper cites Learning high fidelity depths of dressed humans by watching social media dance videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Learning high fidelity depths of dressed humans by watching social media dance videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.816062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.816062Z digest=sha256:87eebac48069782dbf8de0847a2091c596d048bcc25896203933d57b9832a49d

Observation d2601d8f-14fd-4923-a6b3-ebbbfd91843f · outbound

This paper cites Yolo by ultralytics.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Yolo by ultralytics

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.952185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.952185Z digest=sha256:f49adf78fbda1b28b682cdd943ab06be9a049521a947cd975328b767da0f3f60

Observation 7acefeae-4451-407b-95f2-9d10240b0820 · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Dreampose: Fashion image-to-video synthesis via stable diffusion

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.041557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.041557Z digest=sha256:07300b692ad43f73e6247171da9439028ae4d0ad080d8341c800604cf034c96f

Observation c32c5365-7c35-42a8-a1d0-3d0489a46a4a · outbound

This paper cites Text2video-zero: Text-to-image diffusion models are zero-shot video generators.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Text2video-zero: Text-to-image diffusion models are zero-shot video generators

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.189957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.189957Z digest=sha256:5a8ed3dc25166b274cee28d353cc252dfe50a481431fd0cafc24bdfc819724e2

Observation a70eeb23-462f-4a78-8dd7-4da9da819357 · outbound

This paper cites Harmony4d: A video dataset for in-the-wild close human interactions.Advances in Neural Information Processing Systems, 37:107270–107285, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Harmony4d: A video dataset for in-the-wild close human interactions.Advances in Neural Information Processing Systems, 37:107270–107285, 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.335868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.335868Z digest=sha256:f7f335e875b69f4e2fff21ce6ab7a9b8c352016ff233e4226745a0a7512fad90

Observation 27af0b82-7793-48ce-826a-09120dbb2c87 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.441384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.441384Z digest=sha256:e814a551249bac21c7bc870928c148944c7c37306e656f397afd2798bb393ba0

Observation 0f9e282f-c51b-434d-9c89-643ac36b716c · outbound

This paper cites World Knowledge from AI Image Generation for Robot Control.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation World Knowledge from AI Image Generation for Robot Control

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:44.390335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:32.544278Z digest=sha256:c3530c599396e8caf96139b8d8b188b1ca901a189ec36ea9dd5b78b9ad4586a1

Observation e9fad175-df89-4ddd-9db3-16588d30dcb2 · outbound

This paper cites Collaborative video diffusion: Consistent multi-video generation with camera control.Advances in Neural Information Processing Systems, 37:16240–16271, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Collaborative video diffusion: Consistent multi-video generation with camera control.Advances in Neural Information Processing Systems, 37:16240–16271, 2024

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.667049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.667049Z digest=sha256:9a0046e3c0d33c65a3968d44707d5e51593c7ce8048f5b7855247227412dac29

Observation 71d895b1-7f2a-435b-95e4-ea3e890f6c3f · outbound

This paper cites Interactive control of avatars animated with human motion data.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Interactive control of avatars animated with human motion data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.784501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.784501Z digest=sha256:4e8d3c602cd73b31111115b82c5d8b82357f83006ceca4830b9f0975216dc1a6

Observation 78117396-776c-4748-833d-7021096a11d7 · outbound

This paper cites Dispose: Disentangling pose guidance for controllable human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Dispose: Disentangling pose guidance for controllable human image animation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.929080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.929080Z digest=sha256:c7e098f21b23dd54a6dda25173d772cbd003d807792dc19feb85bd5188184355

Observation 1cea47e6-c866-4cac-9f51-982f745f81b1 · outbound

This paper cites Magicmotion: Controllable video generation with dense-to-sparse trajectory guidance.arXiv preprint arXiv:2503.16421, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicmotion: Controllable video generation with dense-to-sparse trajectory guidance.arXiv preprint arXiv:2503.16421, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.077966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.077966Z digest=sha256:b934e1ea2a69858106a5af2a2af9890b3ca8dc4c104af8081257fa1d9de6f413

Observation d78f8ca3-ebb4-4682-8bc1-a86e74cae192 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.241476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.241476Z digest=sha256:f69e5689290871c2ee9fd8f726a150ae9bfca3e4621b12bf9b474c4621a5ee19

Observation 441d9249-c4c1-4ea9-ad25-9693c9ecd598 · outbound

This paper cites Lodge: A coarse to fine diffusion network for long dance generation guided by the characteristic dance primitives.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Lodge: A coarse to fine diffusion network for long dance generation guided by the characteristic dance primitives

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.415877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.415877Z digest=sha256:1867ae05efe626471e04d7a5c7b7c5d37e9c3b37167f8295d323ffb6b93e3ccc

Observation c8951d2e-04fe-4d05-b8e8-5767a0fe2ab1 · outbound

This paper cites Finedance: A fine-grained choreography dataset for 3d full body dance generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Finedance: A fine-grained choreography dataset for 3d full body dance generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.524998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.524998Z digest=sha256:e8ea64b7bcd8df2ec2e3e18d96c790bca382f46d119a54353e948a74419ce496

Observation ae3446fd-c7f8-491e-a946-312f88190ce6 · outbound

This paper cites Evaluation of text-to-video generation models: A dynamics perspective.Advances in Neural Information Processing Systems, 37:109790–109816, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evaluation of text-to-video generation models: A dynamics perspective.Advances in Neural Information Processing Systems, 37:109790–109816, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.655440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.655440Z digest=sha256:a8b533a92b59ed2d27db0fcdab69f420ed4e9883919a35193ebb27bbebe472f2

Observation eba97441-0e2c-47f4-b193-0346b6d5c10a · outbound

This paper cites Open-Sora Plan: Open-Source Large Video Generation Model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.784524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.784524Z digest=sha256:06f413eb402d3a5ce94c115820114c66ce8bf8843af84d25c2265f6005d56238

Observation 561c3ced-6317-4037-96e2-0760abce90d3 · outbound

This paper cites Microsoft coco: Common objects in context.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Microsoft coco: Common objects in context

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.911487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.911487Z digest=sha256:a260807326e12abf761f3bfbdea83b41a88fe12de9cba6003007d54f7f6f2c1c

Observation 5eecf7a7-32db-4588-a5a7-6b9a37951ff3 · outbound

This paper cites Rich: Robust implicit clothed humans reconstruction from multi-scale spatial cues.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Rich: Robust implicit clothed humans reconstruction from multi-scale spatial cues

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.080061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.080061Z digest=sha256:46306996633135a9dd2d9f8b82cc1856779671e348668e8e90eae5489c50320e

Observation 70b16e33-f959-4153-8202-71429bf008e1 · outbound

This paper cites Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.232980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.232980Z digest=sha256:df8a670ee035477101997e81d9f011f878677c0963144d81470386886af86f29

Observation 7c276a2c-8df3-43f3-ac91-74cc9c55b17b · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evalcrafter: Benchmarking and evaluating large video generation models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.359389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.359389Z digest=sha256:426732e04afb445d1d0ce35a9c2e57e6bbf1b59d1baad696e6bdb9e2156e0c3f

Observation 866f235c-51e3-4a8a-90fd-f5a545b88ed9 · outbound

This paper cites Fetv: A benchmark for fine-grained evaluation of open-domain text-to-video generation.Advances in Neural Information Processing Systems, 36:62352–62387, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fetv: A benchmark for fine-grained evaluation of open-domain text-to-video generation.Advances in Neural Information Processing Systems, 36:62352–62387, 2023

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.496802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.496802Z digest=sha256:e14680f82014678c1cef983879f3cecf9c4f6c4a5726b6555fdbfb00ea48ec32

Observation 55ef3d96-bbc4-445e-a9ff-abcd5bbf4852 · outbound

This paper cites Hota: A higher order metric for evaluating multi-object tracking.International Journal of Computer Vision, 129(2):548–578, 2021.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Hota: A higher order metric for evaluating multi-object tracking.International Journal of Computer Vision, 129(2):548–578, 2021

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.644160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.644160Z digest=sha256:1e3c333a202cc6c43b65962ad83ad06c50eccfb11d9f25c9d2d4e36204bdea00

Observation 700ba6c6-ae3e-4c95-970e-c8e24237ad65 · outbound

This paper cites DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.790939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.790939Z digest=sha256:5e84a35065fa22147999077062bdce62b48d69aca41e67962456e451a20aabe7

Observation 2eecacde-9246-4bbe-862e-d50b89449f80 · outbound

This paper cites Notice of removal: Videofusion: Decomposed diffusion models for high-quality video generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Notice of removal: Videofusion: Decomposed diffusion models for high-quality video generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.892799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.892799Z digest=sha256:036622485aeb27f64d8bfc7d2c8408146e1d4627efbc77e00135a44f51d0d83a

Observation bfc92c3e-e5f4-4101-aed4-172674a16189 · outbound

This paper cites Follow your pose: Pose-guided text-to-video generation using pose-free videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Follow your pose: Pose-guided text-to-video generation using pose-free videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.023802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.023802Z digest=sha256:49e0747879eb26ce462891d9b49e158034083d19f1be10628643cebe105fcae7

Observation da0f7099-f544-417d-8f8a-38e68d11bd91 · outbound

This paper cites Foundation models for video understanding: A survey.Authorea Preprints, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Foundation models for video understanding: A survey.Authorea Preprints, 2024

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.160080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.160080Z digest=sha256:78dc3ba820fc71abe21711c8ae90188333595baabf2204b2ca13ddac9a018de2

Observation 9a1ae84c-733b-48a2-bfd4-2a7e08e13cfb · outbound

This paper cites Synergy and Synchrony in Couple Dances.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Synergy and Synchrony in Couple Dances

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.320236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.320236Z digest=sha256:1198bb1d4a453fcdfa2229877db6d9d38dd94be2c185090738760fd686ec9747

Observation b9996c06-4d85-4b26-a5f0-faf946da36cf · outbound

This paper cites Benchmarking counterfactual image generation.Advances in Neural Information Processing Systems, 37:133207–133230, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Benchmarking counterfactual image generation.Advances in Neural Information Processing Systems, 37:133207–133230, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.790830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:35.420035Z digest=sha256:2ad5f88f727edad450386a0e86f82d84e74c901610135e4409cbac4b40bf1ba1

Observation 1d08ed28-1e54-4507-bcb4-c3558f9515f9 · outbound

This paper cites Efficient motion weighted spatio-temporal video ssim index.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Efficient motion weighted spatio-temporal video ssim index

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.621763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:35.511112Z digest=sha256:dc77496222b0f4c0b8ba06ad1fb591597adb36187373c16d0b6a4d3919921b7e

Observation 599b29c2-6318-4b10-8bb1-c61fe0763ba6 · outbound

This paper cites Sora: Creating video from text.https://openai.com/sora, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Sora: Creating video from text.https://openai.com/sora, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.415565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:35.638475Z digest=sha256:95e8edafe103eacfcd66224d8b8f08b44c3814dec1eefa577cea097e35174ba3

Observation 1c2473c0-16db-4e18-9f5e-9aa1cc2a423a · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.757677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.757677Z digest=sha256:c6b108e470ef82a5168f3f2b3d08fe4aefdb93ed4f1b4c3fb9d16457ebe8f869

Observation 7c9b63c7-f018-40b3-a9ba-7c2120a5f46a · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.202694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:35.914744Z digest=sha256:54c6f0aeb19a22cb1758501ede7414e0f2a7b8d78cfb4224f96c6b577eff951c

Observation fe34d260-c08b-4195-ba96-e4d7d70a2aec · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreamFusion: Text-to-3D using 2D Diffusion

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.051233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.051233Z digest=sha256:f9b2233954b59b2bfccf2b5311fcd50d99ef11507c8f45f97e577a84de9e368d

Observation 8174cd73-7e13-4a04-9ccb-9620d9d92d33 · outbound

This paper cites WorldSimBench: Towards Video Generation Models as World Simulators.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation WorldSimBench: Towards Video Generation Models as World Simulators

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.194252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.194252Z digest=sha256:036b5dc2b2bc2be56a02af6ad1d9a933ce5e47d7121fe964a50077ae2d2bccd0

Observation a2ad7b24-9803-44a6-8f01-9e7a52255402 · outbound

This paper cites Performance measures and a data set for multi-target, multi-camera tracking.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Performance measures and a data set for multi-target, multi-camera tracking

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.977423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:36.300010Z digest=sha256:9645d99bc23694f5f1afe88298856918ff88bb2bbc328f525e22110707011c29

Observation 986d20ab-c7c7-4611-8eeb-ad05a8af723f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation High-resolution image synthesis with latent diffusion models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.412598Z digest=sha256:9af78252a6234a1e540a6d6c2dd0c1fa2ddcda19b14ebfeda74703e195dc9fd6

Observation f6261cbb-7924-47d6-87df-5d0653f45bff · outbound

This paper cites Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.824616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:36.511573Z digest=sha256:60f07c7b0f8cc6fc8a1a1e3512a601b50785a8284357d7503402a9ec7ba62073

Observation 682b075f-1a07-4905-a8b1-4919e9eb09c4 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.624194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.624194Z digest=sha256:9d4ec99f9fe9d4ba63192407008d17dc6bd7e626ad90cc2092802b62e4746fe1

Observation fe4084ab-4cf7-48a7-9c17-4e1a95b2bb29 · outbound

This paper cites Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.773038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.773038Z digest=sha256:b4834e569821ab1d025cd0fee6472eb58f54392fb2d4c06c21c673aac27cd4cd

Observation a735f50a-d9d7-44ff-984d-281da8418352 · outbound

This paper cites Transnet v2: An effective deep network architecture for fast shot transition detection.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Transnet v2: An effective deep network architecture for fast shot transition detection

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.631466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:36.918213Z digest=sha256:dd746220d1b7c2d0c4fe3aea429f44df50bead36dbe0314ab341d422b63171c8

Observation 295ef70f-8597-41db-824d-ae06ac631d1a · outbound

This paper cites T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.046057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.046057Z digest=sha256:0411103850b9a446f86c6878bb866164045ff1021ad5d6dd7877da848641651c

Observation 17dab6e0-1d7f-4180-93a1-3b9ef8666ec9 · outbound

This paper cites Journeydb: A benchmark for generative image understanding.Advances in neural information processing systems, 36:49659–49678, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Journeydb: A benchmark for generative image understanding.Advances in neural information processing systems, 36:49659–49678, 2023

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.148856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.148856Z digest=sha256:1c84afdee86d6c1e89dfa4cae23c23cf8c5406bf958b951bebe15ace24b385e1

Observation 0ab2f963-912a-410f-b6ef-81fe8dfa4e11 · outbound

This paper cites DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:44.064186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:37.284377Z digest=sha256:4bf8a894f2358023365372b2a2a2cc5a1f57225026be20d0f829546503af9b13

Observation a8539f1c-a8e2-4549-adf7-af82da83321e · outbound

This paper cites Beyond talking–generating holistic 3d human dyadic motion for communication.International Journal of Computer Vision, 133(5):2910–2926, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Beyond talking–generating holistic 3d human dyadic motion for communication.International Journal of Computer Vision, 133(5):2910–2926, 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.423338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:37.406670Z digest=sha256:1eff613ec48c25da6a0f6b07a9c84308afdf55b7c70bb44af9afcfa9c7a10771

Observation 462060f9-a4b0-4549-9216-66fbc861a471 · outbound

This paper cites Video understanding with large language models: A survey.IEEE Transactions on Circuits and Systems for Video Technology, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video understanding with large language models: A survey.IEEE Transactions on Circuits and Systems for Video Technology, 2025

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.503517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.503517Z digest=sha256:bba4413b595e67310020dc5180677090e8ca7365d1b1d587c50dee819e160825

Observation 49039b20-9780-4dcf-88f4-a11b31bd5e09 · outbound

This paper cites Human Motion Diffusion Model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Human Motion Diffusion Model

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.631482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.631482Z digest=sha256:0bbc811e9ba33bf7e9634451630becab573fa21f4e6478bbc249662f965d40de

Observation 3db3c4ac-a43b-4926-b51d-162747b2aa91 · outbound

This paper cites EMO2: End-Effector Guided Audio-Driven Avatar Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation EMO2: End-Effector Guided Audio-Driven Avatar Video Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.744863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.744863Z digest=sha256:02e50db49d4e20ae499bf9fb18ded0d94325786b148f084337675953f605fc69

Observation 24dcd086-b1d8-49c5-b27f-026e69609a5c · outbound

This paper cites StableAnimator: High-Quality Identity-Preserving Human Image Animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation StableAnimator: High-Quality Identity-Preserving Human Image Animation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.877008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.877008Z digest=sha256:b155efe1430dc6fe49c9fbfbb6fdfabde912e497d952391e4e3e6304b0836a10

Observation 534ead77-e2ec-49f6-9f7f-096e10dc33a1 · outbound

This paper cites Fvd: A new metric for video generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fvd: A new metric for video generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.009818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.009818Z digest=sha256:b086c8a4a7e404dad6b722c29c144e3b9cbd8d1dd7af98a4cf1caa1fd4fbc3d5

Observation 52b01249-347e-4aab-990a-adfe2027572e · outbound

This paper cites Articulated mesh animation from multi-view silhouettes.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Articulated mesh animation from multi-view silhouettes

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.170717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:38.139105Z digest=sha256:b510dfedaa60d8474a881b6842f41fb626d7100b04e33051b18e1946d36e08a5

Observation dfd18d8f-71fe-406c-b47d-1988373d0415 · outbound

This paper cites This&That: Language-Gesture Controlled Video Generation for Robot Planning.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation This&That: Language-Gesture Controlled Video Generation for Robot Planning

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.240073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.240073Z digest=sha256:20827fa2881a7f5fc501b28bdd556efbf033ea0c4f2848184b9b75156962a68b

Observation e9cfa681-e7ca-4641-be4d-fa5de2ea8ff4 · outbound

This paper cites COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.352118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.352118Z digest=sha256:700335d46800ff2d9e4c171be20f8bd65eb52f0e95876bfd7a21de073355a4a9

Observation e864da19-96b1-4111-ab4c-20661dfd5966 · outbound

This paper cites Taming Rectified Flow for Inversion and Editing.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Taming Rectified Flow for Inversion and Editing

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.465996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.465996Z digest=sha256:a35cec8ab8783ec61fd0dcba45dae0b6ed469eaee1ee10bf2bab203aeba38f35

Observation 565a5ffb-4f3d-4970-8bb6-9b8050c1f3fa · outbound

This paper cites VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.608579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.608579Z digest=sha256:e52390b366506c29ce410f27dd4f73539e1d9b968f3d757258dce946ca29a315

Observation 6119652b-97ef-4568-b9a7-3cfbd9be9051 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Disco: Disentangled control for realistic human dance generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.943110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:38.726642Z digest=sha256:b54c8548576fc94e461cb64da4e4e3479ed38df4de93d3e7e0cc7219c146e9ec

Observation 65f317fb-c8d2-4a1c-beca-ac1f96bba92e · outbound

This paper cites Unianimate: Taming unified video diffusion models for consistent human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Unianimate: Taming unified video diffusion models for consistent human image animation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.700278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:38.845741Z digest=sha256:d9df003028cacd3abaf8a6011a8b976d8702cbaacffba542c903b539e1929ddf

Observation faf8ef46-0828-4759-a6f9-a773ad75e8d1 · outbound

This paper cites Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289, 2025

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.951240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.951240Z digest=sha256:b9172cebd704b9addaef3e022a2fa53695b4dd21cebc9a5286b067b5e2633c83

Observation c8bf9644-9762-4233-9e2e-7c962d2e1edb · outbound

This paper cites Instructavatar: Text-guided emotion and motion control for avatar generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Instructavatar: Text-guided emotion and motion control for avatar generation

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.495127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:39.037038Z digest=sha256:2120faec71551e84b2b02e130a84bdb8fa298ad9b7fb510ef0d8b20768479af4

Observation e6bc5f26-51b7-48df-8410-96fa349b471d · outbound

This paper cites Bovik, H.R.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Bovik, H.R

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.157902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.157902Z digest=sha256:4cfbd515e3e24872625e6b6da1bb099fac043c7666f9969cb905c6a528b72cab

Observation 3ece665e-7fe3-49eb-9d56-486174428730 · outbound

This paper cites Humanvid: Demystifying training data for camera-controllable human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Humanvid: Demystifying training data for camera-controllable human image animation

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.285878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:39.307743Z digest=sha256:f52c092365996bca7eb18894511d7c7dd8d6237753e502e99f8a8809d9eb0210

Observation a2ee6168-693d-40cb-8685-840a0cec5c10 · outbound

This paper cites Multi- identity human image animation with structural video diffusion.arXiv preprint arXiv:2504.04126, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Multi- identity human image animation with structural video diffusion.arXiv preprint arXiv:2504.04126, 2025

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.404847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.404847Z digest=sha256:8d88c21bf70a5023504bf494e1885dca47b1f6f439459808a1224afb62665572

Observation 4a53d9b5-b2d4-4d2c-80f0-ecc04d5e0eda · outbound

This paper cites MotionCtrl: A Unified and Flexible Motion Controller for Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.527828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.527828Z digest=sha256:d56da2d6317362a72afe7f267cd73e33df19fca0913c5ee305f1dce8bd7a3436

Observation 02e6b24b-ea57-4867-9607-8d45d5c52528 · outbound

This paper cites Easyanimate: A high-performance long video generation method based on transformer architecture.arXiv preprint arXiv:2405.18991, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Easyanimate: A high-performance long video generation method based on transformer architecture.arXiv preprint arXiv:2405.18991, 2024

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.651709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.651709Z digest=sha256:2a962f60da663c2d696bdbb4901754e49cb33a6a8f56807ce1fa73b0d17deb1c

Observation 547a547b-feb7-4de0-a803-018f84f6dd5f · outbound

This paper cites Xagen: 3d expressive human avatars generation.Advances in Neural Information Processing Systems, 36, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Xagen: 3d expressive human avatars generation.Advances in Neural Information Processing Systems, 36, 2024

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.028384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:39.799853Z digest=sha256:65413fb5113d4fb1865d04d824f5c637bc936945fa19d948faad1fdbb133956c

Observation 85fc7bfc-08dc-4b08-9ab8-cd81ff7f3f8f · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicanimate: Temporally consistent human image animation using diffusion model

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.844906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:39.935004Z digest=sha256:0be63aefb4bfb764bf0fffc56b03ddbe683f4b331b75d737a67a44ecd71f988d

Observation af62cdba-6000-43af-b74d-f697c5675f72 · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicanimate: Temporally consistent human image animation using diffusion model

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:40.070249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:40.070249Z digest=sha256:a02d8ab7c4cb7429ed26c045600a1bc6ed6d0a76aa683bd05d5d0a5657c8338a

Observation 50f1e59c-f42d-4244-90e8-cc3913e8dec7 · outbound

This paper cites Human motion video generation: A survey.Authorea Preprints, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Human motion video generation: A survey.Authorea Preprints, 2024

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.723646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:40.182270Z digest=sha256:eb6d3f51dafad21c43e1cfae071dca22e446b617b280bd7fcc204b72a763d3a8

Observation 2a91e9a4-748e-4ad8-b9f0-f62fe7a174fb · outbound

This paper cites Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:40.295999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:40.295999Z digest=sha256:d1a1f68ecb969df7f65723677595f4bc287139d0d77e1d61f19ff12b2efe8754

Observation 6c5edda7-7994-4a5f-aa32-a100e15d477b · outbound

This paper cites Video quality assessment via gradient magnitude similarity deviation of spatial and spatiotemporal slices.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video quality assessment via gradient magnitude similarity deviation of spatial and spatiotemporal slices

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.474740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T14:39:40.471781Z digest=sha256:d8cf5b98505f844c6d33bfea623ac4c5abf8b4ff68478aeb11e72d8a9a8681dc

Pith citing papers

Observation 5f61ef65-53ed-4d52-ac84-5f8dcce9fa77 · inbound

Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data cites this paper.

Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:53:27.675148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:53:27.675148Z digest=sha256:a9298e58adc773b87174ff48cba6892eba8abf147df71b4f6579e2d1c00ac9f2

Observation f3b5af5a-7f4e-483f-9830-8d65451aa8f7 · inbound

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds cites this paper.

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T21:09:08.459299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:09:08.459299Z digest=sha256:ece69d7b306f1fdcc9a7608b54183e6af581f1bdedc11c595e1220806e254a99

Observation 62c4e2ef-c1ee-42c1-88c8-37f3ddddffe0 · inbound

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation cites this paper.

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:21:23.832707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:28:30.471533Z digest=sha256:8c9e499ba1f71f76bc55aa7d8a9a8a9cce785e394dd58214d1116a78b0094426

Observation b5a0b152-3b45-4e4e-8e3e-791cba101049 · inbound

SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning cites this paper.

SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:17:40.172362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T13:21:00.443497Z digest=sha256:c6d63a13cd65c35fa9c33ddb4a1e99f6a42753d950e519f081f4ae834b13d9e2