Pith. sign in

Paper Citation Record · LEDGER

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

As of 11 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 4 inbound Pith citation observations for arXiv:2501.08682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.08682 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:23:41.484020Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T19:00:44.243335Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7518a99d-d4de-40b5-8042-847ad32b2daf · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.716948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.716948Z digest=sha256:10c43596169825ac58da76ccbe7edd3af0778e9bcea20be12e7bdd53fe9f3f0f

Observation 550e779a-d7f2-43a7-b9d4-1e113b29ed14 · outbound

This paper cites Elucidating the design space of diffusion-based generative models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Elucidating the design space of diffusion-based generative models,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.284218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.733448Z digest=sha256:0d4f9e7c6050fe8dd39106aeef6b1c49ba24717d6bc1122812ce55c8ed6d0a64

Observation befa29aa-189f-4a4a-b5f7-b7ba7fa4cdde · outbound

This paper cites Auto-Encoding Variational Bayes.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Auto-Encoding Variational Bayes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.739601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.739601Z digest=sha256:a8b138dd58ace3b9ba64cb5de127b227b20e8d5fa3fac255cbf19dfaa48baf61

Observation 060082d7-6318-4df1-80e0-7eb74315284f · outbound

This paper cites Single stage virtual try-on via deformable attention flows,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Single stage virtual try-on via deformable attention flows,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.267878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.745924Z digest=sha256:3d141582cabf83bcf55fa145a78e35ea6684be1f54bf882dcbc7f9fc3db9495e

Observation 48563e53-3878-442d-b91e-0dc33e566a53 · outbound

This paper cites Sapiens: Foundation for human vision models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Sapiens: Foundation for human vision models,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.236033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.799954Z digest=sha256:7f5df1810d771ed7af76728fec3ed2619e082dc7e78c66e4b93afe6b3d9484a5

Observation 791e8503-54c9-457f-b44b-3bb8b5e9a340 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Animate anyone: Consistent and controllable image- to-video synthesis for character animation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.189360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.846489Z digest=sha256:2548a49da1a05be6edfd1b177cd9ae9c341ac326a24e94f1ac7d6896abe161e0

Observation b29a7acb-c54a-473d-9087-e49e16b3010b · outbound

This paper cites Dress code: High-resolution multi-category virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Dress code: High-resolution multi-category virtual try-on,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.154725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.851934Z digest=sha256:92a3214123c265686e104bb424334edc9aadcec5a3a5b15a67424fbca37ae78a

Observation c6e2d118-5801-42ba-8f0f-1947efebec8d · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Rerender a video: Zero-shot text-guided video-to-video translation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.115482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.858531Z digest=sha256:906d190cbd24094db917d22af55f4661f26000095eae5dab468235466936a180

Observation b1cb9d71-ce1a-440f-a14e-723fc00fd856 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Diffusion models beat gans on image synthesis,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.050704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.863638Z digest=sha256:6c7fcad2cba3e61497d5371e6fea132f44093bafc278d8592011082bf6f84036

Observation f5bdd9a6-5a08-434c-84f0-435e5a41f38b · outbound

This paper cites Learning transferable visual models from natural language supervision,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Learning transferable visual models from natural language supervision,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.019924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:40.924733Z digest=sha256:cad00031f1996ef3b667ae620ecedbda3e1c1a21e32c3dad6dbc36dc186b5276

Observation 40034013-c75a-417a-bada-db14c0134e48 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.952598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.952598Z digest=sha256:8cc27ae3c8bf803300db7097dee3cc2f0d2f9504166831a5b57346a32bc487d8

Observation 480e9800-5630-407a-84e5-374a1174f5e9 · outbound

This paper cites OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.958697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.958697Z digest=sha256:8096e27e0b287752a740596576510afaad4bc68d8639cc84f75ee0b6f82ff879

Observation 71ef351b-32bf-4cc8-b46d-ffdac378aa70 · outbound

This paper cites Gpd-vvto: Preserving garment details in video virtual try- on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Gpd-vvto: Preserving garment details in video virtual try- on,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.983205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.021799Z digest=sha256:928bc7152e58b4d4dd65be09139756afbff762361859b5f409bca0c60113aece

Observation 83084fb9-fac4-43c7-adc0-cd10c3698f56 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Latte: Latent Diffusion Transformer for Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.051508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.051508Z digest=sha256:4bd128dfd018d4cc05654e455041ec55eedff92c00bfaaff55f027ca952f0e32

Observation e010788b-c20d-48e4-9449-70010462bff5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency High-resolution image synthesis with latent diffusion models,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.937876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.057309Z digest=sha256:8547f093d4249a7da6d07d4a40688ddd717f44cf2c9ede7d7319e938257737e5

Observation 798c1a56-b1dc-4320-ac8c-04b2d4b277e7 · outbound

This paper cites VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.063202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.063202Z digest=sha256:3e85807f0ff11d92499f4c3bf9d6674f4dd8315860ac320ce8a46235e3861c71

Observation d3a4cd6a-4147-4605-9a4a-5bcbf9082ae6 · outbound

This paper cites Tunnel try-on: Excavating spatial- temporal tunnels for high-quality virtual try-on in videos,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Tunnel try-on: Excavating spatial- temporal tunnels for high-quality virtual try-on in videos,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.886479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.068622Z digest=sha256:0ac730df7b306a2ed7499937edfb87438e67617711da2aaffc156f05c9cbec87

Observation 69f62e6b-464d-4c5b-b2a9-ba8f1a416c24 · outbound

This paper cites Clothformer: Tam- ing video virtual try-on in all module,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Clothformer: Tam- ing video virtual try-on in all module,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.845582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.073473Z digest=sha256:aacd410c50b738739199ccd61f1fa117808db2e47fca26e20b26688813ad31c1

Observation 8e9446ca-ff3b-483e-86b0-e44e767e08ed · outbound

This paper cites Improving diffusion models for authentic virtual try-on in the wild,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Improving diffusion models for authentic virtual try-on in the wild,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.819282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.077891Z digest=sha256:75c2a17007c2acee36bc2cc9cffbb81e24c84fd2ff17d3ca7944d3911aa2c026

Observation 1c3237f3-1c29-45e4-aeab-6679c72b1333 · outbound

This paper cites Stablevi- ton: Learning semantic correspondence with latent diffusion model for virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Stablevi- ton: Learning semantic correspondence with latent diffusion model for virtual try-on,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.742148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.141373Z digest=sha256:117a8b9fc046c4ed751ea2c7d3eb45f551b66cd1b7b5e37ac5ab401d102421a3

Observation 2ec9f07f-b1bf-487a-83c9-6a5aa6ad3b21 · outbound

This paper cites Dit: Self- supervised pre-training for document image transformer,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Dit: Self- supervised pre-training for document image transformer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.679155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.207875Z digest=sha256:c9a7ab2a392eebb6d7f5f3de7e065686397d6c7f43050dfaefe443d2f1ff7674

Observation 7ced8ebe-d9a9-4397-bc07-1699087ad35e · outbound

This paper cites Flownet: Learning optical flow with convolutional net- works,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Flownet: Learning optical flow with convolutional net- works,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.615996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.212780Z digest=sha256:d5a4ee7eaf7999f9ea0d06b5e36491256209ed8482c22c0ebf7a31b2c6c61fb6

Observation 572c88e9-71b4-4d9b-827c-a47eb318a226 · outbound

This paper cites Shineon: Illuminating design choices for practical video-based virtual clothing try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Shineon: Illuminating design choices for practical video-based virtual clothing try-on,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.599265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.218167Z digest=sha256:fd892de2e8a9ef329be9959b2176cc554e9fff4bf6ed05483453a65324128fdd

Observation 5d77c051-aaf5-4a6d-8522-4146242327b9 · outbound

This paper cites Mv-ton: Memory-based video virtual try-on network,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Mv-ton: Memory-based video virtual try-on network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.515659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.223168Z digest=sha256:1de464911526d11b83e7299dfc3a10fdb614f06ef56be437cae8925bd31c4a9e

Observation 1f4ad7c9-c713-490d-a639-36fd86fcd7b8 · outbound

This paper cites Fw-gan: Flow-navigated warping gan for video virtual try- on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Fw-gan: Flow-navigated warping gan for video virtual try- on,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.482507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.228081Z digest=sha256:e170d142c9acb2982894c7a4f71e59a3a69fb38f4b6f1b0a1642c04758216e59

Observation 98820a1f-14bb-422e-a67c-d46580be4b95 · outbound

This paper cites Gentron: Diffusion transformers for image and video generation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Gentron: Diffusion transformers for image and video generation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.436364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.233133Z digest=sha256:056fd7a8994da5c55e4d3d860909ec979f0a1397fe767da0848f76e8bee32ba4

Observation 5ae0294a-2efd-4915-874f-05503c61c6a3 · outbound

This paper cites DiVE: DiT-based Video Generation with Enhanced Control.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency DiVE: DiT-based Video Generation with Enhanced Control

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.238260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.238260Z digest=sha256:4012d0050664f37c79f64496f43aa317d9f6baac8099e40fda9ac76373a1d684

Observation 6933059d-7d9f-4a4d-acfa-8daed6c78319 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.243874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.243874Z digest=sha256:efb81dabbac16768abfafe0958335ec54b7884bc1e3dd7d32021ada9de70264d

Observation 8739ab18-7fd0-4818-be72-d880f5617086 · outbound

This paper cites NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.249291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.249291Z digest=sha256:7e36b9138758a34fbcdc9b0d316745b915b50ddff14c2eb413a8a15512b936ee

Observation 93c28c2b-f510-4270-ab7e-0afb0e787922 · outbound

This paper cites Trip: Temporal residual learning with image noise prior for image-to-video diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Trip: Temporal residual learning with image noise prior for image-to-video diffusion models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.398694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.254119Z digest=sha256:25820c880cef5d2db666674f21a08f83adeb4896e0df05ddbadc5034c7750219

Observation 71dd5b71-2eb1-4d9c-b646-7440dc726f38 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.258890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.258890Z digest=sha256:b418799b819eb2deee532f8d70ef40332a91a496fc99e5b1f52b522a0aed6cef

Observation ce4f4262-4437-4568-b641-e7dc5013a8d3 · outbound

This paper cites Multi-concept customization of text-to-image diffu- sion,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Multi-concept customization of text-to-image diffu- sion,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.328075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.264121Z digest=sha256:606c483729a440e86ea89083e1f45b8ab8a3a1758209865f32667c0049512d57

Observation 8ed88b5e-5cc5-43eb-9b1c-cb701fe6c125 · outbound

This paper cites Distilling Diffusion Models into Conditional GANs.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Distilling Diffusion Models into Conditional GANs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.268741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.268741Z digest=sha256:50b3b3f1e1f36b9c50d6b2a9f99a88b007892c9e3100fe71a7e253ee592311c6

Observation 44fece41-f004-493a-b349-5d3cf8640c70 · outbound

This paper cites Toward characteristic-preserving image-based virtual try- on network,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Toward characteristic-preserving image-based virtual try- on network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.302802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.280868Z digest=sha256:9656c182611f8c6fbb73b0b6162f030256cbde5ec5e3b0ca0c95a11cfe8628d9

Observation 8374b284-19a3-4721-93c4-0c4fc3eb2d0d · outbound

This paper cites High- resolution virtual try-on with misalignment and occlusion- handled conditions,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency High- resolution virtual try-on with misalignment and occlusion- handled conditions,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.217866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.336274Z digest=sha256:86cbb0be52370e28ce0d9c733bad57e5b80b2316125ce85508ba8750efee0a9a

Observation 47c9771c-594a-4c43-a5b4-72d9329e894c · outbound

This paper cites Ladi-vton: Latent diffusion textual- inversion enhanced virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Ladi-vton: Latent diffusion textual- inversion enhanced virtual try-on,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.199086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.398737Z digest=sha256:efad8742bb7f46d971266f2e42281ad176d4f4faa0783eaa5dedbfd5254e912f

Observation 8c839f01-dd03-4c51-995b-338fa2e623cc · outbound

This paper cites Tam- ing the power of diffusion models for high-quality virtual try- on with appearance flow,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Tam- ing the power of diffusion models for high-quality virtual try- on with appearance flow,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.118515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.428921Z digest=sha256:a6ed7a17f561f71d30c8f838a73a33da735a251c3bced11b3ee537ee2a9e044d

Observation 5482fa6b-05da-47f6-8a76-a93a23448634 · outbound

This paper cites Viton-hd: High- resolution virtual try-on via misalignment-aware normaliza- tion,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Viton-hd: High- resolution virtual try-on via misalignment-aware normaliza- tion,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.083295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.435225Z digest=sha256:e55e0ab50472bf9da1137c55762c3889d98a3c62cdc5dc6e620232f11d5bd1cf

Observation 647f8ce2-7440-4285-b1af-88ee7cc734b9 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency SAM 2: Segment Anything in Images and Videos

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.440256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.440256Z digest=sha256:cdeee3181271aae74caa31d23631f9779b62c8df8c302b59073754d3c3dc9def

Observation 433d9b05-bee6-4858-89fc-35278f4053e8 · outbound

This paper cites WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.445965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.445965Z digest=sha256:280b3371a20c18f0d280de7a057fa1e74abf6357c3b29cec70a8fcff12c5d966

Observation 0d4582a0-57f0-4dcf-a4bd-20ca3c36c912 · outbound

This paper cites Cat-dm: Controllable accelerated virtual try-on with diffu- sion model,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Cat-dm: Controllable accelerated virtual try-on with diffu- sion model,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.046765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.451757Z digest=sha256:978266ff69b729a0f7a230227e731d747cf26bcd11ab51ab5ab2529486c73d16

Observation b8168b2c-a642-422c-8268-9603ad672aff · outbound

This paper cites Multimodal garment designer: Human- centric latent diffusion models for fashion image editing,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Multimodal garment designer: Human- centric latent diffusion models for fashion image editing,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.966118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.457069Z digest=sha256:1a16d6bedbec4f75e348f765ef7a3731e8e0b201150d0028951b3edad1593057

Observation ec3d2185-6b87-4195-9aff-73d7949f0f55 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Paint by example: Exemplar-based image editing with diffusion models,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.934849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.462209Z digest=sha256:390134f037deaeb0a9c4e10d46a0810659f433dbb0e81db3ba5fdd3f18291c85

Observation 814bbeae-7d9b-4220-a198-0a969c59d755 · outbound

This paper cites ViViD: Video Virtual Try-on using Diffusion Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency ViViD: Video Virtual Try-on using Diffusion Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.467754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.467754Z digest=sha256:e3f299b225443408d7459de3d036a5e22dc710318e84468b040a48cecc7a83c3

Observation fd6e0799-3c37-4d05-a737-8bdc91831b0b · outbound

This paper cites Fresco: Spatial- temporal correspondence for zero-shot video translation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Fresco: Spatial- temporal correspondence for zero-shot video translation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.916071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.473455Z digest=sha256:c8174a44842364de3232187b57bbb5a97b875549cb51f0df083e0a1659d33854

Observation 9c4ea21c-aaf8-448e-ae38-421aefa52545 · outbound

This paper cites Lvcd: reference-based lineart video colorization with diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Lvcd: reference-based lineart video colorization with diffusion models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.819877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.478486Z digest=sha256:6decda63afe8a7c052c04c0cfb287d0d3198e98c496e3dd40ac2e001ea8c9c09

Observation 4b5c0654-a605-497c-ab96-a0ab08d6ca5a · outbound

This paper cites Detectron2,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Detectron2,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.799815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T20:23:41.484020Z digest=sha256:a4516e2164bbc6d092e6eec7d6012a11fc5f33dc9c3fe6da7642a5bd7fdede0c

Pith citing papers

Observation cb931202-32af-476a-97a6-f8f71ed70c7e · inbound

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on cites this paper.

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:39:10.564830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-17T06:34:45.045267Z digest=sha256:a87392692478b53a5374f94881b91c433b15191d11bf2129edaeb590457c0d6d

Observation c504a272-756e-4e04-b20f-0bfaf4d54e58 · inbound

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection cites this paper.

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:08:22.943585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T20:06:30.109866Z digest=sha256:544840f1f312131a0f94697089b8bd6db3bce56b6a8f315fa5303f9e32f85643

Observation e0287f82-ae83-4d57-9d15-1cc25eb3eb4a · inbound

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On cites this paper.

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:05:12.414338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-07T05:43:04.644875Z digest=sha256:50ebf2dcab006a0607c5afc4088adab12269f85c511acb2a6192c0beaa0b9bda

Observation 6058e754-7fc3-4422-bf35-3cb623eac8e9 · inbound

OmniTryOn: Video Try-On Anything at Once! cites this paper.

OmniTryOn: Video Try-On Anything at Once! RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:17:26.409824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T19:00:44.243335Z digest=sha256:3394af198191e3d743dad2cdfe662c6b2fe10c6f85cb9204de0085ab885e2619