Pith. sign in

Paper Citation Record · LEDGER

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry

As of 17 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 0 inbound Pith citation observations for arXiv:2607.18227.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.18227 v1

Coverage vector

measured 92 of 92 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T15:42:04.839943Z

measured 92 of 92 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

92 of 92 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved92
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation db4b9e09-a94e-4a25-a399-18232edd01e3 · outbound

This paper cites Consistent video- to-video transfer using synthetic dataset.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Consistent video- to-video transfer using synthetic dataset

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.216633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.216633Z digest=sha256:960c91a4438fdec55ef0b54a8d4b15e7e3822830a5f6e45fdc97588bc6bc73c0

Observation bb45c8f6-09df-4293-b10b-ae71b8b2fe99 · outbound

This paper cites Prompt-to-prompt image editing with cross attention control.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Prompt-to-prompt image editing with cross attention control

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.309571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.309571Z digest=sha256:5aa0a51cd07d98742f471e18926609ebb51b2d1f59cdfce36ce2225f46be0420

Observation 9dc9ce9b-fc9a-41ba-825a-e4436e22697c · outbound

This paper cites Se ˜norita-2m: A high-quality instruction- based dataset for general video editing by video specialists.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Se ˜norita-2m: A high-quality instruction- based dataset for general video editing by video specialists

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.389862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.389862Z digest=sha256:c6db54154d4fc062af15ecb401bb931ea204882c3c988b8cced6e2ac2d1297e1

Observation 8bb09e9e-2dce-4615-b583-77763e8c81d0 · outbound

This paper cites Scaling instruction-based video editing with a high-quality synthetic dataset.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Scaling instruction-based video editing with a high-quality synthetic dataset

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.476293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.476293Z digest=sha256:8fc30db53366956de34b0a8972ad9371794d8b2a11f69ed224968af6fccde30a

Observation 9d669d84-7cce-4215-b10e-494b14cdcd0c · outbound

This paper cites Insvie-1m: Effective instruction-based video editing with elaborate dataset construction.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Insvie-1m: Effective instruction-based video editing with elaborate dataset construction

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.544673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.544673Z digest=sha256:533189a16e4c0d0e769ee1c75baf7e7a49b9c40fa8701bd6a2c261012b93b339

Observation 0e3ab94c-c493-4979-b1dc-fd69ee435dcc · outbound

This paper cites Openve-3m: A large-scale high-quality dataset for instruction- guided video editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Openve-3m: A large-scale high-quality dataset for instruction- guided video editing

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.644996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.644996Z digest=sha256:8819a16fa5f1f64639fec7170031f946b41dc5a6876750cea1e0980e0751bd34

Observation 04edd82e-593b-4d23-a14d-8d1e98717c8d · outbound

This paper cites A survey on multimodal large language models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry A survey on multimodal large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.752787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.752787Z digest=sha256:11abe11865255408579408206bd1205fc80aa7a576b4f4303de140b97938be75

Observation dc8805d0-144f-4f38-a18c-b4b7a34fcea8 · outbound

This paper cites Vision-language models for vision tasks: A survey.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Vision-language models for vision tasks: A survey

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.837722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.837722Z digest=sha256:9830035e934b02b8c030fcff82bbdafb956d9bb6f509bcad9b734c3d9141422c

Observation 9132d125-fe69-4e05-a142-a134c8d5c319 · outbound

This paper cites A survey of state of the art large vision language models: Benchmark evaluations and challenges.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry A survey of state of the art large vision language models: Benchmark evaluations and challenges

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.874659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.874659Z digest=sha256:ff1f85810b461c62770364c99301fe14adfa8f7ae30230fafc4cb760ee898485

Observation 902a6ec3-b3b8-4cb4-bd7c-e6028a7de5ce · outbound

This paper cites Sam 2: Segment anything in images and videos.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Sam 2: Segment anything in images and videos

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.895487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.895487Z digest=sha256:05b140b61af99c077ba5b093e1eecea3ba408b8e5290975f27e1f4ddecffcc12

Observation 097c0cc7-9512-4dc6-b49f-01d4e86b4728 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry SAM 3: Segment Anything with Concepts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:57.930637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:57.930637Z digest=sha256:d8075685d50fa66754bf4a31d375f75cbdbd25ca839d8be9fcd7898c52193876

Observation e29a3d2e-d272-45cb-9fa0-e4a216646e16 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.005473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.005473Z digest=sha256:d703b3d13f58b740be8f8d651a1a4c1edf37b0f95ce235f0c6330a9127efb586

Observation 12383843-5f11-4ce1-b646-a5fe3208c906 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.117595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.117595Z digest=sha256:df81185657d6007af58df9897a73702d38b3d33f620f88a5cf25c5a5c8bfaf68

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.279994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.279994Z digest=sha256:56a306f44510b1bca3d23d69ea5299ed21f8b26537470af118dbe57a94eb250f

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.433414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.433414Z digest=sha256:b9efa254a8290109dd765dc856365ae4b6983231430d0084196db4a3a29c2419

Observation 2b063174-dfb5-4d46-849d-fcbe204261a1 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Adding conditional control to text-to-image diffusion models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.559802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.559802Z digest=sha256:a99b1045ec0fabf78ecf90e276a4178778673fa676a0d09820065118ec04c420

Observation d17c1185-4208-4f5b-9042-89acc1d71905 · outbound

This paper cites Reward models in deep reinforcement learning: A survey.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Reward models in deep reinforcement learning: A survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.664064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.664064Z digest=sha256:f5f613b6a2be464965baaf8d35a4dc4d21df39a4b6e663effa2c2860e333287b

Observation c3de7022-46a5-4441-927f-6a18d4cf0139 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Wan: Open and Advanced Large-Scale Video Generative Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.802918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.802918Z digest=sha256:7643d10703aa6a6843d2c0ecaf97e65fb8251f4e2d93f5c4f47aa2bd583118de

Observation c1b422bf-e369-48e9-8c2c-5d1ed6766777 · outbound

This paper cites Instructx: To- wards unified visual editing with mllm guidance.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Instructx: To- wards unified visual editing with mllm guidance

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:58.910602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:58.910602Z digest=sha256:88be76e8676e11980460fe10b7d9c3fcec6e65532f38b4f3c29b8e12dd604668

Observation 8076f39d-5fbd-4c3d-8f93-465637181a2e · outbound

This paper cites Vino: A unified visual generator with interleaved omnimodal context.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Vino: A unified visual generator with interleaved omnimodal context

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.030453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.030453Z digest=sha256:89547d20ca022a29a54d02285d43bb5cf9836cef67e7ab0b3262c20b9e6c8235

Observation 4bbb1ef0-d25a-4c57-a803-b2f19a08d1de · outbound

This paper cites Videopainter: Any- length video inpainting and editing with plug-and-play con- text control.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Videopainter: Any- length video inpainting and editing with plug-and-play con- text control

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.199648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.199648Z digest=sha256:e60f661501ac20eedd0062f9e715a98f7f61c9d35f45dd5b3e74425b11bf3f2e

Observation fe60997d-7e33-4d78-a61b-20ae20b55594 · outbound

This paper cites UNIC: Unified In-Context Video Editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry UNIC: Unified In-Context Video Editing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.332116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.332116Z digest=sha256:a4afb0543dda46e0aba22f6da5654792d2d69bf5e179894511d90b997d72a66c

Observation d47b1e4b-adc2-4d98-8090-bc7e559b5e24 · outbound

This paper cites Anyportal: Zero-shot consistent video background replacement.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Anyportal: Zero-shot consistent video background replacement

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.393693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.393693Z digest=sha256:56fdc0648e9f918c1aa94ef138934fd46cb813a11c988704813a631bcfbdc208

Observation 473f52da-8405-4c57-bf27-939dc2d89906 · outbound

This paper cites Generative video propagation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Generative video propagation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.430887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.430887Z digest=sha256:9a65becd6d1e8ec8fb450d7f1b917f558da6aae15e7af0afc41d69963ecb1795

Observation 74c70681-e811-4e96-b91b-88e528340fa1 · outbound

This paper cites Univideo: Unified understanding, generation, and editing for videos.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Univideo: Unified understanding, generation, and editing for videos

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.508246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.508246Z digest=sha256:cc1ea478854f3d8e8931809a572157d81e830ba3f45116a46aa836b680b610d8

Observation d32479a3-12d2-40c3-bd90-aa52ec5fcebc · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction- guided image editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Magicbrush: A manually annotated dataset for instruction- guided image editing

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.564471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.564471Z digest=sha256:76e833954713993705ea2e0854340c25f0d96d79afcbb46d9cac92079ee45281

Observation e69a4e1b-61c9-450f-a2a5-2af79d2057d7 · outbound

This paper cites Pico-banana- 400k: A large-scale dataset for text-guided image editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Pico-banana- 400k: A large-scale dataset for text-guided image editing

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.660118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.660118Z digest=sha256:b65ada33196f1794ffb61cd52b2cf284333cd04da85fee3a1c4785942d0fdccd

Observation 8d95ddc0-def7-4859-a94a-c6d646f64f28 · outbound

This paper cites GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.773307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.773307Z digest=sha256:3e4f7b7f2338e923c63acb33b60038002a416f262193a6697daa765b1c55055c

Observation 069060ca-1a49-4793-94c3-07211f5aabe8 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.890584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.890584Z digest=sha256:c3ca7749e803196c4170e1d3437689de0e93828eb7c13f9bd164a580efac8787

Observation 2680f12b-3e70-43b2-a087-ecf4f94c3ee6 · outbound

This paper cites Learning transferable visual models from natural language supervision.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Learning transferable visual models from natural language supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-01T15:41:59.994300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:41:59.994300Z digest=sha256:07ec51fc58da613a89daa454651b5f4741008c7216bb7d3bed893c0032cd24c9

Observation 132791cd-6678-469d-9ec6-62f6e4eaf2ab · outbound

This paper cites Relightmaster: Precise video relighting with multi-plane light images.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Relightmaster: Precise video relighting with multi-plane light images

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.083678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.083678Z digest=sha256:8f3bd7b4d12999ccd15ad6a92e07ad1b1400fa13e53f349034c8a2c9d563b5be

Observation 7ef07ade-f10e-4026-acbe-28cd9ba96df2 · outbound

This paper cites Recammaster: Camera-controlled generative ren- dering from a single video.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Recammaster: Camera-controlled generative ren- dering from a single video

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.141560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.141560Z digest=sha256:2aace0adab87b08daaa1cd5d4ebaf1119795c873971b31fc3b6ce2d5c59e3978

Observation d6515028-dcb6-4411-82a4-bd546e07a900 · outbound

This paper cites Anyv2v: A tuning-free framework for any video-to- video editing tasks.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Anyv2v: A tuning-free framework for any video-to- video editing tasks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.211910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.211910Z digest=sha256:8010f8ce837818fa8e9f492b80dbfa5a69c282c0d2802d27b6f4b1653126cd02

Observation 4d1a38c1-c1cd-411e-b8a7-092789483474 · outbound

This paper cites Denoising diffusion implicit models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Denoising diffusion implicit models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.292306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.292306Z digest=sha256:b8ef60fa33697f68d47666f799f910ed0c2fbf91b7bebf8e50c84629bb718a63

Observation fc7ffa9e-a825-4a02-8557-aa58b2871bd0 · outbound

This paper cites Videoshop: Localized semantic video editing with noise-extrapolated dif- fusion inversion.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Videoshop: Localized semantic video editing with noise-extrapolated dif- fusion inversion

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.349393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.349393Z digest=sha256:119cb0745a5cb7df0b47879b8bab9f38076d078df037e04f679583b5692e959b

Observation 6e5a6677-e040-4002-bf5f-279feef7a571 · outbound

This paper cites Con- textflow: Training-free video object editing via adaptive con- text enrichment.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Con- textflow: Training-free video object editing via adaptive con- text enrichment

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.430575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.430575Z digest=sha256:f4de6e6791f6474c66a0010bb0f736efc7bced41f25ada9c5561382e023d32d0

Observation bce70e99-11d4-4a5a-9037-21538531cc14 · outbound

This paper cites Flow matching for generative modeling.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Flow matching for generative modeling

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.490828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.490828Z digest=sha256:4d9bbb1d8a70876abd51eae74a5dd3360feb32847cfa033272bba92fd85412fc

Observation b1ef6fe4-9529-4bec-942f-e9c8dcb572e5 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Lora: Low-rank adaptation of large language models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.571605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.571605Z digest=sha256:3436250cca3b5eb712cc45cbbfa48e4aabc11737c16a15d480414fd11057c743

Observation ece7d761-759d-4d6e-bddc-ceeebfdabad4 · outbound

This paper cites I2vedit: First-frame-guided video editing via image-to- video diffusion models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry I2vedit: First-frame-guided video editing via image-to- video diffusion models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.654441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.654441Z digest=sha256:c69b4dee4abcfa309d47db764030bb23c35c854ecdb7dc42b11df8dfc03c64f1

Observation 9aeb0504-5cf4-40f7-a750-65e1448be076 · outbound

This paper cites Vace: All-in-one video creation and editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Vace: All-in-one video creation and editing

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.710693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.710693Z digest=sha256:4bab7d7d153a474a5fd2b58f3beff7ff69789c83e6b82725ab6b8cbbe49e42b0

Observation e1044260-f59b-42f7-828f-1d63cedad4a6 · outbound

This paper cites Scalable diffusion models with transformers.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Scalable diffusion models with transformers

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.787997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.787997Z digest=sha256:fa261f0f895baba13107b90e8c5f501d16b37a7401d0a299aedcc54e9b13add5

Observation 308d5689-58c5-4a7c-8980-d6e8d2d1209c · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Roformer: Enhanced transformer with rotary position embedding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.845057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.845057Z digest=sha256:c038f5a7b21c6112987e4da7cfe91cf93ec3b71a41eb6a1d2930ebf5d8b61272

Observation e7c95de6-ee28-45fb-bbe9-d2b66a21b745 · outbound

This paper cites Light-a-video: Training-free video relight- ing via progressive light fusion.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Light-a-video: Training-free video relight- ing via progressive light fusion

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:00.920461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:00.920461Z digest=sha256:da4e30cad2e25cd6d9facb34028f9af9dc0d778fb370b9dafc38ec37135ba5d5

Observation 9baccbf0-d7de-49da-a7f9-4d4118096b1c · outbound

This paper cites Vface: A training-free approach for diffusion-based video face swapping.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Vface: A training-free approach for diffusion-based video face swapping

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.002588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.002588Z digest=sha256:3a004395d41ca9f9e1bceb5a2cf99a0523a3600ff3c0e5c3576309723e438f4f

Observation b380c154-95a2-4fd3-a9f0-39f6253027c2 · outbound

This paper cites The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.141862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.141862Z digest=sha256:4956f727160d0fc1df40ddb50daf0edcb2bc719ed0a785308020aceacf09df7e

Observation 3d3c9405-a7ca-434b-99c2-3fc9926e6b20 · outbound

This paper cites Rig-your-portrait: Controllable and re- lightable portrait video generation with explicit 3d guidance.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Rig-your-portrait: Controllable and re- lightable portrait video generation with explicit 3d guidance

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.238742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.238742Z digest=sha256:0e57e6576e7469b1a567676abea93cea5a97af63cb1f17f73102a434ab5f81e6

Observation 2dbbc15c-a895-42bb-b53e-859a3cff9773 · outbound

This paper cites EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.295762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.295762Z digest=sha256:b7af7f340b4fbd7db4c75bc011fba24a97bd4a380dd6a713adf529a0013a685c

Observation 90ce21e6-5192-4ada-8dd2-481f6db9b704 · outbound

This paper cites Omni-video 2: Scaling mllm-conditioned dif- fusion for unified video generation and editing.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Omni-video 2: Scaling mllm-conditioned dif- fusion for unified video generation and editing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.349038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.349038Z digest=sha256:2431d99cdde2b7460a8c6b45618b24fe62c30ef1481e375c7f8631d5f58c7ba8

Observation c60929e1-3d70-4e85-9233-cbde4651bf32 · outbound

This paper cites Multimodal Referring Segmentation: A Survey.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Multimodal Referring Segmentation: A Survey

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.408279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.408279Z digest=sha256:5d985867a43b7e2f675cb99fd0086d63f42cc41d86576bb21fc4be0f0fb70839

Observation ac111978-ed64-4c35-99e1-cc3123246abc · outbound

This paper cites Attention is all you need.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Attention is all you need

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.475772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.475772Z digest=sha256:ba8b8d85a3fc02fe1ad1152ae99d4d509a63417ac7345b22df5ad3959ce33ef9

Observation 90d47f93-34ab-49bd-bac6-23a715397341 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Exploring the limits of transfer learning with a unified text-to-text transformer

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.533862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.533862Z digest=sha256:3323b1f1e6c18d92390aa7eeab3889d193f4a8a22dddf64a1430842e5316fb02

Observation bc1d0608-13d9-48d5-8242-5cd1f60e5406 · outbound

This paper cites Five-bench: A fine-grained video editing benchmark for evaluating emerging diffusion and rectified flow models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Five-bench: A fine-grained video editing benchmark for evaluating emerging diffusion and rectified flow models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.592204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.592204Z digest=sha256:c370d1b1ab175bb4857459f424f2f1f74f4ac85bd3b0bebad67250314946d419

Observation 39151521-f7ab-45af-b4a1-d1f0529135e6 · outbound

This paper cites Sparse videogen: Accelerating video dif- fusion transformers with spatial-temporal sparsity.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Sparse videogen: Accelerating video dif- fusion transformers with spatial-temporal sparsity

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.633450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.633450Z digest=sha256:34c98750b2d54ba2e9c98db046acd62bb6591548d487a8dc150ff94dccb9d13b

Observation 39c76206-ff77-4633-8e83-8d71448ed968 · outbound

This paper cites Kullback-leibler divergence.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Kullback-leibler divergence

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.687587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.687587Z digest=sha256:fe757b3199ecf2318c512225d6fa9a09b2fda20a9c94318264466b8d0eae0217

Observation 3d0f420c-d045-4a89-8da2-107e48479fab · outbound

This paper cites Unireal: Universal image generation and editing via learning real-world dynamics.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Unireal: Universal image generation and editing via learning real-world dynamics

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.721201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.721201Z digest=sha256:90199e7eb54fa8d02b4ccfe94bc57a5160ef46e70aac42ca6b2f545f382215ab

Observation 8402ac5c-9fe2-4365-ae60-e5919807b066 · outbound

This paper cites Omnigen: Unified image generation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Omnigen: Unified image generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.764727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.764727Z digest=sha256:0e233092ac1bae2e080c0525f2e949aac55729c2f97cc4caa649d81f282cb51d

Observation c915adf0-ccfb-423a-83d9-a118cbb3b96c · outbound

This paper cites Language as queries for referring video object segmentation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Language as queries for referring video object segmentation

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.877327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.877327Z digest=sha256:360693ccda3533d51c8079c49c80d584a6f9500d9220d7d44ed0f3ea13eafb9d

Observation 068fbd4c-5a11-4744-ab8b-a5311ae12732 · outbound

This paper cites Semantic and sequential alignment for referring video object segmentation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Semantic and sequential alignment for referring video object segmentation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:01.937601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:01.937601Z digest=sha256:b25de6f66c630197412eecb75a5471fc47934f896387da44756361f63571c4da

Observation 885e126b-855d-4857-89ef-8ea50c3df001 · outbound

This paper cites Modeling context in referring expressions.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Modeling context in referring expressions

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.044488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.044488Z digest=sha256:b8258b6a567a058298dc8cbd17564f3383be8f9c1387f3ec07395bcf74ca0a90

Observation debb7b83-fbb4-48ca-a7df-1bb13a57e61d · outbound

This paper cites Referitgame: Referring to objects in pho- tographs of natural scenes.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Referitgame: Referring to objects in pho- tographs of natural scenes

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.129971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.129971Z digest=sha256:dcabcdbfc127cb757820764c2ad9607fef9aff0449b0445ade19bf03912041a5

Observation f58afcc8-b1d0-47d1-9d67-72c26498e48a · outbound

This paper cites Gres: General- ized referring expression segmentation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Gres: General- ized referring expression segmentation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.287822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.287822Z digest=sha256:3a4fa601c8d352c09969a2986557a1d0ae11b8b514127dcf4c69b2408bd1ef03

Observation 2c382caa-896e-4ade-8f2c-8884926d3b61 · outbound

This paper cites Fastcomposer: Tuning-free multi- subject image generation with localized attention.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Fastcomposer: Tuning-free multi- subject image generation with localized attention

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.393858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.393858Z digest=sha256:df731db530b19fafb23f12ce257ba515e02a454e3d4bb82f51a68b9589928e35

Observation 024da12c-81f7-4ac7-a56d-95084efce18c · outbound

This paper cites Do vision trans- formers see like convolutional neural networks? InNeurIPS,.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Do vision trans- formers see like convolutional neural networks? InNeurIPS,

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.454872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.454872Z digest=sha256:330ea14d84b4d558deb59ad315e20c0e641f3ebfd79f7ba3bfb090a988d298a3

Observation 1d1d8e9c-335a-4228-aebb-f53b1f644213 · outbound

This paper cites From Ideal to Real: Stable Video Object Removal under Imperfect Conditions.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry From Ideal to Real: Stable Video Object Removal under Imperfect Conditions

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.522075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.522075Z digest=sha256:117d24acaa0ef5e245929492f0ee074e8aff296ac9aad5952825bc34995c59ac

Observation a42aa916-6d44-4a0f-944e-d1c5d7b7b16f · outbound

This paper cites Deep multi-scale convolutional neural network for dynamic scene deblurring.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Deep multi-scale convolutional neural network for dynamic scene deblurring

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.559921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.559921Z digest=sha256:27f004b85d0b7bdd72ba92cf46a1a7f117cb155b9f6e727232b632f214c35151

Observation ce0f7f52-7dbf-4df0-8378-763e71df8486 · outbound

This paper cites First frame is the place to go for video content customization.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry First frame is the place to go for video content customization

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.651514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.651514Z digest=sha256:0609a0248da69542182a9c766129c4091a8b52f047ab69f64328d3b40cf88885

Observation c48a6c3d-08c4-4b11-a3c4-faa1f8f7af9e · outbound

This paper cites Deep learning-based image and video inpainting: A survey.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Deep learning-based image and video inpainting: A survey

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.721564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.721564Z digest=sha256:307e8560ff53ffe3d2c30ade7d56a22fa0acd5bb2cbd7545d7cc6e01be1e0a48

Observation 88765cda-8bd2-4af9-8ba9-579a9717b0d7 · outbound

This paper cites Unboxed: Geometrically and tem- porally consistent video outpainting.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Unboxed: Geometrically and tem- porally consistent video outpainting

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.765993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.765993Z digest=sha256:4fd2a8d22de276ae1b073ee516d27069b9479e970fb470a19878b08ba5938dc0

Observation 369cb4f9-9d91-4d7f-9480-0e902e5a5739 · outbound

This paper cites Colorflux: A structure- color decoupling framework for old photo colorization.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Colorflux: A structure- color decoupling framework for old photo colorization

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.816873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.816873Z digest=sha256:38e4333a75a593ae1c47db01151979c0f8392c96cc1f777ac8124d3b1f399965

Observation abdd9e1d-96df-4533-9a0e-4e3412fb0779 · outbound

This paper cites Colorsurge: Bringing vibrancy and efficiency to automatic video colorization via dual-branch fusion.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Colorsurge: Bringing vibrancy and efficiency to automatic video colorization via dual-branch fusion

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.887156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.887156Z digest=sha256:9aa813d6ebde15633f132a582f67ebf10c72ddb4ef5299f9eefbbae0d78a0cd7

Observation 6fd347bc-9ec5-4e6e-bf41-61c4d22ea039 · outbound

This paper cites Objectclear: Complete object removal via object-effect attention.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Objectclear: Complete object removal via object-effect attention

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.955598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.955598Z digest=sha256:7a988890e9be4f6aad7d736169815d66899d426e4af71823d8a00659fd5552ae

Observation fcf6de5f-7bae-4346-9673-4fa637c79b03 · outbound

This paper cites Insert anything: Image insertion via in-context editing in dit.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Insert anything: Image insertion via in-context editing in dit

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:02.993948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:02.993948Z digest=sha256:4db7f5b9ff2a633f0cd277d43168955ca11ef8a6d40140d1c9a3abd5f75c1e19

Observation 3f2b1c8c-c7cd-40e2-8a68-3e0d2aa37d9e · outbound

This paper cites Champ: Controllable and consistent human image animation with 3d parametric guidance.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Champ: Controllable and consistent human image animation with 3d parametric guidance

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.085818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.085818Z digest=sha256:ccfc9b782f134b4139577037d1769b2a09dbaccefc2f206aff81bd6b7e7506ec

Observation d0b5bbb2-e48a-4852-b2d2-ab7b4d43a9b4 · outbound

This paper cites Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Phantom-Data : Towards a General Subject-Consistent Video Generation Dataset

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.214892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.214892Z digest=sha256:f54c66e7711a8fd58525c0f95a228be9cadb424b4e232813d0a030651debe720

Observation 3a0c24fa-408a-4714-81e7-a29a5bb819fe · outbound

This paper cites Phan- tom: Subject-consistent video generation via cross-modal alignment.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Phan- tom: Subject-consistent video generation via cross-modal alignment

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.324083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.324083Z digest=sha256:13fb0717f38a26300fad90bddf502ec6cf0b1dba90016986f14437889885bb8b

Observation 8ff40304-302a-4389-b3fb-85f422fe1622 · outbound

This paper cites Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.441093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.441093Z digest=sha256:7f20d79047c1a152629ea6335d66273811fad8a6faf0c4af6c421d8440aaed64

Observation 8c688440-66e0-4cde-8beb-54a04b18f25c · outbound

This paper cites A Survey of On-Policy Distillation for Large Language Models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry A Survey of On-Policy Distillation for Large Language Models

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.593454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.593454Z digest=sha256:587c31330d45c51fa185ed3289d54bb0b22b6ee1f8f77f4844c396b5ca290a89

Observation d8063b2e-6ab6-4e30-9d1e-51f5ff310729 · outbound

This paper cites Reinforcement learning: A survey.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Reinforcement learning: A survey

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.746335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.746335Z digest=sha256:7bac888f1f7bb6f77cbb391defe8091dd386e0744e3e01c6aa1594abcdc52b99

Observation 56001c31-aace-44cd-a66a-47299c962e2c · outbound

This paper cites Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.829206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.829206Z digest=sha256:3aed8503b14bb8aa1d913008f83ba6fc1538932b5f3e59ca137706fb47c97108

Observation fd2ffb50-983a-4100-80e0-3e00f6bbf298 · outbound

This paper cites Scaling rectified flow trans- formers for high-resolution image synthesis.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Scaling rectified flow trans- formers for high-resolution image synthesis

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.903557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.903557Z digest=sha256:32fce356addfcf08d13a290849c6324bd6b1b8c870c907edf187bc1e67c8932b

Observation 5c5b56b1-23cb-47c2-ad6c-9f1f9ce36476 · outbound

This paper cites Enhancing robustness in multi-agent reinforcement learning via temporal consistency regularization: A self-distillation framework.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Enhancing robustness in multi-agent reinforcement learning via temporal consistency regularization: A self-distillation framework

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:03.989549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:03.989549Z digest=sha256:a44498cf1bb963c566f0f0fdd68bf80eb8e84e5bd3830b972a43bddf8348fed7

Observation 2fff4f34-bfc2-4241-9075-c35dec88808d · outbound

This paper cites UniSD: Towards a Unified Self-Distillation Framework for Large Language Models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry UniSD: Towards a Unified Self-Distillation Framework for Large Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.082077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.082077Z digest=sha256:0a335b0c0e31362f9007217be3c52f963d66b8a65a84f7107da7b111625e5ff0

Observation 4b14d1bf-525e-4373-8683-fd5722250de4 · outbound

This paper cites Cccaption: Dual- reward reinforcement learning for complete and correct image captioning.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Cccaption: Dual- reward reinforcement learning for complete and correct image captioning

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.171705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.171705Z digest=sha256:0150d0da64cde25b979c3e41b38dc984765684779874d96bf911a347734b846a

Observation ab000c22-e877-49f8-b07b-4556e75884f3 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry 3d gaussian splatting for real-time radiance field rendering

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.231784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.231784Z digest=sha256:07e69d1bbbd819e5eb8fd80e771b7843f48d7ec11ecc38b9b6629755e3e24a82

Observation 883dfdec-422d-49c4-8220-884b27918352 · outbound

This paper cites Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Human Preference Score v2: A Solid Benchmark for Evaluating Human Preferences of Text-to-Image Synthesis

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.315681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.315681Z digest=sha256:08a24abeddcc5d57de8bcfe13bbab05e63e6488784ffe26acfd0d773588913ab

Observation c911c923-e79e-4cfd-b091-d322e5db12c1 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Pick-a-pic: An open dataset of user preferences for text-to-image generation

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.380346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.380346Z digest=sha256:8223a66d759f351357e980318c3d72b09519bb2ec6a76a9b27799d7f05ebd258

Observation fbec9734-90ec-4c6a-9deb-6718b62060d2 · outbound

This paper cites Stam: Zero-shot style transfer using diffusion model via attention modulation.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Stam: Zero-shot style transfer using diffusion model via attention modulation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.466521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.466521Z digest=sha256:313079ec8aef947ea1bf1b8b443228c3d5d24a8b01a43f295a50d98134218be9

Observation 19e5bd6e-bf03-48da-9fb0-411722176122 · outbound

This paper cites Measuring Style Similarity in Diffusion Models.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Measuring Style Similarity in Diffusion Models

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.526135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.526135Z digest=sha256:2fa5471bd2a7608056f6c11894533cb29b9808393f93496be53eeef137c8d691

Observation f0cd5703-1622-47c9-83b3-2361ea843bb9 · outbound

This paper cites Dinov2: Learning robust visual features without supervision.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Dinov2: Learning robust visual features without supervision

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.604088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.604088Z digest=sha256:43e7121cb06cb81aa5cc7ca9f2140a018c5a772bc76ac10eca6f1dcafb342ceb

Observation 08e24ebb-a40f-4b8f-89d9-042fcf6c91a0 · outbound

This paper cites Classifier-free diffusion guidance.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry Classifier-free diffusion guidance

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.666043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.666043Z digest=sha256:b25bfb90f4012ff94f4772c6e6dbe42a98f4f05cfe355eea2be01de03bf3c695

Observation db8a38bc-31e6-4979-b732-f065c3f0a96f · outbound

This paper cites You only look once: Unified, real-time object detec- tion.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry You only look once: Unified, real-time object detec- tion

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.746401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.746401Z digest=sha256:da5e8433037435562ffab1f14354a77c2f85df2ef302ac92d244545c576874f0

Observation c2ea64d7-54ec-4967-92ac-74862def1255 · outbound

This paper cites CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance.

FlowMimic: Mask-free Visual Editing and Generation with Pixel-pair Warped Flow Field for Online Video Editing Data Generation and Modality Mimicry CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-01T15:42:04.839943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T15:42:04.839943Z digest=sha256:e3b60d9f784bf88ec86018ab464588aac2b9914d099d4877535d76131e92ed9d

Pith citing papers

No inbound Pith citation observations are available.