Pith. sign in

Paper Citation Record · LEDGER

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

As of 7 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 4 inbound Pith citation observations for arXiv:2506.03126.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03126 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:12:11.643982Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:23:05.876881Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T02:40:14.528606Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 389f885a-2c05-4557-aa24-30917415e80d · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.596205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.596205Z digest=sha256:340709d569ab9c39ba6a892b780c57f903e584f4bb7f601840dba1daca33f5ea

Observation dd058a13-c529-4a70-ab25-e2538c4a7612 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Activitynet: A large-scale video benchmark for human activity understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.662318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.662318Z digest=sha256:25bed66cd9b76bca270dc9bcfc51967af9eaee9f35b3f47c4f1d4f84a6863e17

Observation 8af38723-941f-4c8c-a6f2-d8a21683ed7c · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.730235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.730235Z digest=sha256:3be99c4c23e71c0f06bbc15461d33781d612e207e22d823c0479a35760b4dc2b

Observation da75d926-c224-400f-916a-84cd7e90880d · outbound

This paper cites Multi-subject Open-set Personalization in Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Multi-subject Open-set Personalization in Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.805412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.805412Z digest=sha256:3b5dffc54f1849abb978bfe8077ede12c8486bfd852b6f93c51bfe7d8068cdaa

Observation 7a32cccd-bf7b-4da4-9bb8-7731c437384d · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.888630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.888630Z digest=sha256:074b6af96d9a8d9d3f43856b195be787d40f6402dad58c2ef6d8189a4d6b6ee1

Observation 2ad33bc0-f505-474e-b8ce-99f722b5d127 · outbound

This paper cites AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:12:12.106906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:07.996721Z digest=sha256:7392c337288f35a673ad937d8e10d4728acac1086f6079dd5ba814aa5f1d90d5

Observation 42f72ecf-dc28-47aa-828b-08c5f6d867ff · outbound

This paper cites Gemini, 2024.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gemini, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:14.189027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:08.133677Z digest=sha256:71de3b607545bada00c772adbe93aa3a535ded73fa50797bb241cd5efe1a6391

Observation 10bbdf14-7a25-4c48-bb27-3789b2a20813 · outbound

This paper cites CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.227600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.227600Z digest=sha256:3ba7b0a5961c0cc5294b145df96a7694effc135f4cdae8a4b5c2b48e7380f768

Observation 374fa6d5-8d77-432a-8862-f6ea8f81c154 · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.339057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.339057Z digest=sha256:7dfb44678c650153f4faedf2acfd0095b612b1680db37dd863c46f97ac9bfb2a

Observation 8a36178e-afd2-410e-a11e-03b8c7f5c3c3 · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.485043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.485043Z digest=sha256:68661b417695f0dfa240a0e74c22493101699feadfb6ba416875b5dc7ae77f81

Observation d64ba14d-5ff1-4bed-9806-cc8749ca3d35 · outbound

This paper cites The Llama 3 Herd of Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.630129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.630129Z digest=sha256:b7076b4df34573fe858afc2a085a509825ba3fdec9c55718bd1d92911d1b0639

Observation b90d077f-4e9f-4da3-a746-15ea5e9d7b81 · outbound

This paper cites Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.729397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.729397Z digest=sha256:67aff1957004ebdfec73fd104db44c137daa97b8f98b550651cbbee061ffe929

Observation 9356d9de-e9d4-44bb-bcb6-f78c65570da9 · outbound

This paper cites ROICtrl: Boosting Instance Control for Visual Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ROICtrl: Boosting Instance Control for Visual Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.846411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.846411Z digest=sha256:3283cffd9acd690aaecbf24b91bd2ea3928cce9a33abb9cc37efa063cd060533

Observation f573cba1-365b-462e-8c50-832797999e3e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.909571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.909571Z digest=sha256:153d112f45f61b0d89294442d6c35d583914fd46e7b39e5e9f4997a61bd1cad5

Observation 0ffec6cb-a1a0-4858-b45a-43d583ac1ff2 · outbound

This paper cites Long Context Tuning for Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Long Context Tuning for Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.965970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.965970Z digest=sha256:f71086cf2802942bd470152e885d33ab689cb59f44d75918f72404f9788d5156

Observation e051b2b5-5fe8-4eb4-944e-760793cdaaea · outbound

This paper cites AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.048829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.048829Z digest=sha256:07bf0f3f597022b3eb4d26f6be7c2198d8a6b01438061a1b68c4eea80766000b

Observation 7fa80870-4be0-4fc5-a828-286d8f634c71 · outbound

This paper cites ID-Animator: Zero-Shot Identity-Preserving Human Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.126710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.126710Z digest=sha256:6940618938265bb3b0002dd7c129404b6c917697730dea1f1428a7ea17482f6f

Observation aaf0aeb4-19f8-4519-9e5d-d7174043c9cf · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.189653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.189653Z digest=sha256:6d90dcb480fd59db9b470d35d1059b0827630132d85c40cb3d06a0315f99ac9b

Observation a40c66a4-0d8e-408a-a927-93464dff33f3 · outbound

This paper cites Owl-1: Omni World Model for Consistent Long Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Owl-1: Omni World Model for Consistent Long Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.266583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.266583Z digest=sha256:939b702743bb57471bcac9dfae8197eee2ecf473333f33588e5900caacb7aaae

Observation bc92ba84-61f2-483c-bbe5-3ecf2d2c648f · outbound

This paper cites ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.347154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.347154Z digest=sha256:73ef64f2055d6189b6f25762002b984a9b0c06774414477e39473710d25773b7

Observation ae659b54-8ae6-4d0f-8e35-7eb036357a6c · outbound

This paper cites TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.446928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.446928Z digest=sha256:c4fdef3da640e7dba71a230530534c938d7ed8487506b7018cb3a9889f4e7f73

Observation 75efc4a0-733f-4862-8af8-bcd8407e95e1 · outbound

This paper cites AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.494420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.494420Z digest=sha256:42970f216ca8087f7b9924a7e67874549b99601124a922f7e4b0d0fa206bfd75

Observation eb25d9f7-ac88-4b57-a016-a3eefcb212da · outbound

This paper cites Videobooth: Diffusion-based video generation with image prompts.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videobooth: Diffusion-based video generation with image prompts

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.999178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:09.571073Z digest=sha256:3ab5e882604362483049800115e0de48cf9320738d74a8de4379460e8a2ee886

Observation cca380a1-eb77-432b-aa22-f71f09fbdc11 · outbound

This paper cites Miradata: A large-scale video dataset with long durations and structured captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Miradata: A large-scale video dataset with long durations and structured captions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.803816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:09.665314Z digest=sha256:7fa8a0eabea311fed77cf8a1c612c1df7b3d664f49dd6ebd4b95970b6c3b2519

Observation 35e5e471-8e47-4d3f-a3e5-44aaacf52b8a · outbound

This paper cites Animeceleb: Large-scale animation celebheads dataset for head reenactment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Animeceleb: Large-scale animation celebheads dataset for head reenactment

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.699408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:09.757488Z digest=sha256:72123ca767891ce92210acd195641cad0e1bbf557df0023e262cbb9f25b40f01

Observation 24929bb3-2b7c-47be-be85-b7fb14f89aa9 · outbound

This paper cites Segment anything.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Segment anything

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.813019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.813019Z digest=sha256:4939aab5403d00a612f2c0da00643cf2eb34635a8a400a0ab8edde7dc5b5f705

Observation a86e30b0-d6d0-4c61-96be-3e92e6e1f0ac · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.891514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.891514Z digest=sha256:391da1be318ab9888fbb17b3331acb12c8dea1b96e945f0520ec60fe091364e6

Observation 36f8d7e7-a113-43b8-a597-b8278264c7b6 · outbound

This paper cites Anim-director: A large multimodal model powered agent for controllable animation video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Anim-director: A large multimodal model powered agent for controllable animation video generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.577772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:09.969711Z digest=sha256:ee6e0b558843914f827f7a1c37968f350097d34d3c98144813d039c8216d9d51

Observation 820661d1-558e-4c66-b64a-1d735b45ff39 · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Phantom: Subject-consistent video generation via cross-modal alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.052472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.052472Z digest=sha256:ba3369507ab8784d0575f1ac0595eca1b5b7ba6de73bebd805619f65c94cfb1a

Observation bfb470b3-7a96-427b-b7df-d31c0b71d6d1 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.409848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:10.127134Z digest=sha256:e2cc0e4198388585d8d791bd4bb9d7691591db09bc5c41a01ab4163b04c1a559

Observation 54f3858b-f0ae-4f61-9f1b-9fe9c8f76c74 · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation NVILA: Efficient Frontier Visual Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.219918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.219918Z digest=sha256:6538ed67903170dde660c692165e74bf7c35ce94945c37a3a8aa5cae4542f744

Observation ec6cd1af-2c39-4fe7-ba93-a9203c189b99 · outbound

This paper cites Videostudio: Generating consistent-content and multi-scene videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videostudio: Generating consistent-content and multi-scene videos

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.247077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:10.299362Z digest=sha256:d535a920037fc445610dc9c789286dc50a2ad3d8d199d62b74f4df45b899d1aa

Observation 07fdbe30-a898-4f35-893a-7e0a49723455 · outbound

This paper cites Gpt-4o: Multimodal large language model, 2025.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gpt-4o: Multimodal large language model, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.078271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:10.378355Z digest=sha256:763ceef376e48387f386b65060a2a39099c4e02166168bda0f94b421828743eb

Observation 8c0100f0-a661-4ea9-be03-c8bdda53aa7a · outbound

This paper cites Sakuga-42M Dataset: Scaling Up Cartoon Research.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sakuga-42M Dataset: Scaling Up Cartoon Research

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.466249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.466249Z digest=sha256:70e82a544dd48aaeeafc66ecabed0a40ab674c327c1b4e94101dc50fa1c13cfe

Observation 555b527d-8b63-4e0f-9bd6-be90d689a97a · outbound

This paper cites Scalable diffusion models with transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Scalable diffusion models with transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.547419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.547419Z digest=sha256:dab766f8b4d1b8a31c96e5681087aa4b42bd0545fd9a5b0bb034d16118125101

Observation db3ff834-c87b-4a72-8b25-6c91f900c8d6 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.655648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.655648Z digest=sha256:ef7c78fc1cdb6b7939acc6ba6fba31665838d50a496838dee5693e5c747b2a0f

Observation ef22b037-3bb6-4cbe-b768-acdf42d65130 · outbound

This paper cites Learning transferable visual models from natural language supervision.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Learning transferable visual models from natural language supervision

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.735546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.735546Z digest=sha256:dfbb2e16653da6784a6205ab9e90b7704e388ef8b97b5ece571bca4ddf06df33

Observation 67e8c4e3-c52d-46b9-9d04-ae9478b51db7 · outbound

This paper cites Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.866309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:10.816759Z digest=sha256:c9e3868fb76c6d18c359bc348d38a89441bf061faeb75eeb991a59a2882a33ae

Observation d9559444-4464-46cf-b125-2bc3146cfffa · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.899127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.899127Z digest=sha256:f805117b45b3d7b20c90e59df510d9aae820741ac3e097c4697c9eb5f305261b

Observation a3ff8ddc-5acc-4107-97cd-f55f3767ee50 · outbound

This paper cites Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.988953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.988953Z digest=sha256:fcff83d53e073b50f8de44935d83e0a91e55f824cfeebea1f7993c783d06c677

Observation 6eb2fcc0-3d80-460c-8e33-0dc1fa6869a5 · outbound

This paper cites Understanding animation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Understanding animation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.648824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:11.080666Z digest=sha256:5438d3cbcfe99192b7d4ddedddf3a22878daccc42627d8da018cb39f779c5cb1

Observation 6a6aa371-4938-48d2-90ad-025c4183ff3d · outbound

This paper cites Automated Movie Generation via Multi-Agent CoT Planning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Automated Movie Generation via Multi-Agent CoT Planning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.166087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.166087Z digest=sha256:72fbf9574210215c2760abe07495b4d286a3e7b04749727c05d15dcfe0f5d465

Observation 30381dfa-6f88-438b-a9c9-990b24562a6d · outbound

This paper cites Pandora: Towards General World Model with Natural Language Actions and Video States.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Pandora: Towards General World Model with Natural Language Actions and Video States

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.222249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.222249Z digest=sha256:041bfa317edb0e7e51619c8f3ed7baaeff7a72a50637d7a1892384e103719775

Observation f75afb69-c4ce-4946-a2e8-a6659b257310 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.492949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:11.286902Z digest=sha256:766de9b4345a7d3f026bc97a34d60db5ee0390b281aacf6121f00b11f8d43dba

Observation c906f12f-548a-4246-a263-f63b490a5dfd · outbound

This paper cites LVD-2M: A Long-take Video Dataset with Temporally Dense Captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation LVD-2M: A Long-take Video Dataset with Temporally Dense Captions

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.338933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.338933Z digest=sha256:7bb68594919cd3a1dd71e3bf31193b12133f0fb4afd47821e963267270a350e3

Observation 50a82ba6-d370-4019-b301-e78b5afa1694 · outbound

This paper cites Vript: A video is worth thousands of words.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Vript: A video is worth thousands of words

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.358110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:11.402600Z digest=sha256:069691a430c787b86f21a6b38ed86113e6783f02c3bc1cc33adc695780ce2a59

Observation bc59aa92-dab7-4113-ae87-2e9112b87892 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.474850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.474850Z digest=sha256:8f68302661a7002a17f3f543621d96f19aa12f4e47f1706a28b541fd518ad919

Observation 2be077c8-5812-4e90-beb0-51b75fdd6892 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.511487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.511487Z digest=sha256:efa5ed8e330ef5760ed7ce8faab9d480f01bdd44d59b0f86554b4fe984840e02

Observation babab955-deef-496a-a763-52bb4fe2f02c · outbound

This paper cites Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.581622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.581622Z digest=sha256:fad8ffb0d3970800dff2d1d2ac73aca289219ddd6b86802d4988f1b160811647

Observation 3056d301-ef7b-43d8-90cb-904e1dcc52be · outbound

This paper cites a cow is mooning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation a cow is mooning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.237359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:12:11.643982Z digest=sha256:4b9e4a7c6b391c26534d635abad23dd4fe0efd371b5bcf2453b9444e26a5c263

Pith citing papers

Observation 12a9d3e5-c4e3-4a23-a319-0bec9d33f43d · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:0833c8986d2587f42523002ab917f1aefb2bb23bebeb0e61ed3d911f60be6bdf

Observation 03dd8ea7-b774-48bf-ab89-aa092011185b · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:19.182772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T06:28:42.129881Z digest=sha256:f0cbce5904cad11c9bb48affbc35542f05a785dcb0896b5a3b37beee8804c99b

Observation 1cf7ef51-3212-4406-ac33-bd2199c83f1a · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:51:15.171755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T00:50:10.509727Z digest=sha256:84a63c2067c9763c051a753932659a5c5586926f3275e711698cc4c00774adcb

Observation 0495b8e7-e0f5-4b47-8fb6-7ae4308d815c · inbound

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches cites this paper.

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:40:14.532466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T02:40:00.873188Z digest=sha256:c0c5d3f1a1ef74e2ac836b0d6c0ef22f45936784417ef54cba81b5b3cb8f6efb