Pith. sign in

Paper Citation Record · LEDGER

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM

As of 20 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2504.12048.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.12048 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:30.402063Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a31ef44f-82f0-483a-b13d-993cf912ba2d · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM LoRA: Low-Rank Adaptation of Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.348607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.348607Z digest=sha256:19b2d10b50095f2f130add18ec1cf5784436d1dbe0b4464f6491f0effd531248

Observation 27ce3552-4bcb-4387-aade-d3c3e7c168cb · outbound

This paper cites LLM-grounded Video Diffusion Models.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM LLM-grounded Video Diffusion Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.359580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.359580Z digest=sha256:9bf7648c568e757c5f144bc35991cec14365e1b01a81380320aab45c2b2f1407

Observation da526b20-e2af-4894-9bad-7f2e8aa2df57 · outbound

This paper cites VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.364741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.364741Z digest=sha256:ee477922cdcad3aa92054c99a37f855889aafc00623b3cbea1095e8d398548de

Observation ff54fb46-5f72-42b3-839d-9545869c7874 · outbound

This paper cites VideoStudio: Generating Consistent-Content and Multi-Scene Videos.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM VideoStudio: Generating Consistent-Content and Multi-Scene Videos

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.368998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.368998Z digest=sha256:77ecaacdac1cff2b5caa40befe43d577125627027c0ef11238f01e38876b5f5d

Observation dd7a0f79-4d2d-474d-8480-b9688d2f0d77 · outbound

This paper cites FlowZero: Zero-Shot Text-to-Video Synthesis with LLM-Driven Dynamic Scene Syntax.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM FlowZero: Zero-Shot Text-to-Video Synthesis with LLM-Driven Dynamic Scene Syntax

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.373854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.373854Z digest=sha256:0f72a1f32d9513fd9371bda6a84dcdefcc63acc0734a4a4ff8538a7ee01d9b1c

Observation b5fb3320-2e0d-4221-a85a-972eb8a8e50c · outbound

This paper cites FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.377899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.377899Z digest=sha256:d930e35371cf825ddb3c5e969c4dd3317b3c4888da235acd46eaf3274fceda18

Observation ddf8c19b-3445-4aeb-944a-638b30d5a532 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.392437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.392437Z digest=sha256:cef0071112552be4f2e5177a55fb2ae4e8aa7e15553ae5b5868840274917f989

Observation 6bfc736a-5d05-4d71-97e2-eb53f2f5a157 · outbound

This paper cites an unresolved cited work.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-16T12:40:30.595418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T12:40:30.402063Z digest=sha256:038eced9ace98a7a86fba4400bbb201b65700e8203b1e7455ad1b4f23e9c736d

Observation b483865c-2068-41ed-9cff-d5ffa912eca8 · outbound

This paper cites Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Gen-L-Video: Multi-Text to Long Video Generation via Temporal Co-Denoising

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.397158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.397158Z digest=sha256:8d54dc0848674a93e9a48e9ebee3455c850e5a9f47b96c0c84ab9d86c43add24

Observation 42958e16-0f65-4ac4-bd16-a7fb9afa2907 · outbound

This paper cites SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM SparseCtrl: Adding Sparse Controls to Text-to-Video Diffusion Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.337376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.337376Z digest=sha256:cd633048e2a04f8811c5801222c5106f198e00e53dded3dcc9d0e1980952141d

Observation 2575fb15-52a2-4d50-9590-e8acf333a1a4 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Adam: A Method for Stochastic Optimization

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.353816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.353816Z digest=sha256:70a2ec5353a208c60d623f6e0b77a174e4507fe648749985126514ccc248bd5c

Observation eea5614c-0aae-4046-8abe-bb9a102a4791 · outbound

This paper cites In Medical image computing and computer-assisted intervention–MICCAI 2015: 18th international conference, Munich, Germany, October 5-9, 2015, proceedings, part III 18, 234–241.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM In Medical image computing and computer-assisted intervention–MICCAI 2015: 18th international conference, Munich, Germany, October 5-9, 2015, proceedings, part III 18, 234–241

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T12:40:30.612390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T12:40:30.387636Z digest=sha256:2d5c87d60c1c74010db866bc1bd1f256f43f1ec1987a6d5d5b7661a5505d3134

Observation 36ee8dc8-5978-4059-9353-03e61090f352 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.332391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.332391Z digest=sha256:79ddd9bbb22d33fa8eadf87ed8b421fd60796dd11e2533c1455ff415cdf6ea23

Observation 8a56eb9c-6696-46c0-bdb1-5aa14a55b02b · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.382392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.382392Z digest=sha256:26a75422ac2d1f49de475d6436f07f0489a84c7e7d6e2b896b1c28d32147455d

Observation 5cda7574-28d5-474c-9523-4135edf102af · outbound

This paper cites GPT-4 Technical Report.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.327338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.327338Z digest=sha256:71b1c2359da46267d19e64d02d67434a0f780d0930901274c6504365b002524c

Observation 433f0dc1-69cc-49a4-acb5-0f8ec08338d5 · outbound

This paper cites StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text.

Modular-Cam: Modular Dynamic Camera-view Video Generation with LLM StreamingT2V: Consistent, Dynamic, and Extendable Long Video Generation from Text

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:30.343410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:30.343410Z digest=sha256:2b36e01b2ee751451dc1f2003b419fadbcf9f63bc077df535236eb33219814d4

Pith citing papers

No inbound Pith citation observations are available.