Pith. sign in

Paper Citation Record · LEDGER

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis

As of 12 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 3 inbound Pith citation observations for arXiv:2412.02168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.02168 v3

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:50:29.054249Z

measured 85 of 85 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:21:06.510937Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T10:25:42.327411Z

Reference resolution

82 of 82 outbound references displayed

  • verified exact3
  • verified fuzzy44
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 048f2da9-36f7-4042-9c40-1b9d0b48ab27 · outbound

This paper cites https://github.com/black- forest- labs/ flux.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis https://github.com/black- forest- labs/ flux

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.545762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.545762Z digest=sha256:134658b74ea5ab2098b1107708eb97a7976d3e3b960d3311b1f235652c6e6aea

Observation b9f88985-297e-4105-be54-b6e1e90b2894 · outbound

This paper cites https://docs.opencv.org/4.x/ dc/dc3/tutorial_py_matcher.html.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis https://docs.opencv.org/4.x/ dc/dc3/tutorial_py_matcher.html

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.550937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.550937Z digest=sha256:497c7b0ea8591ab0e662b79cdbd364ff61060b069552dee0e24b4d585ea4f1c7

Observation 4c73a320-c909-4c0f-b923-ffe4debd1f3e · outbound

This paper cites https://help.runwayml.com/hc/en- us/articles/ 34926468947347- Creating- with- Camera- Control- on-Gen-3-Alpha-Turbo.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis https://help.runwayml.com/hc/en- us/articles/ 34926468947347- Creating- with- Camera- Control- on-Gen-3-Alpha-Turbo

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.361802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.555989Z digest=sha256:6032fa4ed57a3af396ede2a758ee6ad21be89895bfef65a960e48a5d22b47dad

Observation 5b47b529-5fb3-4b57-873b-a6fa2ad16035 · outbound

This paper cites https://github.com/Stability-AI/ StableDiffusion.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis https://github.com/Stability-AI/ StableDiffusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.347468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.560670Z digest=sha256:fb2c6c391e4983a1a1bf7da8f1b3c9e74888abf86399aa41f6154057d5f5c6d0

Observation 98e1ff6f-a233-41eb-b2c2-d8f6132ce843 · outbound

This paper cites https:// openai.com/index/video- generation- models- as- world-simulators/.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis https:// openai.com/index/video- generation- models- as- world-simulators/

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.332082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.565358Z digest=sha256:a44d5ca873d835f5ab0713e3d49393d9cfd6530e870849339840a6d2a68fb6c7

Observation 9bf87a6b-ceea-47a9-9e97-1161f6776e08 · outbound

This paper cites VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis VD3D: Taming Large Video Diffusion Transformers for 3D Camera Control

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.569888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.569888Z digest=sha256:6c6d6a2b5086cb35f824846c6729cf66f4bfd0a942038c6092d904a461eed4e2

Observation 3a8530e8-25b9-45cc-b074-db5796b2cc5b · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.574952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.574952Z digest=sha256:850337f9e65f71538d03164c724b2eeadb4af3f93e79da9e2241f84f3a7ef92d

Observation 0c4861b3-e88d-46bf-87f4-759c8b278c80 · outbound

This paper cites Any-resolution training for high-resolution image synthesis.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Any-resolution training for high-resolution image synthesis

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.315388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.580376Z digest=sha256:fde8e3e9a5ed5ead5c07acb1ebe9284cde9681058e3ef1d99c657fefdf347719

Observation eb39abca-f22c-45fd-ba85-f579cb0921fd · outbound

This paper cites Tutorial on Diffusion Models for Imaging and Vision.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Tutorial on Diffusion Models for Imaging and Vision

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.585060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.585060Z digest=sha256:bd4e466a1de9f6fcb94cb0066802dbb8e1fa33467c427b5d8bcae04eb6f231fd

Observation 034de7c6-1ccd-4ffc-883c-cf6ebfc1c566 · outbound

This paper cites Image neural field diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Image neural field diffusion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.299930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.589922Z digest=sha256:2f3107003f6c30b7b1f49a3b1a574b836ebd599c01c3bd14c7d708c377d70de0

Observation 11060a5c-c0bf-4b93-aa01-425fecee7e3b · outbound

This paper cites Boosting Camera Motion Control for Video Diffusion Transformers.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Boosting Camera Motion Control for Video Diffusion Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.594702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.594702Z digest=sha256:977de7901476682d69e81babf42fc3a7e3abac86ee028ac49c1398dcf98c161b

Observation 17376856-3fe1-41b4-b1b1-f1352a42e2e5 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:30.283736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.600246Z digest=sha256:da16be7bdadb61035f93b069635d198e9a5628526fc69240035feebafad2c09b

Observation 867bbd28-1b86-40ef-972c-32d31b633922 · outbound

This paper cites E.T. the Exceptional Trajectories: Text-to-camera-trajectory generation with character awareness.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis E.T. the Exceptional Trajectories: Text-to-camera-trajectory generation with character awareness

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.604794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.604794Z digest=sha256:54fa57bdebddbb7f584b30e0797b88055441afaa6a0ccedab087ddc0071f1c26

Observation 6119a0c1-4855-4087-a456-cc57d8c071b6 · outbound

This paper cites Fairchild.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Fairchild

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.268164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.609731Z digest=sha256:9d7360bc62a3a5c80fbb6195ca4028049183a01b19f9b8066a4ea574dff9f930

Observation 85d1dad4-19f9-48cd-a204-8827fe5b0bb6 · outbound

This paper cites Diffusion models beat gans on image synthesis.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Diffusion models beat gans on image synthesis

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.252082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.614372Z digest=sha256:d4bc5945e8a05a688374794a3b6a59cd3e3e6d25c416cb1eb9f6529397d6d05b

Observation 736ced24-7000-4dd2-b073-4374f6495253 · outbound

This paper cites Problems of dataset creation for light source estimation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Problems of dataset creation for light source estimation

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:50:29.420298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.618978Z digest=sha256:f2b46221d390596106ddf4e368826a0801e3d93deb1e05c982bd089aac676b99

Observation a086a210-ddc6-462e-b56c-b2a2c30be071 · outbound

This paper cites Camera settings as tokens: Modeling photography on latent diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Camera settings as tokens: Modeling photography on latent diffusion models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.236810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.623955Z digest=sha256:a2bbf96b4cd0b8d26d339587137aa3779ed96e67520a873a5d864e160cd2ea8e

Observation d81cd21e-035d-4814-a31f-d4f12f543139 · outbound

This paper cites Bermano, Gal Chechik, and Daniel Cohen-Or.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Bermano, Gal Chechik, and Daniel Cohen-Or

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.222177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.628487Z digest=sha256:59d48413124e29dbe64c83bf6f4b8ec7dd86241dee1693ad14059f4bc8979ef2

Observation 9d55154b-b32b-424c-bb1f-6cee67ecbd3c · outbound

This paper cites Concept Sliders: LoRA Adaptors for Precise Control in Diffusion Models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Concept Sliders: LoRA Adaptors for Precise Control in Diffusion Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.633150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.633150Z digest=sha256:1f208e64b857b9a6d2aedde26afd56e8fa3b73b1e888b54d415b904c3243ad57

Observation 9588c85d-06db-4458-9133-ee3614082cf6 · outbound

This paper cites Generative adversarial nets.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Generative adversarial nets

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.206279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.637945Z digest=sha256:49e58a37a5f17412f3d6494693cb33e1ceee6dccd7b8a10a9646e1e69b50158a

Observation e2f33f5b-2c50-4b5d-9522-8c4881847a62 · outbound

This paper cites AnimateDiff: Animate your personalized text-to-image diffusion models without specific tuning.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis AnimateDiff: Animate your personalized text-to-image diffusion models without specific tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.190480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.642871Z digest=sha256:926079e8c3e7ae7e0236975b4500c15985944f2d2d75ec86b88f1d50a7c7471e

Observation 7619e526-bb53-48b3-8b0f-86d15d1cd5f7 · outbound

This paper cites Hasinoff, Dillon Sharlet, Ryan Geiss, Andrew Adams, Jonathan T.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Hasinoff, Dillon Sharlet, Ryan Geiss, Andrew Adams, Jonathan T

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.175023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.647829Z digest=sha256:be64e1357fa9e20ba6d78816d877984088b13561d37ecbed9bfa0b223f2fedce

Observation 923a9cad-65d8-4db5-ae50-647b5e1cd08a · outbound

This paper cites CameraCtrl: Enabling Camera Control for Text-to-Video Generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis CameraCtrl: Enabling Camera Control for Text-to-Video Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.652454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.652454Z digest=sha256:04614d8126f471454009104472b3920ed1c8755853d1e13cdc9c654e30ed4289

Observation 8575858f-0d97-4983-a613-473eeac18237 · outbound

This paper cites Denoising diffu- sion probabilistic models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Denoising diffu- sion probabilistic models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.159427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.657450Z digest=sha256:4569af09a5417f5a6745fc1a948cfbaea18d0c10508edeb4f66a126ecff0dc97

Observation a4aeec71-697e-4a74-93ae-2ee83166db23 · outbound

This paper cites Video diffu- sion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Video diffu- sion models

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.143846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.662424Z digest=sha256:7e9bf17db2796ee4f6c4b8dcfc1bf5732e3cfea02d3f169f7478bf354031e073

Observation a136e7c6-13e9-4b0a-ae7b-d6c27c8842ca · outbound

This paper cites Training-free Camera Control for Video Generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Training-free Camera Control for Video Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.667086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.667086Z digest=sha256:38399b2c896fd6cd6825f1cefb6f0ffeee03424acdac72c4ec7f765306a1ac9a

Observation 8d1283ca-3272-4db3-8d4c-4bf43715339e · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis LoRA: Low-Rank Adaptation of Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.672714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.672714Z digest=sha256:0c957518fc92f2ce8d2c5efdbc6f26d73d6c5dfb8284f2a661e0bfb78d2ff086

Observation 2ba82ac2-3ce2-450b-a7ea-773d93cd46fe · outbound

This paper cites Cinematographic camera diffusion model.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Cinematographic camera diffusion model

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.128046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.677322Z digest=sha256:63c31f69d1019e592f45683e07c0a47e48f9b2a1fac2ded4c8f6061607643612

Observation 5a8c8869-2b9b-48e7-a945-bddec98493f2 · outbound

This paper cites How Far is Video Generation from World Model: A Physical Law Perspective.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis How Far is Video Generation from World Model: A Physical Law Perspective

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.681774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.681774Z digest=sha256:40f68fb2e013791e62f9c580ff3248cf03fddbc546e001939e0621af1961f356

Observation be34e894-446f-4bd2-8975-a5e0b4a69383 · outbound

This paper cites Scaling Laws for Neural Language Models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Scaling Laws for Neural Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.686354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.686354Z digest=sha256:b9e7726540df3bae6ac4f7f85e5f4557e115c08bf2b0326f0296307f14487cd9

Observation e292b237-e7af-4555-8906-53dff3680d2a · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis A style-based generator architecture for generative adversarial networks

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.112380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.691042Z digest=sha256:fa0667b798baee6b232b736181513b61c5a9aa5b73f0d944c3498c90d5ffe2af

Observation f326dfa5-b7ef-476e-addf-77099124c929 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Imagic: Text-based real image editing with diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.095954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.695374Z digest=sha256:275a87418976bf712fa9cce12eea54141703883467c1b47f8285051aa3f5aa1b

Observation ea41ac38-5073-459e-b5a7-b37d41361de4 · outbound

This paper cites An Introduction to Variational Autoencoders.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis An Introduction to Variational Autoencoders

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.699708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.699708Z digest=sha256:a97b629b815db204fb8132095ab81c63b31562c44fa66e0c204923d52039dc9a

Observation 632d3418-c198-411b-852c-56e7bc651303 · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.704285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.704285Z digest=sha256:46680c35910c34aaaba1ad68f86bc744d5740dbc58802cb18e3958780da7b202

Observation e4b7bb23-486d-4628-bdfc-01eeb526dfc3 · outbound

This paper cites Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Collaborative Video Diffusion: Consistent Multi-video Generation with Camera Control

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.823325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.823325Z digest=sha256:e3b39e64f7c6252dcae31838a47f803942974f2411645cae3ece0947cb0bb60c

Observation b00cdd82-94b1-4d75-85c8-e6ded8bf6628 · outbound

This paper cites GLIGEN: Open-set grounded text-to-image genera- tion.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis GLIGEN: Open-set grounded text-to-image genera- tion

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.079742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.828943Z digest=sha256:27c9a0e50f8d380613e7d603fbb5040ac3d0e6c23e94b7a44b8a06fbe66a5aa7

Observation 4712e760-fcb0-4c53-a36c-c3e0990ab0b4 · outbound

This paper cites Salman Asif, and Zhan Ma.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Salman Asif, and Zhan Ma

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.064204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.833455Z digest=sha256:6483b1d1e6aa0d924b9e35de5e1a177245320c0d5fcdbb990ebe3893fecda998

Observation ff9b2dc3-cb34-4267-aa5a-ab7dbdfff22d · outbound

This paper cites Visual instruction tuning, 2023.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Visual instruction tuning, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.048069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.838038Z digest=sha256:66bbc8d38941b11349e05473c27aff8530889cfa9d18569afd596e8d962f7260

Observation 203bce7a-5766-4b46-9dae-31ea0c29a071 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:30.032375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.842500Z digest=sha256:b1a71bc72000938138581e3b24170332ca709200efd37e4c3c410ff1199a21dc

Observation e9a834e4-3a04-4d13-9719-87c93db298b4 · outbound

This paper cites Understanding Diffusion Models: A Unified Perspective.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Understanding Diffusion Models: A Unified Perspective

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.847333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.847333Z digest=sha256:9d7afa81574436406f53c822793328447e7da3cfc6fa76939e7c54d56fb5f5b4

Observation 7119135c-7e6a-40e8-91f3-8e942863653a · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:30.016038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.852074Z digest=sha256:a7b0b0c71ace4cb0a76570c37b827201328761850d470e244bbdc0032816b454

Observation 17b2c3e8-d19f-4659-b4ab-52676a9496f9 · outbound

This paper cites CamViG: Camera Aware Image-to-Video Generation with Multimodal Transformers.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis CamViG: Camera Aware Image-to-Video Generation with Multimodal Transformers

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:50:29.241499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.856584Z digest=sha256:8df4397de47c316b0a3481366497d5e8ab19c0fdecc45398b6154362f02426b3

Observation 3ba190d1-a0a3-4b78-9759-704bc9851bf8 · outbound

This paper cites T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.861311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.861311Z digest=sha256:2233b81706b071d666f17b77256ed03baba43680b3619f25c1ba5b5416acba48

Observation 725923c9-2452-49b9-b3d0-205cd700ba88 · outbound

This paper cites GLIDE: Towards photorealistic image gen- eration and editing with text-guided diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis GLIDE: Towards photorealistic image gen- eration and editing with text-guided diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:30.000500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.866086Z digest=sha256:580a46a9eefb2e356369f4ab6ced2b6da42ea9b06c0f7c45bcf280c38223afc3

Observation eeaaa0d4-43e5-4462-a62c-46e00ead3a7b · outbound

This paper cites Neural camera simulators.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Neural camera simulators

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.984635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.870419Z digest=sha256:c08f6ce13ea14921eb32cbdab25d491c0d8ecb71d51c6bd027a4ea6b95743454

Observation 62c8028e-ceac-42e8-918d-2b6b73e72244 · outbound

This paper cites BokehMe: When neural rendering meets classical rendering.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis BokehMe: When neural rendering meets classical rendering

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.968150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.874979Z digest=sha256:afbff007ace4d0a2935ccf143a79f8ec14c114d1e4c6a3f2853f3993e5d9be78

Observation 7c7972eb-783b-48ee-a71c-4ad119ff9451 · outbound

This paper cites MPIB: An mpi-based bokeh rendering framework for realistic partial occlusion effects.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis MPIB: An mpi-based bokeh rendering framework for realistic partial occlusion effects

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.952060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.880452Z digest=sha256:780779383e330aecceb8b0a186eaef698a27bf7dd335a8dd8b049bed014af620

Observation d43e784b-6ba7-4509-9afb-f200e69deea6 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.936759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.884870Z digest=sha256:f975acfda95d55ef8d71a2c235b9e954021a986bbce21bf1ffbc2d9045f70e69

Observation 45382e95-ddbb-45f5-b236-a29d744828e8 · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Learning Transferable Visual Models From Natural Language Supervision

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.889088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.889088Z digest=sha256:723216ab9adba19faf9e4da2a26fda18b1de68de51c3eb523751b4a5e3477743

Observation 8ff350d0-9a7c-4425-8dd1-c5a2fd5dea91 · outbound

This paper cites Zero-shot text-to-image generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Zero-shot text-to-image generation

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.921558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.894388Z digest=sha256:e2c39297aad63217caf6f7689398cbbd301a230ded2d621ae923d6e1634f1e5d

Observation f277f32b-4ee4-425d-828f-f8de752bebff · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.906451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.898867Z digest=sha256:7ef8d77cf1789203484bf4c2a5b148ef43cd2636acd218c72c6a1b01fd47c394

Observation 6bd86eab-2347-4060-a23d-84aa82f47a4b · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis High-resolution image synthesis with latent diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.891609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.904027Z digest=sha256:d635376daf527fa993221dead6e0ca813b4858a7ee97626f527a1486e653f120

Observation 24351cde-7d27-4fe9-aec9-c089d959dcd6 · outbound

This paper cites DreamBooth: Fine tuning text-to-image diffusion models for subject-driven gen- eration.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis DreamBooth: Fine tuning text-to-image diffusion models for subject-driven gen- eration

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.875258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.908385Z digest=sha256:3162fc52f586814e03ae83bd6fb97a12261443e120f7323f6812656465ab7dec

Observation d023c56a-ce6f-4b6e-88fb-58794a33ae55 · outbound

This paper cites Sara Mahdavi, Raphael Gontijo-Lopes, Tim Salimans, Jonathan Ho, David J Fleet, and Mohammad Norouzi.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Sara Mahdavi, Raphael Gontijo-Lopes, Tim Salimans, Jonathan Ho, David J Fleet, and Mohammad Norouzi

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.858910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.913775Z digest=sha256:9a3ec4b966c218ef9e0c295196f65c05f53c99dd68743635b46e540d2afdcfcd

Observation 9ca4d367-6a02-4818-8d02-0dbb8b22b7cb · outbound

This paper cites Closed-Form factorization of latent semantics in gans.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Closed-Form factorization of latent semantics in gans

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.842963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.918503Z digest=sha256:1d5d524898c75db8312af023fb471c9ef75462842913e016c0f75165accf9d5d

Observation cb723c64-78ed-4c21-bcf3-f58a01442fe0 · outbound

This paper cites In- terpreting the latent space of gans for semantic face editing.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis In- terpreting the latent space of gans for semantic face editing

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.824919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.924338Z digest=sha256:5844e53f28dc0f90dcb0d986dbd23b501647d230adc9f1f12e8d26f1b493495e

Observation 0395b17b-2c16-45a5-85bb-25ebd470b8d0 · outbound

This paper cites InterFaceGAN: Interpreting the disentangled face represen- tation learned by gans.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2020.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis InterFaceGAN: Interpreting the disentangled face represen- tation learned by gans.IEEE Transactions on Pattern Analysis and Machine Intelligence, 2020

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.809388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.928905Z digest=sha256:9acb38ef056ae45526b16d34da755822654975081af5b8c640504a0331b01a3a

Observation bdc13250-6836-4228-aa87-1beddb102301 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.793696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.934664Z digest=sha256:40abea4e9476b547e0265842b81ed8436b06e01030c4996336cc5af333760e31

Observation 6913cd7f-e9dd-48d8-9380-e0999a1ea3cf · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis MVDream: Multi-view Diffusion for 3D Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.939856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.939856Z digest=sha256:c843f618363f5fc6f0cbd2f87c91944de410a55122c8d15f683c5be4b0129bd3

Observation e4c49a4c-e885-411b-906a-7996e90d6e13 · outbound

This paper cites Score-based generative modeling through stochastic differential equations.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Score-based generative modeling through stochastic differential equations

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.945611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.945611Z digest=sha256:81efbaba77b26e58228dac0dd921e9078b76c1440f0c7e7650d25c2712c534a8

Observation 340fdd59-d3d8-4780-b449-8e171f88f66e · outbound

This paper cites DimensionX: Create any 3d and 4d scenes from a single image with controllable video diffusion, 2024.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis DimensionX: Create any 3d and 4d scenes from a single image with controllable video diffusion, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.767980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.950910Z digest=sha256:68074be077c95dcef789ae86a4bf6aab4eb8fcb14f4e09dd9925ec602064bec2

Observation 34983e97-cdf2-4e59-a3b4-c1a72d625d81 · outbound

This paper cites Designing an encoder for StyleGAN image manipulation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Designing an encoder for StyleGAN image manipulation

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.753090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.956153Z digest=sha256:8a6845213f34f38276fc5e18e48e1d8437d49330dd03b6a545940f86e3116bba

Observation 39d9edd6-3b3d-490e-a787-e1b800ff7cd9 · outbound

This paper cites Mo- tionctrl: A unified and flexible motion controller for video generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Mo- tionctrl: A unified and flexible motion controller for video generation

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.738002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.960966Z digest=sha256:d8f77d28dcf627cc9f5127404bf8295dda860d7d7ab739d4602823580d883268

Observation d10964d1-a9f8-4c9f-976e-bcbd5a4411fc · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.723202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.966370Z digest=sha256:a9a680cebc8ed50afdff3085daa39b220110cf5291ed9ee2499d423531b3ec3d

Observation 054bf9cc-b0ae-4611-aada-dfe16c4d0468 · outbound

This paper cites Learning images across scales using adversarial training.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Learning images across scales using adversarial training

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.708085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.971048Z digest=sha256:29c39077eee6c48303965f22cf83d26ce0844ee5380b8c94a0f239ed23c5e00f

Observation 20bc80d4-ea8e-4546-bc21-76d058e51e96 · outbound

This paper cites Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Contrastive Prompts Improve Disentanglement in Text-to-Image Diffusion Models

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-08-11T23:50:29.168224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.975810Z digest=sha256:6f918e1d148ba0a6c927c49ed1034d87d09a04528e6845f6c7c4a5d0cd6ca61a

Observation a32957c0-20a5-4ecc-b15a-f4e9606b666f · outbound

This paper cites Uncovering the disentanglement capability in text-to-image diffusion models.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Uncovering the disentanglement capability in text-to-image diffusion models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.692679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:28.980624Z digest=sha256:fdeadfb929a14536d02b62e38455c160b817013f87e19f47ed81cc528710f127

Observation 84335133-5ec6-4f9c-8070-12f7a350efd5 · outbound

This paper cites SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis SV4D: Dynamic 3D Content Generation with Multi-Frame and Multi-View Consistency

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.985077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.985077Z digest=sha256:b87e26bf5479be7de167d0e846c376ade3f8de843c40ef27cd22801551b145dc

Observation 0a897c70-3e3b-432c-af77-6f62b006ec81 · outbound

This paper cites Cavia: Camera-controllable Multi-view Video Diffusion with View-Integrated Attention.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Cavia: Camera-controllable Multi-view Video Diffusion with View-Integrated Attention

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.990356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.990356Z digest=sha256:8c164eaaed12403dfc223e588f9b22825f526a736b9e42b4eb54d2302425370f

Observation ccc04bae-6d1a-487a-aaf8-e5c3afae8428 · outbound

This paper cites CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis CamCo: Camera-Controllable 3D-Consistent Image-to-Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:28.995853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:28.995853Z digest=sha256:75408df208f3a36d868cd3e553bea868735a1c156b80e80f54674782a53ba4c9

Observation 724e2f7a-88c8-400d-a2a4-5ad403a78b72 · outbound

This paper cites Depth Anything: Unleashing the power of large-scale unlabeled data.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Depth Anything: Unleashing the power of large-scale unlabeled data

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.677382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.001324Z digest=sha256:6b545c8c0b2b3ffb78033b7d9d4fc65dd17bc61ff33acc5d0fc13a53aee4bf6a

Observation e9254725-9986-4753-ae80-e4ea11318693 · outbound

This paper cites Depth Anything V2.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Depth Anything V2

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T23:50:29.005899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:50:29.005899Z digest=sha256:850a7ba9d1cc329f624cf55c0fab604ee6110866153dd4bdd0c469629a2e80cc

Observation 22053b17-ef4a-4ca7-9ea3-b4b9b4d181e9 · outbound

This paper cites Direct-a-Video: Customized video generation with user- directed camera movement and object motion.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Direct-a-Video: Customized video generation with user- directed camera movement and object motion

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.661462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.010622Z digest=sha256:6db4e3e32e8b309da06979e239e82186ee0abf726bfcac61200d7cb56fc746a4

Observation d5c74c11-63b8-4856-9d47-91a7d75a2495 · outbound

This paper cites Adding conditional control to text-to-image diffusion models, 2023.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Adding conditional control to text-to-image diffusion models, 2023

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.646458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.015241Z digest=sha256:547a63fce0f475d2df594ca29a0e79e09defbe7632d9954c612bc1ea876e0fbf

Observation de7aecce-0a15-4c3f-b634-73aed18e16a0 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis The unreasonable effectiveness of deep features as a perceptual metric

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.630931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.020411Z digest=sha256:b9741a6bf11bf5fc760b2587f9669506c822eba6576d5f100780e6cd38603d52

Observation 1c8ab4f4-1a9c-4fb8-b009-84ebab7a3f2f · outbound

This paper cites Zoom to learn, learn to zoom.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Zoom to learn, learn to zoom

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.615917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.025052Z digest=sha256:f0691a5cef0cd941fc18607b9643bdff4db5acf7b5c17eea2cc21ab875379b01

Observation 739ee02b-8cc1-416d-a6ab-57ea5be40ce9 · outbound

This paper cites Synthetic defocus and look-ahead autofocus for casual videography.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Synthetic defocus and look-ahead autofocus for casual videography

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.600460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.029489Z digest=sha256:4eeb71b1a7306644a3d884df19b177534708803a4953f28dafdaa1855a478ac4

Observation 1c3192ba-8b15-440d-9895-029d193aa04b · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.581687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.034110Z digest=sha256:f9f86e1bf1b5d2b27e9ea092a8e3fc147e9b0447aa4e630ccce9d06701dcd97e

Observation 1c0f0e94-af57-4153-9bf4-56402cac3f83 · outbound

This paper cites Camera settings are sampled during training and simulated on-the-fly using physical principles, producing differential multi-frame data without pre-storing large video files.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Camera settings are sampled during training and simulated on-the-fly using physical principles, producing differential multi-frame data without pre-storing large video files

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.565988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.038943Z digest=sha256:e8b196c8aa8a5954e76b282fba0f97c0e090b03290bc98a3b35081010fe4c343

Observation e1a33af1-1b68-4195-8015-a595e14aceb7 · outbound

This paper cites We extract the camera settings for𝐹𝑟 frames using the CLIP text encoder, compute the differences, and then reshape the result into an embedding of size𝐹𝑟×𝐶×𝐻×𝑊.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis We extract the camera settings for𝐹𝑟 frames using the CLIP text encoder, compute the differences, and then reshape the result into an embedding of size𝐹𝑟×𝐶×𝐻×𝑊

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T23:50:29.550151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.044595Z digest=sha256:95fa529378eaa22841371c3d44287426c0ca9b36b5ec3b8edb567a3ea8d1aaa9

Observation cfe3b4c1-c959-4209-ab16-3b6569d468c3 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.532183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.049499Z digest=sha256:41066059de44445b74dd4c3ba183327360ca4226521f3ef9b3395d73017967cb

Observation 4037ac18-012e-4261-a7c5-eac45c040da6 · outbound

This paper cites an unresolved cited work.

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-11T23:50:29.515120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T23:50:29.054249Z digest=sha256:0983d1129bcb3c54f4d208c46aa784a220c9b03c013e4f8e28bef770c1480256

Pith citing papers

Observation d6e15a76-b3f8-43ac-a271-3547a02ecf41 · inbound

Wonderland: Navigating 3D Scenes from a Single Image cites this paper.

Wonderland: Navigating 3D Scenes from a Single Image Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-11T14:21:06.510937Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:21:06.510937Z digest=sha256:ae7256618f000861235449ffcc45f6158ffc970ccf9be73c5014d997c53c2643

Observation 6f9ea93b-cdcf-4724-8018-85e872f5153f · inbound

Less is More: Data-Efficient Adaptation for Controllable Text-to-Video Generation cites this paper.

Less is More: Data-Efficient Adaptation for Controllable Text-to-Video Generation Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-17T19:55:10.008039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-17T19:54:35.190865Z digest=sha256:54db5112e0ca5b62eb46e364941e07d4145783b3af5f7dccb7f8def115c4a59f

Observation 97fbd3ef-4d81-4830-ab87-fd5229489552 · inbound

AnyBokeh: Physics-Guided Any-to-Any Bokeh Editing with Optical Fingerprint Transfer cites this paper.

AnyBokeh: Physics-Guided Any-to-Any Bokeh Editing with Optical Fingerprint Transfer Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:25:42.328928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-07-01T05:27:33.456926Z digest=sha256:995dcf683eb3e238775e71a1df285cf21fa45542d6bac8ec767d2bcaeffefd4d