Pith. sign in

Paper Citation Record · LEDGER

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention

As of 14 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2412.03756.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.03756 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T22:12:35.666619Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact1
  • verified fuzzy18
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 39d3f69b-7061-4ea6-8533-05f76bddce74 · outbound

This paper cites Multidiffusion: Fusing diffusion paths for controlled image generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Multidiffusion: Fusing diffusion paths for controlled image generation

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.288303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.474549Z digest=sha256:0e1631184292f4757e69a1cfbb65fb19e34926db700c65b9748975ad62f1e3c9

Observation 20753695-8249-4216-98ef-5338d103ade9 · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Align your latents: High-resolution video synthesis with latent diffusion models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.270573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.481308Z digest=sha256:d55a9043c034b5c407089b5738153e71a4273de940b13dddd03a655c7c3ee468

Observation cbc3f534-3694-4acf-89d0-f34c5117f017 · outbound

This paper cites Matterport3D: Learning from RGB-D Data in Indoor Environments.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Matterport3D: Learning from RGB-D Data in Indoor Environments

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.487268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.487268Z digest=sha256:b6482b309bb47065ca833ac49229bebea0cea085044cf7d7338bdb444aeabde8

Observation 148561cc-ef3d-4251-b88d-39e182aa6bc9 · outbound

This paper cites Attend-and-excite: Attention-based semantic guidance for text-to-image diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Attend-and-excite: Attention-based semantic guidance for text-to-image diffusion models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.252913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.493080Z digest=sha256:c32ce4476966164522abdb898f544585a62020f44016788596d6441c05d6f577

Observation e9984c13-1734-44cb-9ce8-3a3f8f438455 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.234905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.498195Z digest=sha256:adbdc7cf4802f6a26bcd6775c097fe4e64d7a5ebe2b9ce85d216539911eb57ca

Observation c18e978d-16ee-495e-9201-8fb602cafbad · outbound

This paper cites Preserve your own correlation: A noise prior for video diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Preserve your own correlation: A noise prior for video diffusion models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.215626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.504771Z digest=sha256:35d0569e66366ecef778f88379e916531c40eddb29b5105a013715c7457a0ac2

Observation d26a3603-52ba-4a12-875b-104d59ca959e · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.510659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.510659Z digest=sha256:38632cbab4a9acd464bd7982e5db30acff2cb9e7bb945bed72c5cb56fd96f7d9

Observation 75c44599-4033-4d1c-b3c7-278434a95748 · outbound

This paper cites Reuse and Diffuse: Iterative Denoising for Text-to-Video Generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Reuse and Diffuse: Iterative Denoising for Text-to-Video Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.516441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.516441Z digest=sha256:ddf606c2f9433716bb035fd06dcc216843ce124c273707c43ea22da4a168f003

Observation 31c1c96c-aaf3-42a9-b9e5-dbb831685a71 · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Prompt-to-prompt image editing with cross-attention control

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.197889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.521913Z digest=sha256:5bbe84a966a2a75b25e84a57648c2de62f507eeda9d605ceae9f279874abd2e8

Observation ee30a7e3-2974-413b-a8c4-50a9b912f2c4 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.180447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.536137Z digest=sha256:0973c2b9096967b58079b3123ce99b2d2b155a8180c231ee055ea4b1250ab852

Observation 84b4858a-f0a8-4ca3-b55b-da51df3b0991 · outbound

This paper cites Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Text2Room: Extracting Textured 3D Meshes from 2D Text-to-Image Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.541973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.541973Z digest=sha256:dc15d2ea01061ddf9e3165edd7e4e99ce0d70ccb52b93cf3be4e8494907cb227

Observation b66cda66-ffad-4dad-8780-aa0af6d4e361 · outbound

This paper cites SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention SyncDiffusion: Coherent Montage via Synchronized Joint Diffusions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.547538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.547538Z digest=sha256:ae9d03d7a26046b46e277e48bf61b5d17469d49813c32cfc645335bc4c15f8ab

Observation 3d670d82-7b36-4d8b-8a21-072ef5c4b73e · outbound

This paper cites Common diffusion noise schedules and sample steps are flawed.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Common diffusion noise schedules and sample steps are flawed

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.162984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.554333Z digest=sha256:b7ce7b9b644f1d4764dc294106204074ee9710f8639397d43d10a85a14ba1c97

Observation 89b6d7db-44ef-4b55-8b69-f227506b59ae · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.559616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.559616Z digest=sha256:8b5e8ddf02049d2edbfd81ed79b57082f6e9a090e808a35d104629eabf583a1b

Observation dd1e31fc-0856-4626-8d38-57889c0450f8 · outbound

This paper cites Glide: Towards photorealistic image generation and editing with text-guided diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Glide: Towards photorealistic image generation and editing with text-guided diffusion models

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.142724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.565890Z digest=sha256:354ed13c9e40eff64e7c4b3d851c713b295c4105a73799f7d8880766153544ea

Observation 64262177-2fbf-4342-8f83-526e0daa1d7e · outbound

This paper cites Zero-shot image-to-image translation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Zero-shot image-to-image translation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.120847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.571475Z digest=sha256:a9d2f40faea07e4c5cb6da23115a3251dce8f192cdef6aa6b1ba257f819ccdc5

Observation a27b086c-d8d9-4b33-8ef3-bfe600815192 · outbound

This paper cites FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention FreeNoise: Tuning-Free Longer Video Diffusion via Noise Rescheduling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.576188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.576188Z digest=sha256:d26b8f77d87d0f27028df85c41bdaf419da481059ed752bdc952cd37c8d4e68c

Observation 1b4dedec-6a4d-43f6-b53d-5584fd3d3d90 · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Learning transferable visual models from natural language supervision

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.101972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.581407Z digest=sha256:eb4f92202d15b20dd200a6db4d32bc94ddc5e7f709556dd4be77e0de45c53b6c

Observation 772e6e4c-179f-4e3b-b84e-67610d05d568 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.586714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.586714Z digest=sha256:68fc7bb8c3b4569cd2736de0f02c69de7b2386af7c8b78d7aae83930a4d27473

Observation 87cd3522-8cfd-4da0-b788-34e2ff75fd81 · outbound

This paper cites ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention ConsistI2V: Enhancing Visual Consistency for Image-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.592721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.592721Z digest=sha256:e2fddf58f0a7b97e0c13ba85b0e6b692c256dc4d951ec0068df26b974c34abe3

Observation 74fbb41d-58bf-457f-bef1-bc8d7ce28070 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention High-resolution image synthesis with latent diffusion models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.082395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.598177Z digest=sha256:c065525a9820599427363e8692b0f0696f7f3c1eb0beed762f2901a581dc5ea7

Observation 45bcd2a0-035d-4185-8fe7-2cf23b120b84 · outbound

This paper cites U-net: Convolutional networks for biomedical image segmentation.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention U-net: Convolutional networks for biomedical image segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.064762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.603873Z digest=sha256:325c1a43371bca090c3631f9d3e54cb87782d32aace466832ca93c87f463a6fc

Observation eacd56c9-fd3f-428c-9f54-30203b667688 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understand- ing.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Photorealistic text-to-image diffusion models with deep language understand- ing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.046023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.609718Z digest=sha256:0f146dbd0bc11a114203e2ec4c27b42321012fddf4ad4011c4458783381b2602

Observation deeba151-656f-4c7b-a922-97b32852d700 · outbound

This paper cites MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention MVDiffusion: Enabling Holistic Multi-view Image Generation with Correspondence-Aware Diffusion

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.614836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.614836Z digest=sha256:16e274eb4a17f381d4579d411125e536d58eaad43419ed32b59594ffc31cb4f1

Observation a3b4e943-c53e-4723-b434-998baaf67da0 · outbound

This paper cites Attention is all you need.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Attention is all you need

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.621290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.621290Z digest=sha256:57f7e64e6a419baf7f5d09be1bb5b4cf2e88d4db028c0c1847e81a54cdce6ba3

Observation 84d09a35-daaf-4de3-8419-37e3430e031b · outbound

This paper cites Diffusers: State-of-the-art diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Diffusers: State-of-the-art diffusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:36.016032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.628239Z digest=sha256:6ebecff2a176d4ac92a18f2db80024728f69a937f88a86198b73dc35d6e91b5d

Observation 70fa4d08-60f6-491f-ab08-577ef7cc2513 · outbound

This paper cites FreeInit: Bridging Initialization Gap in Video Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention FreeInit: Bridging Initialization Gap in Video Diffusion Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.634946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.634946Z digest=sha256:5ad949f65b6f9662e9fd3b7cc02ab2a5e59fa37ed780476031b282014aff797f

Observation 470ca0e1-6415-4b6b-9731-4df782bfd3a3 · outbound

This paper cites Freestyle layout-to-image synthesis.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Freestyle layout-to-image synthesis

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.998158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.644795Z digest=sha256:f6348ded3a2125c1ec7b80c4ecd771af3d6365e6f328f6099ab3e9657d10ac20

Observation 5268a08c-3e08-42de-9f0a-baf32f2e8282 · outbound

This paper cites Preserving image properties through initializations in diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Preserving image properties through initializations in diffusion models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.979928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.650107Z digest=sha256:0d7120db4fb105c1355ab020171296cd170ecf6bd378699165c31828c79dd97c

Observation 5f4c8e9a-e8c1-4e23-bcbe-51b0ff19d1d6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention Adding conditional control to text-to-image diffusion models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T22:12:35.961042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.655501Z digest=sha256:8b90d1c4d2255e14d1d460a404de97df40ca0655907f420d8bf38bd12392496b

Observation 3381c663-2fcc-4a53-996b-2d7611e8075c · outbound

This paper cites DiffCollage: Parallel Generation of Large Content with Diffusion Models.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention DiffCollage: Parallel Generation of Large Content with Diffusion Models

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-11T22:12:35.718932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T22:12:35.661069Z digest=sha256:abe7be435b05e1a66b100309f04435ea0b31bf987c54071a6575b2414f30f51b

Observation ddbc1c79-3a4d-4fb4-a32c-2dd18097e10d · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Multi-view Image Diffusion via Coordinate Noise and Fourier Attention The unreasonable effectiveness of deep features as a perceptual metric

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T22:12:35.666619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T22:12:35.666619Z digest=sha256:07bea0fa419625fbc10219088b7396657f9a2274364fab25df5b004caf3f806f

Pith citing papers

No inbound Pith citation observations are available.