Pith. sign in

Paper Citation Record · LEDGER

A Unified and Controllable Framework for Layered Image Generation with Visual Effects

As of 6 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 3 inbound Pith citation observations for arXiv:2601.15507.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2601.15507 v2

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-16T11:54:46.989748Z

measured 82 of 82 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T21:41:27.903998Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T14:25:45.964685Z

Reference resolution

79 of 79 outbound references displayed

  • verified exact36
  • verified fuzzy40
  • unresolved2
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 67ae394b-e89d-481b-97c2-c813ed720a40 · outbound

This paper cites Qwen2.5-VL Technical Report.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Qwen2.5-VL Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.150604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:b1e6ff70c0e346d6d9ea9eb3c33e35b829d45bd2aa45454af15b6afc8a0ed2ca

Observation c23f8edd-376a-4338-93b8-ac7208a43ba9 · outbound

This paper cites Improving image genera- tion with better captions.OpenAI blog.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Improving image genera- tion with better captions.OpenAI blog

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.628525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:84535aefaa4d83712fbb7b5d17fbd2f5c2f44b42bcd4651f49b330e4dfadf732

Observation e40f2fb4-ae12-4f5c-ac45-369dd3fb5494 · outbound

This paper cites Instructpix2pix: Learning to follow image editing in- structions.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Instructpix2pix: Learning to follow image editing in- structions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.616102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:c457f50c23973aae00a20e3a64871a0d2bb869a332746b4d28012a93035f3b17

Observation c3e4773e-577f-474b-a6cf-5444bbf51da8 · outbound

This paper cites Pixart-sigma: Weak-to- strong training of diffusion transformer for 4k text-to- image generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Pixart-sigma: Weak-to- strong training of diffusion transformer for 4k text-to- image generation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.603461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:1daefbfac4f251a24cfa95f0cd75aa273355d7357aec93649e0cd4f96fc3d90c

Observation eb5c57aa-18f9-4342-b2f2-18ab64505394 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.201298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:d7a361e753a67a7b69135da5152e9cc26e9d7fa83a174ab9257d95e9d2950bb0

Observation 080c1d0e-c316-4d7e-b699-9eb34b841fc2 · outbound

This paper cites Unireal: Universal image generation and editing via learning real-world dynamics.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Unireal: Universal image generation and editing via learning real-world dynamics

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.605567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:356d640a00e60b6e9df142370275171e70e813b0dd9983bacb5fadab4a49ef73

Observation 4ea3da2a-1bbc-4272-82a9-df14a010fa72 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.198015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:fa1ab49837362d5a34afef9b48cda12bd854abcd9dcaaf45a87ed8ba5bd5cb09

Observation 4d8b62f9-d255-472b-942a-cb4a81aecf43 · outbound

This paper cites LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects LayerFusion: Harmonized Multi-Layer Text-to-Image Generation with Generative Priors

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.213284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:bf8aaad7ef57e62b0d961ec66cf3f76218e9a537b943e17299720d639a1cc99a

Observation c617c889-c8bd-4254-8cff-26a1b8218a6c · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.248036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:f46b54eeae2414f06dbceeecc66e6e9a5b45ba4895cdd6fedf1d9a20a19f0afa

Observation 33c61263-b67d-47c3-93e7-6ddcbe57937c · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Scaling rectified flow transformers for high-resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.650214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:dda0362771635f6e7b33ba67467dd4602ca5e90143b1df389b7017bd0ae8444a

Observation 261cb8e6-da08-489e-8237-705d81824696 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Scaling rectified flow transformers for high-resolution image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.644693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:1466e0317fd0d7554c6d2c8efad164f6fe743eb15c4fefd20554bcea1026c5d6

Observation 96b3def5-17a3-4030-b591-27c21e3d6680 · outbound

This paper cites Generating compositional scenes via text-to-image rgba instance generation.Advances in Neural Information Process- ing Systems, 37:43864–43893.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Generating compositional scenes via text-to-image rgba instance generation.Advances in Neural Information Process- ing Systems, 37:43864–43893

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.089153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:347159a5834170234ac28f661827d214bba17c69b339be870940fb3f4a36600f

Observation b9c9ce2f-7ad1-421b-9395-3792ee5c00b1 · outbound

This paper cites Multi-scale and detail-enhanced segment anything model for salient object detection.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Multi-scale and detail-enhanced segment anything model for salient object detection

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.636625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:803cc44733021e9df8680ce27634dbdec086247ae440af335e624700e198f96d

Observation 1507d451-6c64-4060-83a8-7e65a2f61f93 · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.227408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:8baf27be64b2227595d4f6a54875e037314f99aa2611d36a92e92f90e8b70927

Observation 63b59802-c70e-4fe9-b7c6-a145b3640c84 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neu- ral Information Processing Systems, 36:52132–52152.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neu- ral Information Processing Systems, 36:52132–52152

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.084649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:41803bc024790c28e964421f972da9cb4e600b1ac1a4773decdfbba31d4a0edc

Observation 89517e56-bfaa-4d8e-9f6a-86af0cd946dd · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information process- ing systems, 30.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Gans trained by a two time-scale update rule converge to a local nash equilibrium.Advances in neural information process- ing systems, 30

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.624307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:5f01814b532fa73754326f77fca10071ea3944e3c381ab6ac42d88f88786247d

Observation d479035b-8d7b-482f-9bcb-60eef5372259 · outbound

This paper cites Psdiffusion: Harmo- nized multi-layer image generation via layout and ap- pearance alignment.arXiv preprint arXiv:2505.11468.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Psdiffusion: Harmo- nized multi-layer image generation via layout and ap- pearance alignment.arXiv preprint arXiv:2505.11468

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.191370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:52bc4441f5a5ab276ef7aa91aa9dfea84c8c9414a905513452c5ac864e462bb9

Observation ecef2ef3-370d-4be4-9f0f-de73d465f687 · outbound

This paper cites Flux family models.https:// huggingface.co/docs/diffusers/main/ en/api/pipelines/flux.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Flux family models.https:// huggingface.co/docs/diffusers/main/ en/api/pipelines/flux

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.082267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:863fa9c7d2ea9f9545f26d7686709c0fc06115cc671403069d89216143a663b7

Observation 625994c4-4569-4e03-9352-24fd8c3ed0af · outbound

This paper cites Brushnet: A plug-and-play image inpainting model with decomposed dual-branch diffu- sion.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Brushnet: A plug-and-play image inpainting model with decomposed dual-branch diffu- sion

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.086894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:13b6e344ace8fd0fc7afbeebf0f29aa760cc12a0630d290da8149b683b061083

Observation f128df37-990b-4358-b610-a0915e62a7e3 · outbound

This paper cites LayeringDiff: Layered Image Synthesis via Generation, then Disassembly with Generative Knowledge.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects LayeringDiff: Layered Image Synthesis via Generation, then Disassembly with Generative Knowledge

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.259360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:77d089cd96670a6fd74e3d23c7e1d9556aefea6c51480318dc4e5a320f20f296

Observation 3973d0ca-1f5a-4cbb-ae0d-1c6ae6bac1f6 · outbound

This paper cites Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Marigold: Affordable Adaptation of Diffusion-Based Image Generators for Image Analysis

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.255751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:783399cd8abf6ac408a58d0dd213a7d1ca269cdc9650f745ec136140b33d214f

Observation 9c26ea02-d108-4858-90a8-c26b33eda0b6 · outbound

This paper cites Segment anything.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Segment anything

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.093289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:087cbdd0f7467fbdb8498bee93358ed536cc42a97906b20d89caa7d5f536662f

Observation 695c805d-27f8-4069-a20c-eed7933d5593 · outbound

This paper cites Flux.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Flux

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.079993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:91afff1576dd149300c0e691d098224a8725ab5954d0dce0342836191d127a28

Observation 893a470a-b52e-4132-bdf8-c3386355dda3 · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.240988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:39a242e1f10d85520af724bc307d28d0f3cc30c2158c37b099de3ea38597c1ba

Observation ea8bcc89-176e-477f-9e31-8b6fd13e24b1 · outbound

This paper cites UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.157306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:3b31423a5e1d410e1277e40042cf7ec9f76b5280724fa4d7d099d0adc75c8b8a

Observation a470ea8b-038c-4420-aaf4-2848f669fb70 · outbound

This paper cites Microsoft coco: Common ob- jects in context.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Microsoft coco: Common ob- jects in context

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.634204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:7d2d0ca114df787d1bcbe6ef0fa62014006bb25256f200d78cda565ec1059d18

Observation aa3214ec-21ac-4b22-95bb-af8efd71f97f · outbound

This paper cites Flow Matching for Generative Modeling.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Flow Matching for Generative Modeling

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.180996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:49018045c3be38f9e44cc967560eeb1ec326e8262235f4306c67959032aa8e2e

Observation 7a465e3d-8337-4053-8d1d-a62ee611188b · outbound

This paper cites World Model on Million-Length Video And Language With Blockwise RingAttention.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects World Model on Million-Length Video And Language With Blockwise RingAttention

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.216612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:94022ed9c84a47b4123250e2d5e26428c418fc8def8e9da46a5150c7ad422dfe

Observation fa367018-3a6d-4128-b2bc-a52e772891a2 · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Step1X-Edit: A Practical Framework for General Image Editing

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.177758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:ca3b38ed8074b8649c56135f3bb3fa4e2160cf9d24fb17046bcf3dfec3d6c5ee

Observation 745c99d1-8653-402f-8681-e9a1ed49985b · outbound

This paper cites Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:31:43.641168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:59cb98ec98193fed57082f1267ab42241809993d7fe944e5ecfe55816650437d

Observation b70eaf92-cb99-4941-b97a-6536680cd353 · outbound

This paper cites Gpt image 1.https://platform.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Gpt image 1.https://platform

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.095310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:a9dd09f84fa7f37677d1c4bf18b2e96d6ca6da5dc9e93f9e702c4e039276522b

Observation 08a076e5-fa43-46db-ae14-86bb26fd826b · outbound

This paper cites an unresolved cited work.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-16T12:00:53.630639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:6f7de0d0d2eb6d6eac3f836da42f946e23f27b0ce2ab3aa926fb5f89895018e2

Observation 48d2aa96-6749-4e74-bb0f-065ca4e665eb · outbound

This paper cites Transfer between Modalities with MetaQueries.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Transfer between Modalities with MetaQueries

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.174372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:bfd4fa30bb062c789a1897dd05ad1cb9c090ac57f16927356a31d7e2146dbf99

Observation 94466d44-3d95-4150-8477-6418fb5431b4 · outbound

This paper cites On aliased resizing and surprising subtleties in gan evalu- ation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects On aliased resizing and surprising subtleties in gan evalu- ation

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.618368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:e97ba9f69b571512abdc7f396f01f33f24c7d09a51fda7a5407cb4a5c870a0cb

Observation 5bca1c96-aa86-4571-871d-1508423c5142 · outbound

This paper cites Scalable diffu- sion models with transformers.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Scalable diffu- sion models with transformers

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.626244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:dad0436bbab8a7db2f11fdbc05ad5d6f6403a95ec0d63e97615ea1d2a0ff348a

Observation 4b1d5d41-d740-468f-9d33-5b7bbc57ba2d · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.184235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:249dd7448ee3fd187507a12d9226344c429bcf55dc56d15da5b29caf1d989fec

Observation 0fdac73d-d768-458d-bb49-7cab7fc43138 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.633028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:a1f3aeb70bbe1b8a4b12b76f32d9fc6641f91c473f6d86e4ec365951f1498e86

Observation b88afb4d-7579-4fad-ae74-0661592f29e0 · outbound

This paper cites TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.252224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:34a37b1f9771dd4dd611550ced636f302e336aabb0aa653938e46df526b84e1d

Observation 8744be27-5bea-4d85-b943-21011f870dce · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer.Jour- nal of machine learning research, 21(140):1–67.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Exploring the limits of transfer learning with a unified text-to-text transformer.Jour- nal of machine learning research, 21(140):1–67

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.647645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:49977882ff1363a7f23086e5d455a66407960c1b504a5a37fabfda0cfe3474ed

Observation e8840ccd-87fb-450e-a4a8-4768f0d4ba02 · outbound

This paper cites Zero-shot text-to-image generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Zero-shot text-to-image generation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.629428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:90bec7ee6c44c993ae0a97993716daaeec1ab747a7c0157b822b0662bfcde5ed

Observation 51a50ffe-3557-4c0e-908e-dd363d24d8b7 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.194336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:57427ca0ea6c28d302b32ced86e22909f8ba38723b5a45bd65af742e663a8ff1

Observation 3fece346-5606-4035-8ca8-fd240b882755 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects SAM 2: Segment Anything in Images and Videos

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.244408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:7276c4c9304d83f167ae9c36309282f27d42fced6740faa54828aefbe1b48399

Observation a0356251-d514-4979-8eac-588f98b7a94e · outbound

This paper cites Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Grounded SAM: Assembling Open-World Models for Diverse Visual Tasks

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.167584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:12eb53a431f610275bec38581975b10455dfdb0f929c34a63ca06978b36be6a8

Observation 2abb2a11-9900-441c-924b-4c68b76b80c5 · outbound

This paper cites High- resolution image synthesis with latent diffusion models.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects High- resolution image synthesis with latent diffusion models

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.075535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:8c9d67dbdb51230543f725f9bf9c52f4850136651bac77fd12a534b1fcabe996

Observation 72660d47-9c10-4417-bb72-148942bc69f4 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.237459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:7ae4477e95600f2c5f7a326900747a0849c2de43b4e61653f7bd438b49a05805

Observation 660da2a3-da14-468e-8d42-54b5f44b1ea0 · outbound

This paper cites Mulan: A multi layer annotated dataset for controllable text-to-image generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Mulan: A multi layer annotated dataset for controllable text-to-image generation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.099633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:ff431d16a9336438750d35a82284addf70a696b69fe943d016b2c82d854cb469

Observation e4920211-a657-45a8-8153-bc5f29c7b4c2 · outbound

This paper cites Unsplash: Free high-resolution photos.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Unsplash: Free high-resolution photos

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.639067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:96d8424a62dd6dcaada11e53876d6f0a0fcab418104065d504efe4ce2561b6e5

Observation 6831a40b-0b95-46e9-a4b1-5bd3a654a024 · outbound

This paper cites ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.209325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:d7c5f5d4cfca9a49f40df3e34cf3d0b53833547224014810d2135b9bee209209

Observation 1b67a0fa-693f-405c-93a8-f3c579525e5a · outbound

This paper cites Instance shadow detection with a single- stage detector.IEEE transactions on pattern analysis and machine intelligence, 45(3):3259–3273.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Instance shadow detection with a single- stage detector.IEEE transactions on pattern analysis and machine intelligence, 45(3):3259–3273

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.073276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:e3b352f86e28d26e2e97aa056ec1da62c249cb646a90f046849553fe880023f3

Observation e2ef605a-db40-4312-be6a-61d3dbd95dc0 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Emu3: Next-Token Prediction is All You Need

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.147690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:88cd561ae4075266bab1667243dc68fdfa06993ec63b46d02a2675f88af79f0d

Observation 2dd93760-c8ed-4735-8f11-c8c45b74e981 · outbound

This paper cites OmniEraser: Remove Objects and Their Effects in Images with Paired Video-Frame Data.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects OmniEraser: Remove Objects and Their Effects in Images with Paired Video-Frame Data

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.161193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:962e27466f3adb2b754aa6642d5a4abe51855356186e5f0e3f26a30f2ecafa0d

Observation 3c358a06-0ea5-4d28-8f77-fb0355be99e3 · outbound

This paper cites Object- drop: Bootstrapping counterfactuals for photorealistic object removal and insertion.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Object- drop: Bootstrapping counterfactuals for photorealistic object removal and insertion

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.068490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:ce1d896b92ee8fb044b2fa82ac56046330a7f425db01f74b04d15a009fd2bee0

Observation 0c92de20-fb3a-4e5d-a697-88bae47b4b87 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 53

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.164174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:75c6201f73da0c92a64cf2476648b2daeead07ea621d3eae0a4191220358cfa7

Observation ff000161-f91a-4438-96fd-b12ab4f1bdfd · outbound

This paper cites Qwen-Image Technical Report.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Qwen-Image Technical Report

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.140462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:73811a5729936a1cea06a8e58b05e668d8b017e0cbc8c18757241f9c6fee0fc5

Observation 0fde9377-a7af-4fc8-886a-23b5b077b765 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.137049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:fe82178ea55a887115e218ab14317915aa2e704f2e811497bd0ffdde5e932492

Observation f5b06fec-209c-45e1-89ea-5fe6b129c838 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.234347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:2ec7d3ce2dcef84239be3d3cf949f6ba863c5fe59ddaffc6b44da3fe7c960719

Observation 7744e0a6-6ef1-44c0-a454-13eff56fe452 · outbound

This paper cites Generative image layer decomposition with visual effects.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Generative image layer decomposition with visual effects

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.070955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:deb2a5962417fe9b8cfce8304e6c08e53f01263c65e36c61f324711b9c9a3a77

Observation 825b9327-b11d-4704-a225-1cfd4c49be71 · outbound

This paper cites $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects $\texttt{Complex-Edit}$: CoT-Like Instruction Generation for Complexity-Controllable Image Editing Benchmark

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.230872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:2d4461a272569c231d67d53266b34acdc061fc529a677dbfba0af778d61553d6

Observation 7fb50555-a184-4aa7-bcdd-7a78773d757f · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.144454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:de1945753141b6f97bc3361cd83f71694ba9628e6ac0b934dac1bfb9b4496443

Observation 6119f748-0f46-4868-9c15-89e22d184bba · outbound

This paper cites Anyedit: Mastering unified high-quality image editing for any idea.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Anyedit: Mastering unified high-quality image editing for any idea

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.063583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:0810434dc260653bbdbad811f3defc9207b6e14574b57349a916ad756af911dc

Observation 02d185b8-6c6e-4ee5-b03b-5ac7f2ba5651 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neu- ral Information Processing Systems, 36:31428–31449.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Magicbrush: A manually annotated dataset for instruction-guided image editing.Advances in Neu- ral Information Processing Systems, 36:31428–31449

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.065919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:06daabf75002cb272324c622ac0b24bd0c805f5bc27914fb77124aa09c8a904e

Observation 93b75084-304a-4469-bf32-0ed663278a30 · outbound

This paper cites Transparent Image Layer Diffusion using Latent Transparency.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Transparent Image Layer Diffusion using Latent Transparency

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.205421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:7049d677f9123b5ee3bc674251cf1ae875c3479b0d31980e211adae03dd1d889

Observation fd63b0b0-a2af-4fe2-9032-110269dbe034 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Adding conditional control to text-to-image diffusion models

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.077809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:b39e560ea8a34e98e11fd54954c8ca8cb35a2a00a7ce6908ba18f325359c1b83

Observation 98fd5801-6277-4739-8b29-a754f838eb19 · outbound

This paper cites The unreasonable ef- fectiveness of deep features as a perceptual metric.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects The unreasonable ef- fectiveness of deep features as a perceptual metric

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.621407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:bd76477087d85a6a48d8bc1f09120433e9bdc40f59a4b7d0096ed820923b1cca

Observation e4304bed-ad23-4032-baf6-652bf75217e0 · outbound

This paper cites Text2Layer: Layered Image Generation using Latent Diffusion Model.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Text2Layer: Layered Image Generation using Latent Diffusion Model

Reference 65

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.223899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:a4a77f8e4e85d7370c39ef7864a10711024c8c1e4a1c9d323da41f385cc23d3c

Observation 79982f45-67bc-4e31-8bc3-4cc076b602c4 · outbound

This paper cites In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-16T16:07:53.247084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:44a8bb3e8592e7908b507b4cbfe46ff5c0fc56e575bdd9aef09022b9add170eb

Observation c16b003c-2dc3-47f6-abd2-9d655e345a4a · outbound

This paper cites Ultraedit: Instruction-based fine-grained image editing at scale.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Ultraedit: Instruction-based fine-grained image editing at scale

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.091327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:763822a67b126b6f5170b5aa0915e9763e1effcfebfa5370ca31fea277a21de8

Observation b95cae79-5dc8-4c97-be75-189e020093ee · outbound

This paper cites ObjectClear: Complete object removal via object-effect attention.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects ObjectClear: Complete object removal via object-effect attention

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:57:50.187534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:aa8f9c2cafe06066927240f8fe2c3483c16196dbcecce96dd96fdfb0ce3b50e4

Observation fbd30a46-3fff-4a34-b3f5-a5e69f040abc · outbound

This paper cites Tp-blend: Textual- prompt attention pairing for precise object-style blend- ing in diffusion models.Transactions on Machine Learning Research.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Tp-blend: Textual- prompt attention pairing for precise object-style blend- ing in diffusion models.Transactions on Machine Learning Research

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.641799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:6b1fea7373f2f202e21b4c68c077995a940f75d15265db4dbb3897cb3ac2d17e

Observation 2c7f30ac-1ed0-4244-87a0-5431c7eb66b1 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:57:50.171082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:329ae24c56493a6024a67fc57680940199a36ff9ea236da91ec1e388c87c65d9

Observation 3bbc9cb1-8497-4818-b821-dad3700fc643 · outbound

This paper cites Addition.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Addition

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.592348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:aa9bd7491f5f5aab45bfd1bd51b37e10cfccf455bb26a425ccca264f9450caf8

Observation 9497fd29-402c-43b7-8ab1-c4a2f848c3a3 · outbound

This paper cites Addition.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Addition

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.061422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:9698851197bd927f7de98ba2b7c1721287b8767d942c8eb46ae5450d2d1ac63b

Observation b7837a65-fff4-406f-9a0a-19f989e6573e · outbound

This paper cites 11, we provide more results of our model under three different modes (FG Gen, BG Gen, and Text2All).

A Unified and Controllable Framework for Layered Image Generation with Visual Effects 11, we provide more results of our model under three different modes (FG Gen, BG Gen, and Text2All)

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.097484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:68a9ef1664aec508796aee7f5f3cb021dea9610b05b50dec85f31972402f0b36

Observation 4b8ae453-037b-46dd-ae30-3444619ad012 · outbound

This paper cites 12, we provide more samples from LASAGNA-48K.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects 12, we provide more samples from LASAGNA-48K

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.101606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:24e0663d21fe4048a12d3a29234b51a3a81f1ab9e45b450fc1db788ef126423e

Observation 22ba6195-69d9-461c-98bd-5e2512f85245 · outbound

This paper cites 13, we provide more samples from LASAGNABENCH.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects 13, we provide more samples from LASAGNABENCH

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T11:57:51.059433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:07f0b3a2301b526e5031e03d26cb8a9cc91426c5f4a5f5e051185b6489d4a3b1

Observation 529b7732-04bc-4bb5-a6eb-56a6361bc56c · outbound

This paper cites GPT Score.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects GPT Score

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.626810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:c38953f5282ba9e5f27f3b8542d02b9fe078398eed9c53dfa42a5a49825ffd20

Observation e42a5b44-1002-46e6-a43a-ad856e3fc0f3 · outbound

This paper cites The results show the superiority of our generation paradigm.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects The results show the superiority of our generation paradigm

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-16T12:00:53.616272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:8cecf4f5caa7765c4fd41ede44f309d6d0c6a59b1145b532399ead794754dd7e

Observation e07b29cb-a68e-4174-97c1-4712fbe83ebb · outbound

This paper cites an unresolved cited work.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-05-16T12:00:53.635083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:4d46dfbbfa1a2f1fccb90d1dad119e8d0a44d982d56bb04399e151ad2785d8c0

Observation c6a998a6-abc8-43c2-9a33-7f52a39b3e3f · outbound

This paper cites FG Gen” denotes background-conditioned foreground layer generation, “BG Gen.

A Unified and Controllable Framework for Layered Image Generation with Visual Effects FG Gen” denotes background-conditioned foreground layer generation, “BG Gen

Reference 79

Resolution
malformed identifier
raw_fallback, observed 2026-05-16T12:00:53.611488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T11:54:46.989748Z digest=sha256:51ef4c6c9f05cdf44eeed3ea199bfe46855acaa8bc0c5ffb849dc306f17e24c7

Pith citing papers

Observation 3c0945a2-e5d0-4976-bd99-9561d6d2a9ee · inbound

LiWi: Layering in the Wild cites this paper.

LiWi: Layering in the Wild A Unified and Controllable Framework for Layered Image Generation with Visual Effects

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-15T02:03:29.118096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T02:02:08.897553Z digest=sha256:11ff49ce1742a75935c14411c3b42c69dc0730d3ac0447e8f198dedff5c0b91b

Observation 380cb195-4f5d-4e21-8b0d-b1ad54983eb9 · inbound

LiWi: Layering in the Wild cites this paper.

LiWi: Layering in the Wild A Unified and Controllable Framework for Layered Image Generation with Visual Effects

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-22T10:24:47.055987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T10:23:18.785888Z digest=sha256:0d4a03626c77748589a06da8ad0bb486696ccef336f780c85e4888769358d352

Observation 922ee30a-8615-4ecc-9587-f394c11962ec · inbound

LiWi: Layering in the Wild cites this paper.

LiWi: Layering in the Wild A Unified and Controllable Framework for Layered Image Generation with Visual Effects

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-01T14:25:45.965976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T21:41:27.903998Z digest=sha256:864dc5d51f693b9da347ce90db8b905f955f57f71308413088f82722388c5aee