Pith. sign in

Paper Citation Record · LEDGER

Preliminary Explorations with GPT-4o(mni) Native Image Generation

As of 16 August 2026, this Paper Citation Record lists 100 of 219 outbound references and 3 inbound Pith citation observations for arXiv:2505.05501.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.05501 v1

Coverage vector

measured 100 of 219 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:46:03.070172Z

measured 103 of 103 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T05:30:32.163273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T09:34:57.138862Z

Reference resolution

100 of 219 outbound references displayed

  • verified exact6
  • verified fuzzy0
  • unresolved94
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2f74821-6a5a-413a-b5c7-9576da5ff7dc · outbound

This paper cites Defocus deblurring using dual-pixel data.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Defocus deblurring using dual-pixel data

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.638890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.638890Z digest=sha256:077831e822dc0bdcdb84bb5b986fbee6a6463712cae5e50b0fc5cd49e46c6b87

Observation 329d17a9-d9db-4cc5-ae02-d9f49f503222 · outbound

This paper cites Learning to reduce defocus blur by realistically modeling dual-pixel data.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning to reduce defocus blur by realistically modeling dual-pixel data

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.644112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.644112Z digest=sha256:71baf204ddf824f8945c0f14164aa4e2b6fdd77367961d24c4fdd22a31930e0d

Observation 00c433ad-b03b-4ad2-a622-27d6839b8735 · outbound

This paper cites Ntire 2017 challenge on single image super-resolution: Dataset and study.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ntire 2017 challenge on single image super-resolution: Dataset and study

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.648456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.648456Z digest=sha256:2211138b6f6742b2e926e1863822e5048a19f3d3df1fc1bfb68e345b4409fe84

Observation 3aaddf03-7b7d-43ff-80ed-c46a7c64207d · outbound

This paper cites Dream360: Diverse and immersive outdoor virtual scene creation via transformer-based 360 image outpainting.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dream360: Diverse and immersive outdoor virtual scene creation via transformer-based 360 image outpainting

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.652848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.652848Z digest=sha256:5081ac3b04f9eb877639fed89d133ba5d4c15a5afbc0281448a0162fafced8d2

Observation edf06fe7-0192-412e-9f88-721dccf7e23d · outbound

This paper cites Single-image reflection removal using deep learning: a systematic review.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Single-image reflection removal using deep learning: a systematic review

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.657073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.657073Z digest=sha256:4d0d677b78e07f5a83c48e093aa41105a1d4692fd3a94fb6c4bc465687c2d720

Observation c13648ec-b1fb-479e-bdc2-850edac22983 · outbound

This paper cites 2d human pose esti- mation: New benchmark and state of the art analysis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 2d human pose esti- mation: New benchmark and state of the art analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.661347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.661347Z digest=sha256:90784611128c03b1744bc1515de32270cd8c676f5b6ecced6059068526a07c5e

Observation 7ac9330b-c9ed-4d16-ac33-65b32d158f4d · outbound

This paper cites Change detection techniques for remote sensing applications: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Change detection techniques for remote sensing applications: A survey

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.666252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.666252Z digest=sha256:adc54a58b1163eae8c3bf5c84fa832f38fec5124fccbad4f2cb748c235364591

Observation 3ae4da77-3dd8-433d-b42d-94571df53b03 · outbound

This paper cites Rethinking inductive biases for surface normal estimation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Rethinking inductive biases for surface normal estimation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.670557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.670557Z digest=sha256:c2423f3cdfe34ae4af9cc55d4cc8e18c13b49e874fb515c1e2d61be2d3abb973

Observation 030a4bb0-bb52-4a85-a8a4-eb7fd7198e7b · outbound

This paper cites eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers.

Preliminary Explorations with GPT-4o(mni) Native Image Generation eDiff-I: Text-to-Image Diffusion Models with an Ensemble of Expert Denoisers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.674974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.674974Z digest=sha256:fafb0928072cb0bf0cb00c9919cd23720e426dfb68001485755fc3304695f2b7

Observation d0518307-9150-49df-a996-025b8665ed12 · outbound

This paper cites Diffusion Models Through a Global Lens: Are They Culturally Inclusive?.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Diffusion Models Through a Global Lens: Are They Culturally Inclusive?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.679724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.679724Z digest=sha256:9b09edb273bd92eab5c6c2e40c0ba2b3081ded7ab258e8638f5e6db2d40d3149

Observation dd4f7c4f-44d3-44d5-844a-01e6252cd2f6 · outbound

This paper cites Adabins: Depth estimation using adaptive bins.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Adabins: Depth estimation using adaptive bins

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.684520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.684520Z digest=sha256:ccea5e8e350e1c0bf640cefdafa4fd80f5ce4f8af9184fe8b6f781a88589d5ea

Observation d20c8806-8848-479d-becf-562849163015 · outbound

This paper cites Ledits++: Limitless image editing using text-to- image models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ledits++: Limitless image editing using text-to- image models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.688621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.688621Z digest=sha256:1c378646791a0402e33b82f9c42fee9bd3bd2e3594ed9a9e1d7c79834452a2cf

Observation 1099dc0d-9e68-429b-bead-154f303e1082 · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Instructpix2pix: Learning to follow image editing instructions

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.692633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.692633Z digest=sha256:53d1f8d0f54070fed07916156ed63896f4d919ece6e40de9e00a56215681b922

Observation 18e18725-8c52-4fe0-9862-9367d3146fb8 · outbound

This paper cites Learning to generate realistic noisy images via pixel-level noise-aware adversarial training.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning to generate realistic noisy images via pixel-level noise-aware adversarial training

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.696479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.696479Z digest=sha256:51b5b8a10e2379f375f59b56b4970f0a2d8658711a5d757faec99f91f7aca016

Observation 79f1762c-f203-41b7-b1d1-344874892682 · outbound

This paper cites Decoupled Textual Embeddings for Customized Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Decoupled Textual Embeddings for Customized Image Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.918935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.700930Z digest=sha256:7c1039bbce9c3cf15c64f095cc6ad44dd94c722dc605331ebcdd0d70d75492e5

Observation 1dd26251-8747-4f02-b62e-4449266e6874 · outbound

This paper cites Controllable generation with text-to-image diffusion models: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Controllable generation with text-to-image diffusion models: A survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.705414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.705414Z digest=sha256:f90ba1d058c79fffe8b5c00917524d70f42ee6ec3bebddd735fa46f5ce1e6500

Observation a28b3082-2557-4e4f-a468-60d249ca8477 · outbound

This paper cites End-to-end object detection with transformers.

Preliminary Explorations with GPT-4o(mni) Native Image Generation End-to-end object detection with transformers

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.709903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.709903Z digest=sha256:039606df1474ae31f5b26a5e4695b35f442be6db304aac164fee56593f85de6a

Observation fde5357f-4cf0-474c-a43b-7b0ad7bacbd5 · outbound

This paper cites Generative novel view synthesis with 3d-aware diffusion models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generative novel view synthesis with 3d-aware diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.714366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.714366Z digest=sha256:a8a588ad3ebb6f34ffc35c58b7b74b013ac7e17b2a2543478efda96f6b9758ea

Observation adf74f5c-5fb7-462f-915c-658c64b779ec · outbound

This paper cites Spatialvlm: Endowing vision-language models with spatial reasoning capabilities.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Spatialvlm: Endowing vision-language models with spatial reasoning capabilities

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.718644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.718644Z digest=sha256:ab5f5af9ed0118fca1c19e46e382798875a9b887f50a944a1624814674b6cbf8

Observation 73c824a5-04fa-4b49-9113-54dc16c64d83 · outbound

This paper cites ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit Adaptation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation ArtAdapter: Text-to-Image Style Transfer using Multi-Level Style Encoder and Explicit Adaptation

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.832209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.722900Z digest=sha256:da5910fc8204755e05a7237a4077b5ef7f5024084ffc0025f582d5361c399ea1

Observation ad939b62-f772-4283-b493-c7287b43361b · outbound

This paper cites DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DisenBooth: Identity-Preserving Disentangled Tuning for Subject-Driven Text-to-Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.727261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.727261Z digest=sha256:022f6b156ad9f881d565d3fac6605ef61dabfac65dbb2607fd80e3065fd77ab7

Observation 59912438-66d6-4ff5-bb62-79e632c62f3d · outbound

This paper cites TextDiffuser-2: Unleashing the Power of Language Models for Text Rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation TextDiffuser-2: Unleashing the Power of Language Models for Text Rendering

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.731711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.731711Z digest=sha256:84970d37b3cc9bf352658fe0b76750c9461afd06952a36f2cf22d24699c7f8ca

Observation 37680e1e-3a6c-4d2b-b740-7fd746b94214 · outbound

This paper cites TextDiffuser: Diffusion Models as Text Painters.

Preliminary Explorations with GPT-4o(mni) Native Image Generation TextDiffuser: Diffusion Models as Text Painters

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.736690Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.736690Z digest=sha256:f8ed753e04e0ab346bd35720eec4cd1913133892d3b20f55ea401e82f0e73d52

Observation d264e7f9-0cd0-4d58-bffa-56c9768eb36f · outbound

This paper cites Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Semantic Image Segmentation with Deep Convolutional Nets and Fully Connected CRFs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.741344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.741344Z digest=sha256:89e603739129cae6550fa94895793c16e5c225996a9ce7a9edfd3a76f4de7b8a

Observation 144adc4e-f03e-4738-885a-6ecade1076d2 · outbound

This paper cites An Empirical Study of GPT-4o Image Generation Capabilities.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Empirical Study of GPT-4o Image Generation Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.746041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.746041Z digest=sha256:5e2a442a939e1e55d333e36dc1225f5e9bf674f43111a3f88baeca2ef5a0fb4b

Observation ad71bfe2-44f0-49f3-8971-e9aa8831e946 · outbound

This paper cites Manga Generation via Layout-controllable Diffusion.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Manga Generation via Layout-controllable Diffusion

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.750697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.750697Z digest=sha256:79ceb02040df88ad23b387a7c9557c4961208a344cb569136e1a3636c6f303f5

Observation 1bf12b18-de75-432b-b3ff-dac921afb197 · outbound

This paper cites All snow removed: Single image desnowing algorithm using hierarchical dual-tree complex wavelet representation and contradict channel loss.

Preliminary Explorations with GPT-4o(mni) Native Image Generation All snow removed: Single image desnowing algorithm using hierarchical dual-tree complex wavelet representation and contradict channel loss

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.755132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.755132Z digest=sha256:0db8b47b77b818fadfa15cdffcb88ba2a76a8e13d58c65f03a769523d1e208bb

Observation 0e070bc9-c686-4a74-990b-ba419fdfd1d0 · outbound

This paper cites DreamIdentity: Improved Editability for Efficient Face-identity Preserved Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DreamIdentity: Improved Editability for Efficient Face-identity Preserved Image Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.759651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.759651Z digest=sha256:22bc34b1dc810b8da36efc5a0be3da94a5defd3aaac16f6bb6b93b140e94718e

Observation 771d326f-3b91-491c-8196-7d2d99db9355 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.764185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.764185Z digest=sha256:3cdcde4ec4dab6c170b92f6ea4457acf98f6820a2c4d0300d5bb9eb99d6bcb5b

Observation c0a62679-c050-4b44-aeb2-565d6e1f5ac8 · outbound

This paper cites Snow mask guided adaptive residual network for image snow removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Snow mask guided adaptive residual network for image snow removal

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.768725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.768725Z digest=sha256:8cf79b3b653918b32538532f411d373d82cda6457a3361eea94228b1ed1a4344

Observation 81906c70-4c57-4731-9167-1584c2f88e7c · outbound

This paper cites Masked-attention mask transformer for universal image segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Masked-attention mask transformer for universal image segmentation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.773805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.773805Z digest=sha256:626bd168bb91606ce17e2ac9f08655c502be9f3b7c2c69cd0c42367a28f58e04

Observation fb144973-a305-4ff1-9e7b-53ed587bd2a9 · outbound

This paper cites Per-pixel classification is not all you need for semantic segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Per-pixel classification is not all you need for semantic segmentation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.778216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.778216Z digest=sha256:2d07829a45450c4716eb9d9580238d45110c9452c214cd449ab8a7d7076c83e7

Observation aa57dda6-cfb8-4972-822e-7452d1676e04 · outbound

This paper cites Object counting and instance segmentation with image-level supervision.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Object counting and instance segmentation with image-level supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.782665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.782665Z digest=sha256:f7c807728af5c6fb01e20746ffb8aef9befad94ef15ac49b28cb19b82864d7f7

Observation efb880de-d4ed-4f27-a38b-dd511d5a384e · outbound

This paper cites Generating diverse agricultural data for vision-based farming applications.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generating diverse agricultural data for vision-based farming applications

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.787222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.787222Z digest=sha256:79d1580e86eab4cde41ee6d6de4a3a95108745bd465046b07f45e99fbc39b79a

Observation 842a30dc-5498-40c7-a371-30b5d1dfa2e1 · outbound

This paper cites The cityscapes dataset for semantic urban scene understanding.

Preliminary Explorations with GPT-4o(mni) Native Image Generation The cityscapes dataset for semantic urban scene understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.791409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.791409Z digest=sha256:2ef7f43f957e0f046ce97b67f84cc07cc768f8222b2ecdefd2bd533c7cb85d2b

Observation 9dc81256-0f63-4a7c-974e-cb0e54898864 · outbound

This paper cites Latentpaint: Image inpainting in latent space with diffusion models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Latentpaint: Image inpainting in latent space with diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.795343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.795343Z digest=sha256:b471497c9492d9e970c784196c046850d8a2e195029706060bdd8c425c4f1e0d

Observation 35d55de4-b92f-443f-b11b-34312494d835 · outbound

This paper cites Deep learning based 2d human pose estimation: A survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Deep learning based 2d human pose estimation: A survey

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.799583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.799583Z digest=sha256:100bbdba38d0226bb5c8825246d476467f2b37eeec70131bad317377dc70136e

Observation c94acfc6-3603-4f26-b205-f1706bac0966 · outbound

This paper cites 3d-aware conditional image synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 3d-aware conditional image synthesis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.803714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.803714Z digest=sha256:d42b5944da79bb856b94ee2fa73ea5374bae9187d6806b74219fd5cf8ca13579

Observation f1049d54-fde6-4363-9609-7d23088a7cec · outbound

This paper cites Towards intelligent design: A self-driven framework for collocated clothing synthesis leveraging fashion styles and textures.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Towards intelligent design: A self-driven framework for collocated clothing synthesis leveraging fashion styles and textures

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.808389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.808389Z digest=sha256:fb720a9ba53bbfce1dcc3d2c60fe0c9f874bc2240a0963cf7e36464fc7bb0b53

Observation b9e021c8-a6b4-42b7-9377-b0e7818556dd · outbound

This paper cites DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DreamArtist++: Controllable One-Shot Text-to-Image Generation via Positive-Negative Adapter

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.812799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.812799Z digest=sha256:1c7ca0e0d9366c41721a5f5bc83cf51d6367e901aad90d9209445c6bc16eba73

Observation cc8644ff-a1cf-45ff-b6d1-c9264d3d3e35 · outbound

This paper cites Discovering novel biological traits from images using phylogeny-guided neural networks.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Discovering novel biological traits from images using phylogeny-guided neural networks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.817421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.817421Z digest=sha256:c5b9f4e943b7377951a783295f1f4fc4bf9f91070afea92b5f159acaaec22d63

Observation 3b902b97-9c4c-4c7b-9c9c-1dcb46892af8 · outbound

This paper cites Everingham, L.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Everingham, L

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.821802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.821802Z digest=sha256:02a5228cc2c712a7239ddd652f52e224bff44a3fc67b25a129d1bb98db6e3d92

Observation 795e3154-f504-4469-86ce-559b3197ed52 · outbound

This paper cites Everingham, L.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Everingham, L

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.825899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.825899Z digest=sha256:84851c67b8143f0a69211920fd8a29da037c467a402f4b83f7d5df4a7f243319

Observation cfea3f1b-a514-4cd3-814a-cfaa7b5a2dd8 · outbound

This paper cites Guiding Instruction-based Image Editing via Multimodal Large Language Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Guiding Instruction-based Image Editing via Multimodal Large Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.830158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.830158Z digest=sha256:6404c43ea4076f108dab0ec1cf62e8217a8fb043b211b9eafdd76ca28b367673

Observation 3579c4b0-bea9-41da-a16e-4b22a62178bc · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.834485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.834485Z digest=sha256:e4887a318ea4bd3bccb9a55d7271886bc1a1acb8f214f86233fa49fd0e94e6f3

Observation 7d6494be-079c-44b0-af90-c7f83bcdd069 · outbound

This paper cites CascadedGaze: Efficiency in Global Context Extraction for Image Restoration.

Preliminary Explorations with GPT-4o(mni) Native Image Generation CascadedGaze: Efficiency in Global Context Extraction for Image Restoration

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.838880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.838880Z digest=sha256:e21212863a33d51ca25a5d17fa7596b6c988ebebe101939c6320ccccc91cef9b

Observation 038a5edf-4984-4878-aae7-9146e0f93afe · outbound

This paper cites Fast r-cnn.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Fast r-cnn

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.843410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.843410Z digest=sha256:f2444b3dd6c727f1b5ec3665d3c4d8c34691a29d4a3966f29dd5353b7d74659d

Observation 44601787-8338-494b-bfa9-f15a2d3eff96 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Rich feature hierarchies for accurate object detection and semantic segmentation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.847545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.847545Z digest=sha256:f4751d573234430784629f61ae1109066e9a479418b489f4c4a5962a00c3dc19

Observation f1bfd7b5-69d9-495f-abe3-3c61337ed373 · outbound

This paper cites Talecrafter: Interactive story visualization with multiple characters, 2023.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Talecrafter: Interactive story visualization with multiple characters, 2023

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.851631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.851631Z digest=sha256:d3f075b4ffe5ccef959f9926396f3ac445d66fea82cf3e1d5f0af51a7142353e

Observation 27b339b5-6703-467f-8f51-62bbc13b7210 · outbound

This paper cites DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DART: Denoising Autoregressive Transformer for Scalable Text-to-Image Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.855641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.855641Z digest=sha256:6009fe8a1354cc74e6cc8ed3e8a889737d5773a2a277b3787124b1ecc15ae6ae

Observation 719a037f-9686-4e88-a9f5-1b7e6a4fe024 · outbound

This paper cites Modulating Pretrained Diffusion Models for Multimodal Image Synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Modulating Pretrained Diffusion Models for Multimodal Image Synthesis

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.628774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.860484Z digest=sha256:ad14fd66dc03ab8bde6333302882cf510cd97611918b67ea3e30eaee75bdd524

Observation da1b9c2a-45b8-49c6-920b-b19e59bd6b0c · outbound

This paper cites Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Infinity: Scaling Bitwise AutoRegressive Modeling for High-Resolution Image Synthesis

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.864916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.864916Z digest=sha256:af287df834a22a9504db5d3f2a8b3976d6961ea01e081660c35a4d7a070544ed

Observation 530f8f8b-2a42-4e3a-af28-5c699994804b · outbound

This paper cites an unresolved cited work.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.869304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.869304Z digest=sha256:c9dc7cf502339e33834152e8fdaf7b0cf4477b3b82179d40a14151582bd67612

Observation c14053b7-d2c7-4b55-8457-9cce2b8d67ad · outbound

This paper cites an unresolved cited work.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.873445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.873445Z digest=sha256:caed26de73802308d0f0240de5de6eedcc2b8f14e028a3f72d2d68106df217ca

Observation 1785a553-6d7c-464b-85fc-eb57d11e5020 · outbound

This paper cites Dreamstory: Open-domain story visualization by llm-guided multi-subject consistent diffusion, 2025.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dreamstory: Open-domain story visualization by llm-guided multi-subject consistent diffusion, 2025

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.877554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.877554Z digest=sha256:9ae35867ac9535e95f6866c979f1078ae758471302241be49547d4e77cdaa05a

Observation 37d306bd-4176-4657-add1-b92b5857256c · outbound

This paper cites Styleposegan: Pose-consistent virtual try-on via pose-guided style transfer.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Styleposegan: Pose-consistent virtual try-on via pose-guided style transfer

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.881449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.881449Z digest=sha256:a6ce32daacd4ce84f95b90bc9d40f871f5c800a67c0a5440f149b32777722721

Observation 27fc4bf7-d9d2-4bfb-b952-2f079060611c · outbound

This paper cites Dresscode: Au- toregressively sewing and generating garments from text guidance.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dresscode: Au- toregressively sewing and generating garments from text guidance

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.885537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.885537Z digest=sha256:1eec97678331bbb77a6d94abe1fac6b54442ba198e0002e9fec97bde8e167460

Observation 3b483f99-2dc4-4887-bb4f-0d86f3b9a311 · outbound

This paper cites SynthSet: Generative Diffusion Model for Semantic Segmentation in Precision Agriculture.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SynthSet: Generative Diffusion Model for Semantic Segmentation in Precision Agriculture

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.889670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.889670Z digest=sha256:131597714c28a7003a892eca47269a900d208b5bee3d4f123d681ad01002dc44

Observation c3b79798-25d7-44b9-b6a9-eb4f08b76709 · outbound

This paper cites Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.893977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.893977Z digest=sha256:c934500f5047f023742d7932740c641bd87f9dc09f0977a33dcf5feedf28ff49

Observation 9cb529da-72ca-48b1-8cae-28ce021f1691 · outbound

This paper cites Composer: Creative and Controllable Image Synthesis with Composable Conditions.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Composer: Creative and Controllable Image Synthesis with Composable Conditions

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.898202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.898202Z digest=sha256:c393c26a2b3315f5dc8d69c40ff7048e9fc347ff2222f2f4af056694bc30d599

Observation f996091a-1cd1-40ce-96fb-96f203d1ad48 · outbound

This paper cites Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Learning Disentangled Identifiers for Action-Customized Text-to-Image Generation

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.555542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.902399Z digest=sha256:e96e028d81c8bf9ea1decc529ba079d18ca42e2088b0bae003ce47cc5c37c9fe

Observation 0659121a-ae89-4980-a33c-7b0621a351bb · outbound

This paper cites AutoGeo: Automating Geometric Image Dataset Creation for Enhanced Geometry Understanding.

Preliminary Explorations with GPT-4o(mni) Native Image Generation AutoGeo: Automating Geometric Image Dataset Creation for Enhanced Geometry Understanding

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.906527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.906527Z digest=sha256:4649c09013cbd4a85332e657b39a5aa70c64be7c2e97168b14db5a2aef748b27

Observation da0625f8-2d9a-43a4-bd93-3b8c12fade59 · outbound

This paper cites ReVersion: Diffusion-Based Relation Inversion from Images.

Preliminary Explorations with GPT-4o(mni) Native Image Generation ReVersion: Diffusion-Based Relation Inversion from Images

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.910917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.910917Z digest=sha256:5182e047ebd276ee9eafdc1a13fe5ce82b79d98ec206c7324e19ffd36c220e29

Observation 8f837654-7f41-417b-a035-eea6134b0574 · outbound

This paper cites Liteflownet: A lightweight convolutional neural network for optical flow estimation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Liteflownet: A lightweight convolutional neural network for optical flow estimation

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.916052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.916052Z digest=sha256:f5c85eacc668717a0e34ce30ee0b770833cff85c6013ec1292a4c9ec2e9561b7

Observation f9e57c1f-b25e-47a8-be60-48d7052a8e4c · outbound

This paper cites Vision transformer in industrial visual inspection.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Vision transformer in industrial visual inspection

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.920065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.920065Z digest=sha256:f6d52308170669fda15a4b4efe8dd945e105671186d291cf8e427ce454f20a07

Observation 63a02054-893f-420e-a2da-d9a46e4bf82a · outbound

This paper cites Flownet 2.0: Evolution of optical flow estimation with deep networks.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Flownet 2.0: Evolution of optical flow estimation with deep networks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.924276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.924276Z digest=sha256:6d15825aa4ad4078d4de741fdcebaae6a61d1a3841dc34ff658747e379b2a463

Observation 5c9a6014-9305-4151-b7aa-9856591d1143 · outbound

This paper cites Desnowgan: An efficient single image snow removal framework using cross-resolution lateral connection and gans.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Desnowgan: An efficient single image snow removal framework using cross-resolution lateral connection and gans

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.928240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.928240Z digest=sha256:f8a8487cf2f5c9b90e6ce9257655d7ee02921f1bd60ae52b61bc9d532be777b4

Observation bf5fb43c-4d2b-4f47-9e21-a10750de7164 · outbound

This paper cites Remote sensing change detection in urban environments.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Remote sensing change detection in urban environments

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.932328Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.932328Z digest=sha256:563b52bd216098389f3127d359060771af589e11d5a53f8f4c6eee3e92123cbe

Observation cb8dcaf4-c535-472d-b974-0afacab16b0b · outbound

This paper cites Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Culture-TRIP: Culturally-Aware Text-to-Image Generation with Iterative Prompt Refinement

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.936502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.936502Z digest=sha256:556c53d9a68c2ff0c06a4ab8c66c8bc2899f614765b942ef79be632ed2bdca60

Observation 66f47646-64f3-4701-8919-89ac17e4823f · outbound

This paper cites SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation SSMG: Spatial-Semantic Map Guided Diffusion Model for Free-form Layout-to-Image Generation

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.494729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.940679Z digest=sha256:dc33fe451a8de804cecd4773df270cacb7538c6f3d197d60f9f4e562a042559f

Observation c98d2440-8761-4f94-aacc-064690ab2357 · outbound

This paper cites Lumen: Unleashing Versatile Vision-Centric Capabilities of Large Multimodal Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Lumen: Unleashing Versatile Vision-Centric Capabilities of Large Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.945352Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.945352Z digest=sha256:b76ddbe8a65d15cb085a234fcf0d79df6529d388eaf44263e15486d0d44c4ed8

Observation 7a29afa5-3cec-408f-99d2-d47299c4a2f5 · outbound

This paper cites Beyond Aesthetics: Cultural Competence in Text-to-Image Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Beyond Aesthetics: Cultural Competence in Text-to-Image Models

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.949916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.949916Z digest=sha256:c0f71db10238755906fe4c3422671ce521755c15f93a9209b12f37f1b67c49a6

Observation 557623ef-19d9-4a24-aaf0-6202808f6732 · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation 3d gaussian splatting for real-time radiance field rendering

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.954334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.954334Z digest=sha256:882e9efa36f7e19bf097bf689001c18292cfdca762c75c9e168daab925cc8430

Observation 6a5b976d-9aed-453c-b850-aa770c613df1 · outbound

This paper cites DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models.

Preliminary Explorations with GPT-4o(mni) Native Image Generation DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.958763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.958763Z digest=sha256:55960bc403b87920f912aba0fe0d431e0bb105cf666db750d4590d9c0f939e14

Observation e722c22a-6969-44e7-8556-d2540879507f · outbound

This paper cites Probabilistic modeling for human mesh recovery.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Probabilistic modeling for human mesh recovery

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.963353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.963353Z digest=sha256:f855536b4adda668a620cdb95c5e11e891bc129427ea5cb03702e3f0b74e0368

Observation 2a3ae3de-47d3-4b5e-850a-3139a0907cae · outbound

This paper cites Raindrop-removal image translation using target-mask network with attention module.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Raindrop-removal image translation using target-mask network with attention module

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.967398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.967398Z digest=sha256:a6be582d12e89157eb5ed99521b9ed1515925884b09513039ec650005f5ddf10

Observation bc586949-05d3-4f58-a8c3-67e684898143 · outbound

This paper cites Dicti: Diffusion-based clothing designer via text-guided input.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Dicti: Diffusion-based clothing designer via text-guided input

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.971373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.971373Z digest=sha256:7abeb3e41812f54e12f37f5d254b0008f50a14a603d136aa23e3059cf12a119e

Observation 5d7bd6e6-486c-4f55-ae82-b522f0b294ec · outbound

This paper cites Physics-based shadow image decomposition for shadow removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Physics-based shadow image decomposition for shadow removal

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.975366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.975366Z digest=sha256:40d1933de1ef7e45f1a75893d5b41a564c478be4d1f76017bd40bd189682d9f0

Observation 60341660-f287-4588-858c-a35a363177c4 · outbound

This paper cites From Text to Visuals: Using LLMs to Generate Math Diagrams with Vector Graphics.

Preliminary Explorations with GPT-4o(mni) Native Image Generation From Text to Visuals: Using LLMs to Generate Math Diagrams with Vector Graphics

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-15T23:46:04.434842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T23:46:02.979690Z digest=sha256:7afe83a91acde042b4ff213aafbd4e06722c597b80cab7711cf33e2b9c82610c

Observation 9d49abaa-04a5-4f99-adba-a64617eb183d · outbound

This paper cites An underwater image enhancement benchmark dataset and beyond.

Preliminary Explorations with GPT-4o(mni) Native Image Generation An underwater image enhancement benchmark dataset and beyond

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.984144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.984144Z digest=sha256:74e4ee37818f07890daf5531a167b9677603e2056a4e9d5c34d926a9259eaace

Observation 6c8fa3de-94c1-46cb-ae53-a46c538b7287 · outbound

This paper cites Real-world deep local motion deblurring.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Real-world deep local motion deblurring

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.988400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.988400Z digest=sha256:6cbdea590c459bf99dea28ee5c710d88f28e3762859c43a6013df514cc7f6875

Observation f4738e64-9a59-44a8-b5a6-8aa89ae0307a · outbound

This paper cites Cheffusion: Multimodal foundation model integrating recipe and food image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Cheffusion: Multimodal foundation model integrating recipe and food image generation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.992330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.992330Z digest=sha256:f3e2620deb34660cf1e35d758ce5276c1bb9f8332b88e7d666547aaf2c0da13a

Observation 3c9fd2a9-a618-4930-9f8a-7e0113366c74 · outbound

This paper cites Image content generation with causal reasoning.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Image content generation with causal reasoning

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:02.996386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:02.996386Z digest=sha256:129fc3f25277e06b685ad7c0b220c5fddd736afab31cd9f53bcb4e9593e672e7

Observation e0b3a3b6-8b60-4731-979f-4eb3fca83de0 · outbound

This paper cites Exploiting reflection change for automatic reflection removal.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Exploiting reflection change for automatic reflection removal

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.000457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.000457Z digest=sha256:f6d95f2b536c6d972ce2dc7e90f524306113741ec07bde03f9c0552696096d4a

Observation ff29fc25-f131-4327-8a2c-6f0ede33dd9d · outbound

This paper cites Generate Anything Anywhere in Any Scene.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generate Anything Anywhere in Any Scene

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.004601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.004601Z digest=sha256:8b80f44bfd31ecb7e3ee178ff67c71523afe59d038c1a877e9cffbe77324ac0d

Observation 309c2a77-bd40-462b-9d10-8de4fcfb5d38 · outbound

This paper cites Gligen: Open-set grounded text-to-image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Gligen: Open-set grounded text-to-image generation

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.009540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.009540Z digest=sha256:d2895cc45df048e304e2a9dbfc58d7213e2fb83deac4d81b82bed6d39f043ab9

Observation 17d7d127-b48e-41a2-818d-979823c3d95e · outbound

This paper cites Swinir: Image restoration using swin transformer.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Swinir: Image restoration using swin transformer

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.013949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.013949Z digest=sha256:2b618a34d2b4ca68820368379543d01e2e7f9fe219e06b7fcba569c82fc3f511

Observation 3dcaaf6c-f497-4a4b-b579-c84e6b9b14f9 · outbound

This paper cites Object Counting: You Only Need to Look at One.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Object Counting: You Only Need to Look at One

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.017973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.017973Z digest=sha256:b7a27f0d0bc8d12cc302f55018b12bea48cb90e6e59b5572d4929af271f8d4be

Observation ae60a7ab-10be-4e17-ae25-0d380b23dc62 · outbound

This paper cites Phys4dgen: A physics- driven framework for controllable and efficient 4d content generation from a single image.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Phys4dgen: A physics- driven framework for controllable and efficient 4d content generation from a single image

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.022186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.022186Z digest=sha256:d13c0345059490ef442128e9cbdfeb6e20a8afb7c8848399b342cb7696d4c385

Observation be5619c1-a47e-42dd-b6c4-f69a150adc2a · outbound

This paper cites Microsoft coco: Common objects in context.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Microsoft coco: Common objects in context

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.026366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.026366Z digest=sha256:18b09f8c31904a76fec8c99800bc19561e1341096dd88b00bceb1bec49beae55

Observation 630eb4f5-70be-4fbf-9da6-1b9f6b190442 · outbound

This paper cites On the cultural gap in text-to-image generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation On the cultural gap in text-to-image generation

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.030724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.030724Z digest=sha256:abab6054a25f1a0ff5fb45abc27338df672ceb857b0255a7cdd260cd372f2e61

Observation dfce535c-f1ea-494f-9e9e-cffa8163f5d0 · outbound

This paper cites Generative Physical AI in Vision: A Survey.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Generative Physical AI in Vision: A Survey

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.035004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.035004Z digest=sha256:c924df0060fbe2e053b381523d1745c0111f83ec69003ec2353c1b0d661cee17

Observation e2282a71-9935-46f7-a685-e1547c4a58b8 · outbound

This paper cites StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter.

Preliminary Explorations with GPT-4o(mni) Native Image Generation StyleCrafter: Enhancing Stylized Text-to-Video Generation with Style Adapter

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.039629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.039629Z digest=sha256:b2061113fc794eac8ffc19e403f059e19e86842b7e6c6f4748154948329ecf3c

Observation b3e7ee1f-17d1-4baf-a65f-e0c854a84b01 · outbound

This paper cites Structure matters: Tackling the semantic discrepancy in diffusion models for image inpainting.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Structure matters: Tackling the semantic discrepancy in diffusion models for image inpainting

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.044224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.044224Z digest=sha256:8b88a376955c53ab2b5ef9f12a537a52f115fa1417e0119ef56db8a3a13cf56c

Observation 54215261-fb6c-42ec-a503-37627b0ccd0b · outbound

This paper cites One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization.

Preliminary Explorations with GPT-4o(mni) Native Image Generation One-2-3-45: Any single image to 3d mesh in 45 seconds without per-shape optimization

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.048723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.048723Z digest=sha256:7f5d8550f0ebd48a467d852ba8337b69da0daedc2f27f8a484d5ca65d9e6b570

Observation c419dc49-cd53-4dba-b0cd-23f047d9b483 · outbound

This paper cites Git-mol: A multi-modal large language model for molecular science with graph, image, and text.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Git-mol: A multi-modal large language model for molecular science with graph, image, and text

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.052810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.052810Z digest=sha256:2f517db08679376b068f1afdd48fbc889a44c920aa23cd8aaa9c728652ee05c8

Observation be9f98fa-9853-4479-ba25-35e94013e4ce · outbound

This paper cites Character-Aware Models Improve Visual Text Rendering.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Character-Aware Models Improve Visual Text Rendering

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.056819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.056819Z digest=sha256:4d2d2ccb37b1b3b17bb65a6c27790c2dbd9aed0ab7040a0ed15adcf825034a52

Observation 6a51503c-58dc-48ce-a7c9-8abaa2446b08 · outbound

This paper cites Zero-1-to-3: Zero-shot one image to 3d object.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Zero-1-to-3: Zero-shot one image to 3d object

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.061601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.061601Z digest=sha256:3ae13f2e7cbda80fc204806d505725285ecfff81e628e8fb5e14636793e04c81

Observation eb0c1020-6dcf-4345-bf62-bd5e33e9411e · outbound

This paper cites Physgen: Rigid-body physics-grounded image-to-video generation.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Physgen: Rigid-body physics-grounded image-to-video generation

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.065904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.065904Z digest=sha256:a6b1b095cc6840af71dbc464dd8c2ba51037790127c17db88ca57d7bb59a4b84

Observation b08867eb-ef58-45ba-a0d0-4e02857f91d1 · outbound

This paper cites Ssd: Single shot multibox detector.

Preliminary Explorations with GPT-4o(mni) Native Image Generation Ssd: Single shot multibox detector

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-15T23:46:03.070172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:46:03.070172Z digest=sha256:9a656ef26d19976b0fe2216789998091acbb7fd03ad3c5fa857111a0f23ffdde

Pith citing papers

Observation f09b4bd0-cc60-4329-b3ee-9ccb02314bf8 · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:10:59.092148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T17:46:59.581486Z digest=sha256:87d716962a3d61793b52c56539e8214f382c23327e96a2db2bcac8885678bab6

Observation 2727aeba-fb0f-4b09-af38-c50b6465e54c · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:34:57.141773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-21T09:34:16.292323Z digest=sha256:092e3a4bd0b1c4d018bd0fd430365ffc8de979887ed674a6d6b31be39c3eeb17

Observation 92c38713-8390-44c4-b780-4939cd4b66a9 · inbound

PosterHarness: Turning Scientific Poster Generation into an Auditable Instruction-Following Benchmark cites this paper.

PosterHarness: Turning Scientific Poster Generation into an Auditable Instruction-Following Benchmark Preliminary Explorations with GPT-4o(mni) Native Image Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T05:30:32.163273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:30:32.163273Z digest=sha256:b8f5c659788ab2b361a79e14d120b00eec0161ab0e70860df8fef3a7b0a212d2