Pith. sign in

Paper Citation Record · LEDGER

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

As of 7 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 7 inbound Pith citation observations for arXiv:2605.11061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11061 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T07:30:53.939221Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:46:23.336119Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T16:59:58.072695Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact3
  • verified fuzzy39
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch25

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1fb42cc-2f91-4e74-b6e1-bc990d373d7b · outbound

This paper cites Qwen3-VL Technical Report.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Qwen3-VL Technical Report

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.671231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:3e057b1d4d664e6db85e803a057a3450bf17dd1ed8072ed5170455bbb20838bb

Observation a268411b-4ec7-45a1-a5c3-dc428dfecbc5 · outbound

This paper cites HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T17:13:39.137651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:5a8b395deb3a895ed45bb37af12e9c638da5c2113388c5a08dba305ab042f4b4

Observation 27a4dcd7-9984-4154-8190-7705781966c1 · outbound

This paper cites IEEE Transactions on Image Processing (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer IEEE Transactions on Image Processing (2024)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.510490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:9f09c62c765b61715a8823fea955f1ca1969576710a8454328dd10390a8cac80

Observation 5c6bd396-b527-4ba4-af0e-c97c61658af9 · outbound

This paper cites In: ECCV (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ECCV (2024)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.514587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:c76622bb8f324e73e6bf8e113d814deb6ecb4399a05fea822a047c97b5c153b1

Observation 7f266e11-021d-4f5e-856f-0dacaccfa619 · outbound

This paper cites BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.665378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:c90a1c4fb4af202d5a220d83a36863cf7c332fbc65f3ffaba5a97ca21c63c27a

Observation cc9faee1-1340-44ad-9a61-c304c3207aa8 · outbound

This paper cites In: ICLR (2024) 22.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICLR (2024) 22

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.518594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:2a9bbbd2d4b473b72200c8f6e17323d8070346321c8ef5b997eb84b0a2d587f5

Observation c58e6441-eafc-4645-a827-ce671e78440d · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.659035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:014d85663afada0844b7ada9936decb434f667e3b62477bb12bd543e4806b12d

Observation 9452ff85-47dc-466c-b7de-e940990143e7 · outbound

This paper cites Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Region-Aware Text-to-Image Generation via Hard Binding and Soft Refinement

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.676226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:ef17516403d89e449f84967938d6bb4ec0a52a2311f88e15dfbb8bf0ed0dcc0d

Observation dca98682-6c96-4ed9-af04-0b76425374d2 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Emerging Properties in Unified Multimodal Pretraining

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.647389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:9b549bf20f3dad00fe68af311aa355cbabd6e6e4d0438e16a69a9d37b8886f9d

Observation 5ffcc56b-5c24-4464-9a6f-82f8ccbcc1da · outbound

This paper cites The Faiss library.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer The Faiss library

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.652839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:19013fbd0549a0a1adca3dc2cac79061aca62ff076844841b956bed917e4c1ff

Observation a7d83bad-d66e-4afb-a0dd-a33196e45e7e · outbound

This paper cites Textcrafter: Accurately rendering multiple texts in complex visual scenes.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Textcrafter: Accurately rendering multiple texts in complex visual scenes

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:32:29.544298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:2f4d4dd638da1cbf158591022201c2db004d9738fe9803e296a9a7bc28bb7a0e

Observation 43bbd542-371f-4f18-8a75-f97f8ff72d41 · outbound

This paper cites In: ICML (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICML (2024)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.465507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:b0e52ae9e75c37a15aeb4af633052dd69a46b212c3d10a9d5fc49ce35fc62e0c

Observation a3c5b42c-02a1-4d9c-b151-b7003dcd0d6c · outbound

This paper cites X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer X-Omni: Reinforcement Learning Makes Discrete Autoregressive Image Generative Models Great Again

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:32:29.598535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:00fab30c2043f3d96e6742f8278bd2e132d7dbc53d60c047f72796569989a42d

Observation a256db51-9c80-4cfa-b861-f4031ce526f3 · outbound

This paper cites In: NeurIPS (2023).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2023)

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.489482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:9302c2218c8094ab1b8680fc88bc868f7db627090085b208dba9609cf9ff539a

Observation 27d25ff8-dede-4758-b9c9-f325ee42c3a5 · outbound

This paper cites https://gemini.google/overview/image-generation/ (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://gemini.google/overview/image-generation/ (2025)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.469584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:937b6821c63b3afa5f1b877e8a579f44e05b5a224359a8e1f5987f9d7b1c2e03

Observation adb59fe0-7f20-4585-b94e-df0571979689 · outbound

This paper cites https: //blog.google/innovation-and-ai/technology/developers-tools/gemma-4/ (April 2026).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https: //blog.google/innovation-and-ai/technology/developers-tools/gemma-4/ (April 2026)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.499884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:246742e2828af1fa04e48e0ede09517ce2d04b83266a89b86cbd168505aa62ed

Observation 74446acd-b424-469b-9c49-e6452fec12b8 · outbound

This paper cites In: NeurIPS (2020).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2020)

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.453213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:1070d7dcbd598995eaff38540e26145f852f25cfe935c668c2741f8b45a7a20c

Observation b24eb85d-2d0f-4d81-96b6-7a4e04ef2012 · outbound

This paper cites In: CVPR (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2025)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.356506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:9dd75241792378be9f86522941eef07004f22920386961067885d25927fb37bb

Observation bce87889-9f27-4c14-a2a9-efd15e3e9397 · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.579242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:b3700cd29eef7cb8cc2d6b92a2bb7bdbc72bbdf445046d4f51e86e73d0e4ddab

Observation 9154e600-e148-43f2-9d23-fb2a4be881b2 · outbound

This paper cites In: ICLR (2014).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICLR (2014)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.481888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:061e9928e529ce48fee1f217da55cb22cb9a8dae7fa2e3913c31cc569876f24d

Observation 23cc2dc6-1f3e-4c7c-90aa-91c104fec71b · outbound

This paper cites In: ACL (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ACL (2024)

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.461793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:cab779749f683fef1f8b35e015cdff288bade0aac24d51f5e665e4919974c4a6

Observation 93c9a460-489d-491a-82ad-a9d156bd91ca · outbound

This paper cites https://huggingface.co/black-forest-labs/FLUX .1-dev(2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://huggingface.co/black-forest-labs/FLUX .1-dev(2024)

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.448683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:37968a5b3bfe6b74f10b512e5772dcabd35046523003e3117f773ea186f1ed2d

Observation 60f97e5b-9aba-4527-983d-2f621cb8814b · outbound

This paper cites an unresolved cited work.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-05-13T08:17:32.492703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:df46fc83a7bc5f8823ed665c08005f69d4ed0c89419caa8177a5fb056a96bb88

Observation 9820ca73-c620-4e73-a237-5557071a865d · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:32:29.621485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:a445271b9ff588c47481d863f665e43a2090eb4076df01359006f0396d58aee0

Observation 468fb069-db85-467c-9a7e-28c94ac8f3f8 · outbound

This paper cites https://github.com/LAION-AI/aesthetic-predi ctor(2024), gitHub repository.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://github.com/LAION-AI/aesthetic-predi ctor(2024), gitHub repository

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.496287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:cf1db8a649075dfbfd27e9654a68607d18b15d36e09ce30f1110f2c9e856aa01

Observation 6a67e039-9482-45db-9104-8d9f55e2104b · outbound

This paper cites https://github.com/LAION-AI/CLIP-based -NSFW-Detector(2024), gitHub repository.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://github.com/LAION-AI/CLIP-based -NSFW-Detector(2024), gitHub repository

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.382380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:ccf6205484b8976f18357532094edfab4ce29aa67d846b04a31c2cfc9de6dbf7

Observation 9c8b9787-480d-48de-baf2-96dbfbd966f5 · outbound

This paper cites https://github.com/LAION-AI/LAION-5 B-WatermarkDetection(2024), gitHub repository.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://github.com/LAION-AI/LAION-5 B-WatermarkDetection(2024), gitHub repository

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.360618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:56772b22c6c1dc66c76891521792bc9bbb1570bbbc8dd0cc3cb2b81ad2e9f0c8

Observation bcf1b7a8-d449-43f1-8056-77ceb73bb902 · outbound

This paper cites Back to Basics: Let Denoising Generative Models Denoise.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Back to Basics: Let Denoising Generative Models Denoise

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.591394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:a4e688eece4649f977e07c7fc2e0339f62b6a8f642a9f4b0c543b24f6bbc9442

Observation 9d5a4e95-03ff-4132-8a32-06f8e89223cb · outbound

This paper cites Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Hunyuan-DiT: A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T14:58:37.566044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:747e3b639145d69b979c64c61d59489d8191d2aa622bc3b394424b55f7dbc044

Observation ae1ef9dc-505b-491f-ad74-21551f2d61eb · outbound

This paper cites In: NeurIPS (2026).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2026)

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.370076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:545b782c80d6d4c0ca096cd8ced72c7ca3a6779ea34f6a364e862af278697c16

Observation 3fb64024-787f-4e87-8272-6e948da6cd59 · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Step1X-Edit: A Practical Framework for General Image Editing

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.531987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:b093eaab8bd5d3801ff7721d1d688bbef67c8179a8c1a40266c32fec254b1761

Observation 99f182e1-2c93-46a7-94c8-bb316d6f7dd8 · outbound

This paper cites In: CVPR (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2025)

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.377758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:235c80e1db123c3c3027dbbf8ee182da5b2ec381967ff850ddbc76d96af4c436

Observation c5be07b6-eeef-4801-88ca-84d3256d1b2b · outbound

This paper cites In: ICCV (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICCV (2025)

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.444719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:eb9c4e76f439341f587dc983fb076e1f92c9c9bd98e8e11c4476fa1b6b6f04f6

Observation bb919cf6-f5db-4efa-bdb3-477d3ee66a54 · outbound

This paper cites arXiv preprint arXiv:2508.15772 (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer arXiv preprint arXiv:2508.15772 (2025)

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:29.560958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:7100ac0e84453284bb61822ab5345c3c604f61dc7b7365f12c9ef68868ba00bb

Observation 35b65d5a-2178-4b0a-8bd8-5c3341e6ffc2 · outbound

This paper cites https://openai.com/research/dall-e-3 (Sep 2023).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://openai.com/research/dall-e-3 (Sep 2023)

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.405444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:5cc4053e0c5cbe9a99a0ca2da64f83580b91ae702af4f3c390fb56d3437a2d2b

Observation 8165d0c7-f382-4cd4-8351-a2c831575824 · outbound

This paper cites https://openai.com/index/introducing-4o-image-gen eration(2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://openai.com/index/introducing-4o-image-gen eration(2025)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.374055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:21ad19f667d4165f0d5d99a560460a69fa2fff52db05f9db0a2d2c5639dfb158

Observation 016da473-5e59-4b05-bdfe-2952d6c7b0b8 · outbound

This paper cites In: ICCV (2023).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICCV (2023)

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.364922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:7c85714250714343868c31accfeff389d0007e7c7540a57bf10d22394a7664bc

Observation f0e266e1-cf2d-40de-b6ab-01bc9d4892e2 · outbound

This paper cites In: CVPR (2022).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2022)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.425399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:70e5a644dfd8b3b6abfa014ae39186b78b20f46a9fb81c561d56af75afd9069a

Observation 4f87267d-49a1-49eb-8ac7-aede43ad6df7 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.610426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:a4dffe8c9044444a09c65c1bdd48f5b80189db9d5bc6987c76e4ac01497d798d

Observation 43665c71-0b3d-4d79-8dca-aac461f56e88 · outbound

This paper cites In: ICML (2021) 24.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICML (2021) 24

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.412736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:68fde9ad920316c0d6e50ad4b19ce5a76485b77a3e5c95554f0e9ffcba68c172

Observation 97d7a366-2576-44c2-b432-4d7f49bcb8c3 · outbound

This paper cites JMLR (2020).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer JMLR (2020)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.400986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:370903e4e8f6bd686b1f297bfdb96a095b20d8b4abf1585cb80a56b052109c9c

Observation c61d1ed0-9308-41e4-bd81-02d159546de8 · outbound

This paper cites In: CVPR (2022).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2022)

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.386140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:1ce660f4590b80ffbba14eee0c284060ee4f4e59b983ef4a970ea1aa06813f17

Observation 5cbf63d3-4aad-4267-bac8-e2833cd20959 · outbound

This paper cites Seedream 4.0: Toward Next-generation Multimodal Image Generation.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Seedream 4.0: Toward Next-generation Multimodal Image Generation

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.615987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:edabd7f96bf29bd0669aaf5612bb287753d49d2479a59a24111dd43e1290ee2a

Observation d25aad1e-c1a4-42a2-a647-e71adebd6cc8 · outbound

This paper cites GLU Variants Improve Transformer.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer GLU Variants Improve Transformer

Reference 44

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.626212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:1582df4fcd7204b7f5d877c3d27ae26820bafff7f5a43bccc4a25b1db56e1ebf

Observation 6422eda1-78aa-4224-968f-29ef9f1e2ec6 · outbound

This paper cites In: CVPR (2023).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2023)

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.457522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:24b244f6962e6c2eca50f4c2f9bf31ce05fc526aeba4bad9aeb16a0f445b0f97

Observation c185cd1f-f978-4aaa-94a1-14a9704782ca · outbound

This paper cites Neurocomputing (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Neurocomputing (2024)

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.421498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:c4a8038be6fd625ee3f38c4c8bd2f913947473309429caaf9435471b16080dff

Observation 2078cf8d-408a-4d12-90ec-83363d006897 · outbound

This paper cites https://app.klingai.com/cn/ (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer https://app.klingai.com/cn/ (2025)

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.429379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:1eb019c8113576b33fafae3c0e1231f98bb47af5ec80ae9eec83d0d7b447e8ce

Observation e9b3cfab-41d8-4c3d-af72-011b99644dcf · outbound

This paper cites Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Z-Image: An Efficient Image Generation Foundation Model with Single-Stream Diffusion Transformer

Reference 48

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.555288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:25cb0e274d0d9f82550a7db1b34c48c7397ac888504956d1303bc66a5caaf2d6

Observation b3a7d686-3b0e-4660-8277-8b974026f496 · outbound

This paper cites SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features

Reference 49

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.566822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:fe1621504e2d972303ce864f22297f433c2810b85b73ee5f666ccc6026ddb905

Observation 203a4629-c506-40b1-95c7-6fcfb3c8b415 · outbound

This paper cites In: ICLR (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICLR (2024)

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.434234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:7b2045143968fb0c5c646fffdfe629aebcd75d8d9a5b75ca88cdac52d32ae4c2

Observation 3e383fbd-8d9e-4b13-b95c-92907063b594 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Emu3: Next-Token Prediction is All You Need

Reference 51

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.538240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:276a75aeb270026d6c8ede99447d88286649231f8da583f7dbb85d4641bcf4ac

Observation a69cbe8a-95f7-42eb-953b-17f17f639ad3 · outbound

This paper cites Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling

Reference 52

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.572856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:ecf0c46632f1de452d6b22c7eaa6c424985af212611289f8f524a8727c65c56b

Observation 8e92a159-38c2-46c8-8648-4e9bc6221ff3 · outbound

This paper cites Qwen-Image Technical Report.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Qwen-Image Technical Report

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.549443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:76976ca35fc86a2bffa61b71a812af7b7546cfaee883854809cca4b915317fac

Observation cdfc8470-3fbd-4dda-8c09-5149de8c4930 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 54

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.636734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:9e594ffd71a6250fb66d31f4ca572c1f018f943b1ebd2efef6cfe335776e74d4

Observation ab2dc494-82ce-4b4f-8278-ecf47f15a06d · outbound

This paper cites Dreamomni2: Multimodal instruction-based editing and generation.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Dreamomni2: Multimodal instruction-based editing and generation

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:32:29.585504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:0c9e48eaa3e0545182703650b2c9755a5b8ea633291ec28a44a05f45e78af697

Observation 7477b3ce-c5a3-4628-a231-aca8af8a4cec · outbound

This paper cites In: CVPR (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2025)

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.473659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:0d058eb2da37752b608140e955ec44e5c0f7a965d826996aa232360144087a93

Observation 5573006f-2d1b-4f44-a2ee-0ec082133e03 · outbound

This paper cites In: ICLR (2025) 25.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICLR (2025) 25

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.503507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:a13a6e24caeafe14ec24dcdcf335dc012d831ba36655683fe308fa851c3bf9cc

Observation 2982a70a-14f9-4342-a1c6-7fe1933be2b5 · outbound

This paper cites In: ICCV (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICCV (2025)

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.477658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:2e5b3ffdc7c617b543003cfc25ba90e5cd36ff7e565654de31786922ffb4525e

Observation e6f3e836-e3b6-44dd-b520-5a428ec346ac · outbound

This paper cites Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:32:29.604208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:7c3b8ea641b745205b342be4e6cc34d614d2cab2853f29464d88da82d87d107b

Observation d9128e45-ad7f-46fc-9c97-2100d66e12b1 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T07:32:29.631186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:88aeaf468884dd5504a8afc3ceac3065aa3b615fe73dba696889b12abae879a1

Observation fe45bf9d-c02d-4410-86c3-4c05307a13db · outbound

This paper cites In: NeurIPS (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2024)

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.485756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:dd65b316cea57f7ba71455dfd0971f23691c1617737993ae2a43e5e7bb5c544a

Observation dcf373ce-94b9-4cd5-aa47-b8b4c624181e · outbound

This paper cites In: NeurIPS (2019).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2019)

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.417266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:81a6e91d07325e96fe8ab7ed1823bd3c631233c8d30411e81d71df96cce31865

Observation de7575d5-74f1-4162-8156-c548dbad2634 · outbound

This paper cites In: CVPR (2018).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2018)

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.440984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:c410a4eee30d7b71edb4b9f2021babc93e2be2d82cb9a84b671395415f51feb5

Observation f17b5f57-04d9-41ca-b0d4-c7584fe0f8b1 · outbound

This paper cites In: NeurIPS (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2025)

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.409293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:59ca376224a82622162f78cf0bdcafe1356c5ce0ece4dc8987016b2196ac10e5

Observation 28a75958-6235-47de-8e2b-b2ed87465a26 · outbound

This paper cites In: ICML (2025).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: ICML (2025)

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.397045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:2cbe60d6001f64931f2a89f52014fc447d2fb6ead5d297dcd7eb4a4a1988b672

Observation 8389586f-d5f0-4e84-9b5a-46348f880015 · outbound

This paper cites 3dis: Depth-driven decoupled instance synthesis for text-to-image generation.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer 3dis: Depth-driven decoupled instance synthesis for text-to-image generation

Reference 66

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T07:32:29.642149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:8d18c1ead31040baf5177541cbc0309f2b0820fd31e63915ea900ea3cfa11435

Observation d0cbfb05-184b-4380-9bd7-e4c47abf3c3d · outbound

This paper cites In: CVPR (2024).

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: CVPR (2024)

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.507089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:b4125033f629b017b12797d4e947ccc19e6c82da6b5ec984c0707364dca4c728

Observation 3ea6c37c-51fc-494b-8114-0a2e06b9586a · outbound

This paper cites In: NeurIPS (2024) 26 Appendix A.

HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer In: NeurIPS (2024) 26 Appendix A

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T08:17:32.391035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:30:53.939221Z digest=sha256:85496de261b9ef2c34ff6ae8481d926405eb86888cbc5cc9c14e7ac92c2ebe52

Pith citing papers

Observation 0f9da1a7-4a34-4082-940f-c732308e2808 · inbound

Toward Native Multimodal Modeling: A Roadmap cites this paper.

Toward Native Multimodal Modeling: A Roadmap HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-06-29T23:04:01.951693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:58:38.610609Z digest=sha256:115eb9a5cee9b6585cd098da2f53950cb329c682d75c2c562fd46a21d93786c8

Observation bb2a5259-2619-4d5c-89ba-806ec6451423 · inbound

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization cites this paper.

WeGenBench: A Multidimensional Diagnostic Benchmark towards Text-to-Image Model Optimization HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-04T03:39:29.342711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:54:09.656061Z digest=sha256:8e23872b724144e39020dc0dae533c33a9dc8c0e49f55d664c86bfa9399a84be

Observation 19428b20-9a48-46a5-b5a0-e7084c4bf3a8 · inbound

DiffusionBench: On Holistic Evaluation of Diffusion Transformers cites this paper.

DiffusionBench: On Holistic Evaluation of Diffusion Transformers HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 213

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T16:59:58.074828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T00:06:11.951205Z digest=sha256:969239965a12d0c447d9c6ed549d0a515180fffcfea993e0bc6d6e1e2ec4a506

Observation 7c448b51-6336-4511-b2b3-f0c670e2ca42 · inbound

Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget cites this paper.

Boogu-Image-0.1: Boosting Open Agentic Multimodal Generation via Understanding under a Minimal Budget HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:13:50.834073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:13:50.834073Z digest=sha256:48b1960a1eca32a99a957d4e9c69ba40d0c39b23085787f7db263de6d1454d2f

Observation 28ae26fe-1b61-4390-a6b2-06388b361900 · inbound

Pixel-Space Diffusion Transformers cites this paper.

Pixel-Space Diffusion Transformers HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-01T17:35:35.295064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:35:35.295064Z digest=sha256:a4f6b2b7f629edb2ae41ebb49406f3e87455fc4953421d1ff6202ba08702fb6b

Observation a6902df7-c41c-4bd3-a0a3-b84a0323650e · inbound

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results cites this paper.

The Second LoViF 2026 Challenge on Real-World All-in-One Image Restoration: Methods and Results HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T08:26:05.138511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:26:05.138511Z digest=sha256:514b6c4ddf03d4cc762648e897d56b701528fff9782f6ba8e61c2522fe1f46db

Observation 6cbb03c2-804e-43c6-96bc-03936f7f091a · inbound

Test-Time Curriculum for Open-Set AIGC Detection cites this paper.

Test-Time Curriculum for Open-Set AIGC Detection HiDream-O1-Image: A Natively Unified Image Generative Foundation Model with Pixel-level Unified Transformer

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T00:46:23.336119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:46:23.336119Z digest=sha256:fe2c453da58dd4592ca4abc15a4f7d19f228bb49c13e77e30106cc451d93f3b0