Pith. sign in

Paper Citation Record · LEDGER

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing

As of 18 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 3 inbound Pith citation observations for arXiv:2603.11593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.11593 v2

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T18:25:59.357898Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T21:23:41.271521Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T21:28:58.273805Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved95
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 497a3e10-6c61-4278-a297-7ab13d62553e · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.186388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.186388Z digest=sha256:70dfd54e5ffb720dc87d8147203509db5b066134fa64a1d541788d8c4115974e

Observation fa3916fd-18c8-4fc3-95e1-c7e840d87ed8 · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.287784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.287784Z digest=sha256:0166c7cbdb3fa6396bb78a5635d24783268f7aaef81de06646051d819009f579

Observation 146445f9-fe9d-435d-a2d8-f9b32b0608bd · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.345344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.345344Z digest=sha256:8c1112da765c9cc6522b5330bd712c424c4b9edee2c53f14553dc6d07d706c37

Observation 45f39006-f063-4ff9-b389-ceefe0151d11 · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.413670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.413670Z digest=sha256:bc130a46100fb0ca366ab8e4909e1f252b3706a06e7d739bba4cff153ad7e7af

Observation 684a42c1-beb0-4463-901b-d4805feb1d91 · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.495216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.495216Z digest=sha256:b5bef193d4efcae8900074b8b942fd0be59f62aa7c9955b4535a59ae0908b9ab

Observation 568aecae-e2c8-4b06-af5d-15381dcefd73 · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.551475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.551475Z digest=sha256:32824a9e0db68f2497f26c4a856cf8716508b94bac6a89e1c430a44d2a19ba32

Observation 8de09d47-effc-44f4-a828-ee9e6c65fd7a · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.662214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.662214Z digest=sha256:9fa51f80052053e40ab33c05f474eba8f878fc3481be2a2a0a7679ca5ed8b087

Observation 99e45f93-8860-4851-a687-514092a880b0 · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.733512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.733512Z digest=sha256:3b4ef28f2817f704fc0c4db97f63d34a7eb6c0345b5a0e61f8f1fb15a2434090

Observation f97792dd-c1c8-4721-b4cc-e8940d11700e · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.804854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.804854Z digest=sha256:b5d99630ffc90360018eed809490f4035a7ee64f165a2b883cb457fe5c92d41a

Observation ad7fb4d6-b8ec-4fdc-bcb6-f471c42695fb · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.867889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.867889Z digest=sha256:1fa6cc994f947167856db442f303fe8e580488e7cc56b9ab37b8dbaa46dd5d15

Observation 0ca1954e-efa5-4bd0-8c96-4b7f3a0e05cb · outbound

This paper cites an unresolved cited work.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:50.939236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:50.939236Z digest=sha256:12030eee2e769716701a7e8beef5b6ba42053b2ccbc29ff0cc2643dfecf5cdda

Observation c128f45f-1d81-4b06-bc87-7dadb1a09a27 · outbound

This paper cites https://blog.google/innovation-and-ai/technology/developers-tools/ gemini-3-pro-image-developers/, 2025.11.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing https://blog.google/innovation-and-ai/technology/developers-tools/ gemini-3-pro-image-developers/, 2025.11

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.026361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.026361Z digest=sha256:a584fa6768a72ccba43c5d246f0ee2b957dc69b8acedfc11f3fbb49d9919df67

Observation c835e926-ff75-476b-bb2c-d70d00eff150 · outbound

This paper cites Qwen3-VL Technical Report.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Qwen3-VL Technical Report

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.149194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.149194Z digest=sha256:fbfdd39b9c3ce1c12e681c87e31636f20ea9a75efdff9d4eceacc6572baac49c

Observation 7c023666-e061-4857-8ae5-f1f9810cbc0f · outbound

This paper cites Instructpix2pix: Learning to follow image editing instructions.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Instructpix2pix: Learning to follow image editing instructions

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.304223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.304223Z digest=sha256:f0e037adfad36d87a8abe69fe0053f481dee9130256ac757ee19dfbea33bcda2

Observation b40c3b83-26c3-4cf3-b1bb-bc1931208efb · outbound

This paper cites HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing HiDream-I1: A High-Efficient Image Generative Foundation Model with Sparse Diffusion Transformer

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.463871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.463871Z digest=sha256:ef29f3bb3bee7351d747ae3fa995a373410892442d534c76c5e56e6a79f84d5c

Observation 2567134f-1279-4405-a264-17c4562c088e · outbound

This paper cites HunyuanImage 3.0 Technical Report.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing HunyuanImage 3.0 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.624195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.624195Z digest=sha256:a6e10e96cae2bb316fc5a0267ab70d5bd2acd02c7b2b2bd87665113d5b9bd4a5

Observation bf30f40b-4c19-46a6-ae56-9b63bfd7883e · outbound

This paper cites ByteMorph: Benchmarking Instruction-Guided Image Editing with Non-Rigid Motions.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing ByteMorph: Benchmarking Instruction-Guided Image Editing with Non-Rigid Motions

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.802196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.802196Z digest=sha256:1dd55cc1030065d29b59d4706044b631db42b61dcc5bce6424802120c471197f

Observation 456bbd3a-cab7-4738-a983-a598b0185aa9 · outbound

This paper cites Textdiffuser: Diffusion models as text painters.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Textdiffuser: Diffusion models as text painters

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:51.947062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:51.947062Z digest=sha256:b09838a8987970c74812f75202da2e3898164ec488eeefd383bcbcfa70be89d5

Observation 4e31a50d-ece1-4863-8faf-69584b48a3fa · outbound

This paper cites Textdiffuser-2: Unleashing the power of language models for text rendering.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Textdiffuser-2: Unleashing the power of language models for text rendering

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.104904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.104904Z digest=sha256:d56a7e58a6f51c6d244d3c2ab4ef5c77be6b052ea37e90be92c071d03e8147af

Observation eb830778-b908-4c77-8592-84a3490ae6d5 · outbound

This paper cites Pixart-α: Fast training of diffusion transformer for photorealistic text-to-image synthesis.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Pixart-α: Fast training of diffusion transformer for photorealistic text-to-image synthesis

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.342580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.342580Z digest=sha256:32f1c3cc46d427ccdaf2d26db6a4262c7d0e9f3c1294395a51c20c1990e6b22b

Observation e2391c0d-6086-4f42-b788-8402b6c26b5d · outbound

This paper cites ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.501547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.501547Z digest=sha256:bd7c854fa1a8cd5e0fabe2b06f06523d4e22180a705be1668c6d11ef505c8fe5

Observation 35772a41-bbc4-4737-b365-352d08d46ff9 · outbound

This paper cites Emu3.5: Native Multimodal Models are World Learners.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Emu3.5: Native Multimodal Models are World Learners

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.604284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.604284Z digest=sha256:2f854b4fbe108496c846ed194c4829453026d2e820a0b0b069a6af7ed1bf96a3

Observation 1a42ac87-73fd-44eb-81cb-aa8f4a00da49 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Emerging Properties in Unified Multimodal Pretraining

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.779316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.779316Z digest=sha256:6af505ad2f6b57522d82c0106003334f02f9d5123be53533895631fe215f7368

Observation 51b6c7f2-3be6-4c69-82ed-ed014c7f86d9 · outbound

This paper cites Diffusion models beat gans on image synthesis.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Diffusion models beat gans on image synthesis

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:52.978785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:52.978785Z digest=sha256:225512c98b156b1b63d9e01a0440c96082ba521cfbd9afe172adef8690890770

Observation d3a00fa5-fd29-44a6-82b5-16b0c64d03a7 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Scaling rectified flow transformers for high-resolution image synthesis

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.215287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.215287Z digest=sha256:30d5dc2c94f71502a14e8292cfc9348eae5d0ca76974aba786d1983a0da1d4e4

Observation a1c7b072-6116-4d0d-9a7a-7cef17bd31ae · outbound

This paper cites Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.388773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.388773Z digest=sha256:202ae3e841b35abec2dfecc75bc8fdab95776be20e327bea992d81429c0ebb85

Observation 919939fb-0d9f-4c8d-a67c-7a9c0b4a7902 · outbound

This paper cites Seedream 3.0 Technical Report.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Seedream 3.0 Technical Report

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.536048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.536048Z digest=sha256:f1f3f505e0f4e3ad341262db6c6c96f915a52875f3e69d45aac9425f3852aae1

Observation eadedf43-2507-4b54-b214-32650a25e9df · outbound

This paper cites Unireditbench: A unified reasoning-based image editing benchmark.arXiv preprint arXiv:2511.01295, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unireditbench: A unified reasoning-based image editing benchmark.arXiv preprint arXiv:2511.01295, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.663986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.663986Z digest=sha256:dd5d74dbdfadc858e20949a41c9785ba68c139e81b6821e9e2e874c4ed87e0ad

Observation 09c77f50-4dae-46dd-8e64-9410fe17cf2c · outbound

This paper cites Prompt-to-prompt image editing with cross-attention control.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Prompt-to-prompt image editing with cross-attention control

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.819052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.819052Z digest=sha256:52c3f0ab74d91939984d9255609637c6f39e28058277cebf97b61f8f1757e1d0

Observation 331bf65c-f5cc-49c9-9333-b1972737565e · outbound

This paper cites Denoising diffusion probabilistic models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Denoising diffusion probabilistic models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:53.968106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:53.968106Z digest=sha256:1d08cd88c35d5b8c54d1153eb8e3df0bedd8a78bcac453b24c0f92ea11ff75bf

Observation d1649735-bfc7-4142-8104-870464c24b91 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Lora: Low-rank adaptation of large language models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.068225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.068225Z digest=sha256:991de1a2bc2c6bce854e4e14ef2360b3ec968c4199e69bc0043d3c6e40d29542

Observation e56911de-a6d2-4c4b-be43-06436abed170 · outbound

This paper cites In-Context LoRA for Diffusion Transformers.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing In-Context LoRA for Diffusion Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.225573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.225573Z digest=sha256:9eb6bc4d46ca4ba98ea3adbe37445d4a57af68777787a31a4ff026f985f222b7

Observation 3fdad22e-0ecd-430e-a9d5-0526859c900f · outbound

This paper cites Hq-edit: A high-quality dataset for instruction-based image editing.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Hq-edit: A high-quality dataset for instruction-based image editing

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.311943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.311943Z digest=sha256:fe3cf0a3e220956de3b4fce91917dc5b66f867c4340011bb35c6ad8213b7a26b

Observation 872371d5-98bb-46c7-8d72-2500494c21c9 · outbound

This paper cites Leopard: A Vision Language Model For Text-Rich Multi-Image Tasks.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Leopard: A Vision Language Model For Text-Rich Multi-Image Tasks

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.474338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.474338Z digest=sha256:392203ceffa394483f7c074c097a26e9c60cfb9fc007e8d657cebba43b59bb6a

Observation 06e74491-959a-417f-8b03-cd91c75ea3e3 · outbound

This paper cites Auto-Encoding Variational Bayes.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Auto-Encoding Variational Bayes

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.610097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.610097Z digest=sha256:e025724791381802afaef6745b9a14d40802e90ed3d7594a1769bbc5debd546e

Observation 5cfd87b4-bac1-457d-9fe1-a919a15b76d8 · outbound

This paper cites Pick-a-pic: An open dataset of user preferences for text-to-image generation.NeurIPS, 2023.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Pick-a-pic: An open dataset of user preferences for text-to-image generation.NeurIPS, 2023

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.745982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.745982Z digest=sha256:5de92151460ae93262a29822bb96c4e5af5ffd183e69c275f68bad9514121564

Observation c334d8dd-f2d8-43f7-b8c7-452b0e6eb3c9 · outbound

This paper cites Flux.1-dev.https://blackforestlabs.ai/announcing-black-forest-labs, 2024.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Flux.1-dev.https://blackforestlabs.ai/announcing-black-forest-labs, 2024

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.809757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.809757Z digest=sha256:36d343327082f26432b10887f32603c7d1a5f2bbbc818503a8213908d196da12

Observation 495d9cf3-c497-4509-8b50-beb7b27a7293 · outbound

This paper cites Flux.2-dev.https://huggingface.co/black-forest-labs/FLUX.2-dev, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Flux.2-dev.https://huggingface.co/black-forest-labs/FLUX.2-dev, 2025

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:54.871940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:54.871940Z digest=sha256:db1f4f0962ddcfc575a3de637721df69c139ff882df851c3058bb811d930e447

Observation 6bf21ce2-0a0d-407f-9625-e62dfefbb164 · outbound

This paper cites FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing FLUX.1 Kontext: Flow Matching for In-Context Image Generation and Editing in Latent Space

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.060368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.060368Z digest=sha256:6bde087c78e22b1627d8611a5d5928562c7143f76a9e32f757ea6d622eb52690

Observation bba047c2-70b0-4cae-b344-e39b02c06f8e · outbound

This paper cites What matters when building vision-language models? NeurIPS, 2024.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing What matters when building vision-language models? NeurIPS, 2024

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.213315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.213315Z digest=sha256:0d807ed8152b2644fe1098156d25aa7ebf3f7648923379f5a4acbc665cc9f0f0

Observation 9082ce14-9fd9-4036-84f6-92f924a6e56b · outbound

This paper cites Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Unlocking the conversion of Web Screenshots into HTML Code with the WebSight Dataset

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.359076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.359076Z digest=sha256:65a364714b9d0037cc2b76f325e8bab70fd8d01f3e79b8a1de8952079e611e28

Observation b53cd30e-aeaf-490c-a020-e02ce5c46acd · outbound

This paper cites MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing MixGRPO: Unlocking Flow-based GRPO Efficiency with Mixed ODE-SDE

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.487831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.487831Z digest=sha256:e8c7dbfbf824c65d22686046c912ab0ea2a7f397260889499bfa3b91018e7012

Observation 30d4fdce-1a39-4ca0-ae5f-eba2bd783e86 · outbound

This paper cites Hunyuan-dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding.arXiv e-prints, 2024.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Hunyuan-dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding.arXiv e-prints, 2024

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.562254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.562254Z digest=sha256:a0bcc77bf8b7efa240301e85c0cd99a64519feb684d00a8d23cd567cb4c6df29

Observation 46220900-93bc-4278-bf64-e650f2e2257e · outbound

This paper cites Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Uniworld-V2: Reinforce Image Editing with Diffusion Negative-aware Finetuning and MLLM Implicit Feedback

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.636781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.636781Z digest=sha256:8cf61d3e279fc2beed7ea55981d3e7c17586b230eabba2dda5acd7c3a8e6e963

Observation 7b1be1e1-25e7-4926-9bb4-68e645d6a4c9 · outbound

This paper cites UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.733988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.733988Z digest=sha256:27f048e3b3e3bc80e8c8454e0b13c3e0318116a7d5e3944503b61ba7579c2d84

Observation fa2482c6-3010-442f-898e-3e3ff850af99 · outbound

This paper cites Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Playground v3: Improving Text-to-Image Alignment with Deep-Fusion Large Language Models

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.817521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.817521Z digest=sha256:cd1e64af957dbed36d8cfb4202ef0c3544c9e07c2af99bcb1b4b77ad06ce0feb

Observation 1fe61889-8328-4bf2-bd87-4fe5ac2784db · outbound

This paper cites Flow-grpo: Training flow matching models via online rl.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Flow-grpo: Training flow matching models via online rl

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.869754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.869754Z digest=sha256:3aeb516ba55cd0caa2c353a6487a0b038d36a56cdafb8888543d5d1a2c5c93af

Observation c844d8df-2da4-48e5-b9c0-3d29eae4278d · outbound

This paper cites Step1X-Edit: A Practical Framework for General Image Editing.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Step1X-Edit: A Practical Framework for General Image Editing

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:55.936773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:55.936773Z digest=sha256:92f951532b4ad5a9650ac078012ce433241679eeea46ddbdda6d3535d6a126a2

Observation 1fcd1dfb-8d87-4f65-8286-f13486f64d2d · outbound

This paper cites Flow straight and fast: Learning to generate and transfer data with rectified flow.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Flow straight and fast: Learning to generate and transfer data with rectified flow

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.011659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.011659Z digest=sha256:6c41701253fd914546b3506385abb433d8a1a2ce5cfe48955f792c746be0907e

Observation 2dc8c338-17d4-4d84-8b48-7411e84012f7 · outbound

This paper cites Glyph-byt5: A customized text encoder for accurate visual text rendering.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Glyph-byt5: A customized text encoder for accurate visual text rendering

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.085866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.085866Z digest=sha256:4619f8d5b29c2d1eaf12c25f0d7c79c90705c0ea17f024b1e51bf8603a3baeca

Observation 5b01989d-772e-4e26-a67c-ca269beabb75 · outbound

This paper cites Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.191301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.191301Z digest=sha256:7e2d57166736687610c18be4f49338f37c7156226a38c1b1ecf392f4fc4ef2e2

Observation d4c4c8d9-0b1f-44d7-b4d7-3212ace3aede · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.NeurIPS, 2022.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.NeurIPS, 2022

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.283730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.283730Z digest=sha256:2e67fc7f0c835ced76c7c005d89649e71b1c40e27c2fdce1278ec6669607d228

Observation 53b7500e-ca88-4854-8171-9f19b3084096 · outbound

This paper cites GlyphDraw: Seamlessly Rendering Text with Intricate Spatial Structures in Text-to-Image Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing GlyphDraw: Seamlessly Rendering Text with Intricate Spatial Structures in Text-to-Image Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.357238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.357238Z digest=sha256:b82c71174efb1ab11e95d957d4cb8701c7bca322cf5c49feec55d648de4c12f0

Observation 7dc32abf-7197-457c-8f8c-9f16fd6096e9 · outbound

This paper cites X2edit: Revisiting arbitrary- instruction image editing through self-constructed data and task-aware representation learning.arXiv preprint arXiv:2508.07607, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing X2edit: Revisiting arbitrary- instruction image editing through self-constructed data and task-aware representation learning.arXiv preprint arXiv:2508.07607, 2025

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.457872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.457872Z digest=sha256:aafd7fe75d76bfa8da655611f10006fda3b1104f61441ad4de826b6df9418903

Observation 0294cc2f-ca87-4f7d-aa2c-ee30d5908766 · outbound

This paper cites Scalable diffusion models with transformers.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Scalable diffusion models with transformers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.564178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.564178Z digest=sha256:d8b88b4ec5c523d5486ee45c36fc3bd5ab41f64a1b1b61b93585b17ce2bee4b6

Observation dc4062f6-1556-4adf-97c1-c8ac0c018501 · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.633486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.633486Z digest=sha256:ccaa91a462b23cbcb6316638ec77f762bfbb38a9ffcb81d464d4e310294eb826

Observation fc0b4974-ded4-4683-82f7-ea997df2c7ee · outbound

This paper cites Pico-banana-400k: A large-scale dataset for text-guided image editing.arXiv preprint arXiv:2510.19808, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Pico-banana-400k: A large-scale dataset for text-guided image editing.arXiv preprint arXiv:2510.19808, 2025

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.702842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.702842Z digest=sha256:d4fd8f6c603b3f3e43f5c5abe427006dc1bc6bdbca5c785e1a2df1f3e1136359

Observation 5a01faa9-f004-41db-904e-213313b699ea · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing High-resolution image synthesis with latent diffusion models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.790489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.790489Z digest=sha256:ae8c84d035797c69932f38d783ed5cb670a355cce5798b312392b5eb097fd225

Observation 08ea00fd-9087-42be-9b16-2a47ae0e877d · outbound

This paper cites Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Dreambooth: Fine tuning text-to-image diffusion models for subject-driven generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.867973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.867973Z digest=sha256:f8372d2cd878bde11cbd33ce366943409830995cee8dc213198b5d7e81cdb52b

Observation 1384296f-078b-4a63-8a99-3c3f469c2269 · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Photorealistic text-to-image diffusion models with deep language understanding

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.934873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.934873Z digest=sha256:eb6886e1bec15172937ffa819311b3ec7bb387502a1db3887a5ec80da62f0c4a

Observation 0c15b93a-2886-430f-8b47-b98ff3dabaa2 · outbound

This paper cites Denoising diffusion implicit models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Denoising diffusion implicit models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:56.988929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:56.988929Z digest=sha256:4be295b83dc3e0429f4bf411ac274a04bca84f1e8627f83c5e9bc98e31d8fdb7

Observation dea0fc42-9d44-41b5-88cb-4a6ae793da5d · outbound

This paper cites MIT press, 2018.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing MIT press, 2018

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.094195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.094195Z digest=sha256:2a7b600db56a24222db48ca9c80e0cc8e3155346119f1d882df5a3cae6572e09

Observation d21b5d1b-19af-4291-9c77-d77f5f3010fc · outbound

This paper cites Ominicontrol: Minimal and universal control for diffusion transformer.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Ominicontrol: Minimal and universal control for diffusion transformer

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.180839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.180839Z digest=sha256:9bd458d6cf62456657ed31d47e6e3a89fbb839af128bd364bcce2a3184825ada

Observation 90ecf533-48e8-4a48-bf94-91fda535ab25 · outbound

This paper cites LongCat-Image Technical Report.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing LongCat-Image Technical Report

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.227130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.227130Z digest=sha256:bd8ea7d1dd86ffe0394e01e45d9c76596c5fee7155f9fb32034ad54377a43093

Observation d20801b6-3697-433d-bce6-aad5659dc08a · outbound

This paper cites Firered-image-edit-1.0 techinical report.arXiv preprint arXiv:2602.13344, 2026.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Firered-image-edit-1.0 techinical report.arXiv preprint arXiv:2602.13344, 2026

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.312745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.312745Z digest=sha256:452e00c5e611c5440df8a68bc90893ad803b2a44b12ed847e4675d62e19b8b02

Observation 22e4bc87-69f7-4303-8334-90d22498be55 · outbound

This paper cites AnyText: Multilingual Visual Text Generation And Editing.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing AnyText: Multilingual Visual Text Generation And Editing

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.379685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.379685Z digest=sha256:032019ef8f498e865be8ce42754fab00487d0b879a980d087f369ea3180c0c4a

Observation 3c6ff124-2ebf-49e2-be4e-9ed7b45cf401 · outbound

This paper cites I2i-bench: A comprehensive benchmark suite for image-to-image editing models.arXiv preprint arXiv:2512.04660, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing I2i-bench: A comprehensive benchmark suite for image-to-image editing models.arXiv preprint arXiv:2512.04660, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.439520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.439520Z digest=sha256:ea53abc73aad8d7a35e4837197813f67f0847d95942bd661c1dd1b4e199c2821

Observation f74e0d7c-be25-4471-8ba8-b45fe48e28f6 · outbound

This paper cites GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing GPT-IMAGE-EDIT-1.5M: A Million-Scale, GPT-Generated Image Dataset

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.474554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.474554Z digest=sha256:4450019643e915c9143ae73741ce7d60e04c889448e979bf2a0ba5d2e3a3f792

Observation 1d2f2c6e-9ede-441e-8e68-bdae169a014c · outbound

This paper cites Omniedit: Building image editing generalist models through specialist supervision.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Omniedit: Building image editing generalist models through specialist supervision

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.523572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.523572Z digest=sha256:4eec6a3284ea9bb2c0afed4e71368488185af133cd3aef2800d86ac3d949ceb5

Observation f009c49c-861d-4c5b-9dde-b8404fe561d4 · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.NeurIPS, 2022.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Chain-of-thought prompting elicits reasoning in large language models.NeurIPS, 2022

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.573589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.573589Z digest=sha256:c7539df8530a113f9acd442df65756e345906db7e12f4d09e5d067905bdc1c39

Observation 411af40f-e151-46cf-b848-fdc318aa1ce5 · outbound

This paper cites Qwen-Image Technical Report.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Qwen-Image Technical Report

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.635235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.635235Z digest=sha256:905739880f00a54fcb7e09e90269f35df0505bbb0f7f691ee3273a5f210461b5

Observation a3bdcf0c-3a3a-4055-8e2e-283613f01ae2 · outbound

This paper cites OmniGen2: Towards Instruction-Aligned Multimodal Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing OmniGen2: Towards Instruction-Aligned Multimodal Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.683143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.683143Z digest=sha256:d4cb13e0a62a7618a767d6fc6727bee4047a8e838ccc607438381e5642bf71e8

Observation 063dd370-6886-4a33-91fd-024be025a392 · outbound

This paper cites Chronoedit: Towards temporal reasoning for image editing and world simulation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Chronoedit: Towards temporal reasoning for image editing and world simulation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.758850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.758850Z digest=sha256:a9b67b5a2c2bc8326eec92f90b112f9b8eba6415fee0c8aef5086fc7f3cb5888

Observation 14b1cc6b-39cb-4900-bcbd-b2a2c54db953 · outbound

This paper cites Editreward: A human-aligned reward model for instruction-guided image editing.arXiv preprint arXiv:2509.26346, 2025.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Editreward: A human-aligned reward model for instruction-guided image editing.arXiv preprint arXiv:2509.26346, 2025

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.918642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.918642Z digest=sha256:fb3f44311f05acfb4c99e666e4138a3cba65a2d2227585ea837a5740b5a33d8a

Observation 80dd6ab2-3b5f-47e4-a4ae-00b74a611bd7 · outbound

This paper cites Less-to-more generalization: Unlocking more controllability by in-context generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Less-to-more generalization: Unlocking more controllability by in-context generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:57.965240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:57.965240Z digest=sha256:cfd00cf3f636d0aa527f0d8d7bcb60e82bbd7641811dabdcc4feedd9945c07a2

Observation 4d28bdc7-a608-4efe-9acd-af421db85222 · outbound

This paper cites Omnigen: Unified image generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Omnigen: Unified image generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.019139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.019139Z digest=sha256:0d03555f8eb3d6fe828eed6f2e9c5612c634bfb63c00d2fa084ee4996934b456

Observation eb5f8bdd-de55-4791-ac88-71976b7a85ae · outbound

This paper cites Imagereward: Learning and evaluating human preferences for text-to-image generation.NeurIPS, 2023.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Imagereward: Learning and evaluating human preferences for text-to-image generation.NeurIPS, 2023

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.095839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.095839Z digest=sha256:2f2211dc8ba88c8b06f01b387090032c6c10352b6348dfa54fc98d3cfa501e15

Observation e6718861-7cda-4028-a645-ffdad54657b9 · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing DanceGRPO: Unleashing GRPO on Visual Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.162492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.162492Z digest=sha256:b03d2cebe223dc3a7021f68dde65aaffec803525a3ac262f75491f68d338fffa

Observation 4cac92d7-2605-46d4-b882-db3732ea784e · outbound

This paper cites Glyphcontrol: Glyph conditional control for visual text generation.NeurIPS, 2023.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Glyphcontrol: Glyph conditional control for visual text generation.NeurIPS, 2023

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.243104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.243104Z digest=sha256:785504df3ae8c9363c4faf8cf6dbd35bae0b0901d92f3cb4e7d55856a57faf61

Observation 431cceef-dc42-4ef3-a2e3-06fdd4459c29 · outbound

This paper cites Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Echo-4o: Harnessing the Power of GPT-4o Synthetic Images for Improved Image Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.290824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.290824Z digest=sha256:17f13188b6fbb107b97405e0780a18c22a0d3256a93e70d257e5cc55b896b51d

Observation 0990e47d-aaf8-45e9-885f-c7077b2a7913 · outbound

This paper cites ImgEdit: A Unified Image Editing Dataset and Benchmark.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing ImgEdit: A Unified Image Editing Dataset and Benchmark

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.358063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.358063Z digest=sha256:b46978d9f1218edc70de1e809bf9864b0279479d78de4aee5f7d31bb444445f8

Observation 51df0a60-8feb-4791-9d38-1068ef9c86ca · outbound

This paper cites Anyedit: Mastering unified high-quality image editing for any idea.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Anyedit: Mastering unified high-quality image editing for any idea

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.438027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.438027Z digest=sha256:708407fa3a3eae31ed370e8e3319f753d30e43e79a4a4ebdeae075954cd72639

Observation 12839497-cd86-48bf-817d-5a082a2eab51 · outbound

This paper cites Creatilayout: Siamese multimodal diffusion transformer for creative layout-to-image generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Creatilayout: Siamese multimodal diffusion transformer for creative layout-to-image generation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.514838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.514838Z digest=sha256:dff6cbb028543aac162d13b9446d98b12ca9a154a873b6f2d8eae4e2538a91cd

Observation f7b425f2-8f5a-4f8f-b815-58e6af5e864c · outbound

This paper cites Creatidesign: A unified multi-conditional diffusion transformer for creative graphic design.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Creatidesign: A unified multi-conditional diffusion transformer for creative graphic design

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.566345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.566345Z digest=sha256:e46671b5bf9261ecdf438205fc775281124b20565cd215dc9e2813dfbc714e2f

Observation 7540a95b-56ab-4e79-8db2-7e0449bde025 · outbound

This paper cites Magicbrush: A manually annotated dataset for instruction-guided image editing.NeurIPS, 2023.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Magicbrush: A manually annotated dataset for instruction-guided image editing.NeurIPS, 2023

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.625258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.625258Z digest=sha256:9dda72d0befc98980df316caca7a9435fd4d688471243fe67191c9ea6793dad7

Observation 2ee31b4d-eeb3-460f-9210-3b8c60e2bde9 · outbound

This paper cites Enabling instructional image editing with in-context generation in large scale diffusion transformer.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Enabling instructional image editing with in-context generation in large scale diffusion transformer

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.691643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.691643Z digest=sha256:1075ffdca194034275c94d5af685624e172f773cbd2c6ae977354cb991d2ba61

Observation f601a3ca-b9ea-47a5-9d97-14c509e5cc1c · outbound

This paper cites CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing CreatiPoster: Towards Editable and Controllable Multi-Layer Graphic Design Generation

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.771116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.771116Z digest=sha256:e4e17da6b17f609fc6690f833e4c5f465f402c9b858ba9f815487ef1149d9ec1

Observation 989c3d6b-5231-4c63-9f74-9d3b9353c53e · outbound

This paper cites Ultraedit: Instruction-based fine-grained image editing at scale.NeurIPS, 2024.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Ultraedit: Instruction-based fine-grained image editing at scale.NeurIPS, 2024

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.839910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.839910Z digest=sha256:4482e4a4f02680527fc3d727e583feae3b3be5e41936fc67cb475e050e212090

Observation 1cf4fa16-dfd6-4b08-b7f6-d075b5abe4a3 · outbound

This paper cites Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Envisioning Beyond the Pixels: Benchmarking Reasoning-Informed Visual Editing

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.919055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.919055Z digest=sha256:633677b7170f3efda32522d307cecaa3b12813f3e428af29e65e2ebdef3b4d0a

Observation a1834dcb-3031-4e85-b99b-d94bf50e2f80 · outbound

This paper cites Udifftext: A unified framework for high-quality text synthesis in arbitrary images via character-aware diffusion models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Udifftext: A unified framework for high-quality text synthesis in arbitrary images via character-aware diffusion models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:58.963382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:58.963382Z digest=sha256:78c5bbb2cf9c98611ff613bb62585ff15e7b60f78b9c3c550fdbff8edecfd69d

Observation 06b307c4-bef8-4c60-994c-6850f400a37d · outbound

This paper cites DiffusionNFT: Online Diffusion Reinforcement with Forward Process.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing DiffusionNFT: Online Diffusion Reinforcement with Forward Process

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:59.032945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:59.032945Z digest=sha256:5f10aa173049ee327fd4c20218433dc4ce52b799bf75b93920105c555c79548d

Observation 1522fcd8-63bc-4efe-b417-781b6cd67933 · outbound

This paper cites Cogview3: Finer and faster text-to-image generation via relay diffusion.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Cogview3: Finer and faster text-to-image generation via relay diffusion

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:59.113573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:59.113573Z digest=sha256:332a6f7eed216ac4d28d3662bf73c053950d6423279b7092a5165a69e4cec594

Observation 872000da-919a-4644-9c49-aa51132d62dc · outbound

This paper cites Migc++: Advanced multi-instance generation controller for image synthesis.TPAMI, 2024.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Migc++: Advanced multi-instance generation controller for image synthesis.TPAMI, 2024

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:59.205393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:59.205393Z digest=sha256:f1a6d5349f7f3ec48e6744158af68af8de227d502358e64733123542861a0c49

Observation 7c2f8687-c5f5-4576-a668-b86236d8a422 · outbound

This paper cites Migc: Multi-instance generation controller for text-to-image synthesis.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Migc: Multi-instance generation controller for text-to-image synthesis

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:59.288837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:59.288837Z digest=sha256:d8f8b1ab0f887d3c38275f23f5dea79e5a331631ff62153b8f52f09fab2cc45b

Observation 05e95358-a391-4f02-8e55-547ce4fd6325 · outbound

This paper cites Dreamrenderer: Taming multi-instance attribute control in large-scale text-to-image models.

WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing Dreamrenderer: Taming multi-instance attribute control in large-scale text-to-image models

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-02T18:25:59.357898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:25:59.357898Z digest=sha256:9472789ae9c90b05d85e87a1d82bc774f1a6390abf363ba90b9572277e804a5b

Pith citing papers

Observation df92818b-f87d-41e2-82f7-820b887676e6 · inbound

TextSculptor: Training and Benchmarking Scene Text Editing cites this paper.

TextSculptor: Training and Benchmarking Scene Text Editing WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-20T03:19:15.682913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T05:16:43.756525Z digest=sha256:607b264393c57e96750092fd26505188a689aa145096954bb77b2cc3810b84ad

Observation df6303bf-c7bf-4db1-8a01-e1b46657565a · inbound

GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images cites this paper.

GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-20T03:19:15.682913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-02T13:56:43.671622Z digest=sha256:6244ba3ac7703fae8c647f230f332509b95f44f4f1c2e128330bdba49cc4e165

Observation 2f1b7dfd-0f28-4722-907a-2227c894b1d4 · inbound

GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images cites this paper.

GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images WeEdit: A Dataset, Benchmark and Glyph-Guided Framework for Text-centric Image Editing

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-20T03:19:15.682913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-03T21:23:41.271521Z digest=sha256:22dcd600607ee868da605f107a4e0dd247cff7060793942401fb5527d3598aae