Pith. sign in

Paper Citation Record · LEDGER

VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2312.14867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.14867 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:43:11.291929Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:59:52.750964Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a2c2b204-02f5-46ff-9b8f-feecd4f956e6 · inbound

VideoPhy: Evaluating Physical Commonsense for Video Generation cites this paper.

VideoPhy: Evaluating Physical Commonsense for Video Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:34:37.805267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T11:34:37.599691Z digest=sha256:bea2ad001c8fbd57580e9df32d0acde78bbcf37ea935c0892fb73294dd10908d

Observation 8da70949-c9f3-4256-9c3e-a6a79f2ac986 · inbound

InsightEdit: Towards Better Instruction Following for Image Editing cites this paper.

InsightEdit: Towards Better Instruction Following for Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T12:19:22.467898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:19:22.467898Z digest=sha256:de8714a6fc1ac64eae8c48fb583695984fb76bf9b0fa61452077f1897fe888fa

Observation 1cabf55d-34c7-4467-922c-2bf70c40988e · inbound

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion cites this paper.

PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T11:36:17.191697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:36:17.191697Z digest=sha256:21e8cb6d8d828171b9cb650bf8abe05b9924acabd36716d1efab8a4ca233ac48

Observation e61fc56c-1aac-4393-9111-0fc58d5d9837 · inbound

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts cites this paper.

T2I-FactualBench: Benchmarking the Factuality of Text-to-Image Models with Knowledge-Intensive Concepts VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:02:43.325540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-23T08:00:12.781392Z digest=sha256:a4e8a6608b96f44d9b544a96b01bf8f232e86bdb79c91fb1ad528592726f0829

Observation 7501a7a5-02b6-47e7-8b49-5ce6e3ca7999 · inbound

EvalGIM: A Library for Evaluating Generative Image Models cites this paper.

EvalGIM: A Library for Evaluating Generative Image Models VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T15:50:16.739684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:50:16.739684Z digest=sha256:49b3f10fa11dade20b2526c60656796a146a9d2bbd900d8b56b5bf3fd102f4fa

Observation acccb5c3-1323-4519-a148-a8eaf814951d · inbound

DreamFit: Garment-Centric Human Generation via a Lightweight Anything-Dressing Encoder cites this paper.

DreamFit: Garment-Centric Human Generation via a Lightweight Anything-Dressing Encoder VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T05:23:02.048441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:23:02.048441Z digest=sha256:691108740fe30d2090df5a95e2fdf4ec743845b912b9ea0c742c13d8739b668c

Observation 15ba3053-9b20-4b7c-bad6-930cb533feee · inbound

Explainability for Vision Foundation Models: A Survey cites this paper.

Explainability for Vision Foundation Models: A Survey VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 161

Resolution
unresolved
no resolver link, observed 2026-08-10T17:26:35.594156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:26:35.594156Z digest=sha256:4bb52001a60229fcec12b6ac870afbf7ca6b9d766b5c187dd87097748aa64558

Observation 67f98038-984a-4cae-804d-c04f74b6cff3 · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:36:41.668122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:8cc627e016e71753995e9e29442b91b604aa8026ae02b652e6b08a441c53bf0c

Observation f3732ed8-7007-4827-86dc-cf4d00058154 · inbound

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer cites this paper.

In-Context Edit: Enabling Instructional Image Editing with In-Context Generation in Large Scale Diffusion Transformer VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T16:07:53.134117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T16:07:53.054355Z digest=sha256:3e536d31c540675c6d23cdc7542969c926774b8b22872aff72e44b2b02f0c042

Observation b977744d-a7d0-4741-bb3a-c694fa0650ab · inbound

Multi-Modal Language Models as Text-to-Image Model Evaluators cites this paper.

Multi-Modal Language Models as Text-to-Image Model Evaluators VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T04:43:11.291929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:43:11.291929Z digest=sha256:e088f9d0e984f0e6db546286fbcba0c7be384f868fffc7ffd2b0d30bc471c7ed

Observation 5d592bcc-2256-48dd-a416-eaf99345832e · inbound

X-Edit: Detecting and Localizing Edits in Images Altered by Text-Guided Diffusion Models cites this paper.

X-Edit: Detecting and Localizing Edits in Images Altered by Text-Guided Diffusion Models VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:52:32.826546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:52:32.826546Z digest=sha256:1e66011f3b7f240962f02e13b95cf95d4c01fadb42a16d9e5ff111246bcac76f

Observation 1e5abc39-9021-49d6-8b63-15ce95ed4fee · inbound

Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation cites this paper.

Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:16.408545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:16.408545Z digest=sha256:39dd13a4d8271ac3f398d3df1e4d88550e4378a1ba7ac1dab9a45a6a12f1d871

Observation 9d4ebb82-c564-4ef3-968e-ea2e5158af25 · inbound

ImgEdit: A Unified Image Editing Dataset and Benchmark cites this paper.

ImgEdit: A Unified Image Editing Dataset and Benchmark VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:17:45.259977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T18:17:45.123690Z digest=sha256:87b2e972752be5f45364160e90fbb8c0154eb013b23731972bcacba41bba9a68

Observation ad65a177-2452-4707-8dac-e53ac9c1a002 · inbound

R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation cites this paper.

R2I-Bench: Benchmarking Reasoning-Driven Text-to-Image Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:48:52.532034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:48:52.532034Z digest=sha256:8af6720ab950c0bb501e479dfcfeeae0645fabe8db254e0274dd92c01d2cce1c

Observation 080b9fb2-f8da-4dbb-87e7-17dcb316a9b4 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 149

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.739032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.739032Z digest=sha256:180ee0fe6aa873a6ed84a9418889d3869d5e327443779f372235c439f2b7708c

Observation b2226db6-c3d0-4fa5-80af-3cf959513b83 · inbound

Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing cites this paper.

Balancing Preservation and Modification: A Region and Semantic Aware Metric for Instruction-Based Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T00:43:14.231201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:43:14.231201Z digest=sha256:4b5e9a299d6343fcf0741657139e84a219b6536ff08b1c55c8a5b1fc186b3d13

Observation d23d590f-3542-48ac-8cf6-b63796687287 · inbound

JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent cites this paper.

JarvisArt: Liberating Human Artistic Creativity via an Intelligent Photo Retouching Agent VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T19:09:17.992169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:09:17.992169Z digest=sha256:f40af1132bbbad4fdb0acf2ed09884ef335ac0732b6a579f0db4e7fdf115bd9a

Observation ef2fe87d-4d27-4650-80b5-8d47954eb964 · inbound

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation cites this paper.

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:01.313709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:03:01.313709Z digest=sha256:c02fc304c263459e3773b0e27061cb287c387c3245a438fae52a3afc2594cc69

Observation a3bd7ea0-470b-434b-9038-2b132119e827 · inbound

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation cites this paper.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:40.925409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:40.925409Z digest=sha256:2eebf08fce57ce1fdde97c4bd55dc24088499fb97202ecca10bdd7bc87d7ca7f

Observation 4bf251e1-f89a-4031-8745-a2ac7ded0574 · inbound

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation cites this paper.

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T12:54:34.950914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:54:34.950914Z digest=sha256:52b221de32422a8febbc34f2c6179f33a032b8176b5b499a8567bb825c5b188d

Observation 2e70bba0-9261-4dd3-a136-34ae5d070c2e · inbound

WebMMU: A Benchmark for Multimodal Multilingual Website Understanding and Code Generation cites this paper.

WebMMU: A Benchmark for Multimodal Multilingual Website Understanding and Code Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T17:16:05.053807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:16:05.053807Z digest=sha256:052e4bf28f61c112307c03e9bebf5a2836f7703110d02a1bd92876b53730602e

Observation 17025fd9-d7d8-4042-afa8-5a8b61467602 · inbound

Edit in 2D, Verify in 3D: Reinforcement Learning for Multi-view Consistent Scene Editing cites this paper.

Edit in 2D, Verify in 3D: Reinforcement Learning for Multi-view Consistent Scene Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T19:14:07.671722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:14:07.671722Z digest=sha256:984b3da09f87b82d3a874b3fa91d658090f1300efb5d79ecb8eaf677bf6dde3e

Observation d48f321e-3acf-45a1-943f-faf5e2b1b736 · inbound

Meta-CoT: Enhancing Granularity and Generalization in Image Editing cites this paper.

Meta-CoT: Enhancing Granularity and Generalization in Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:41:19.200450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:30:28.636915Z digest=sha256:61d29de5a90b852e1d46b4a414832ce7b2b7c5d9990b07f31b11cce497f18a45

Observation 70bc142b-cbb0-4b3b-946a-33184d4f70d9 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T08:39:53.770233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T08:36:27.676888Z digest=sha256:4eebb43b23fd904735b222f4d047903340ae33578efe3bb84433d81f5b34e116

Observation 1d5ead44-912d-415a-9ff1-b58cfc52d902 · inbound

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison cites this paper.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.344776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:8b751c94544ed40be28e4138bfdd5f9ff069b140abf9f63abc183c95b4a39029

Observation eef3ef19-e9b8-46ed-9519-061dc413ac1e · inbound

ProductWebGen: Benchmarking Multimodal Product Webpage Generation cites this paper.

ProductWebGen: Benchmarking Multimodal Product Webpage Generation VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:14.716074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T17:26:44.297446Z digest=sha256:eac77a049dd3c15bba911bd437ee08dbce302428231d3b7075d8686fd7912879

Observation d1942f95-0450-444b-8a74-468e590a48e0 · inbound

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching cites this paper.

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:46:18.517174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-28T15:07:29.897089Z digest=sha256:ce9de9a17cea4f490270abe22de678255f9497285e71c9fcdef2dd7005cdf877

Observation eaf774bd-b86d-43b4-ad0e-8fe17f5fe678 · inbound

SpatialFlow-GRPO: Where Spatial Credit Drives Image Editing cites this paper.

SpatialFlow-GRPO: Where Spatial Credit Drives Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:52.752437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T05:37:23.569942Z digest=sha256:d281ac63b9eeda94ff08ca223e298e9f873d68329e9c8d94c309640b3bcdf0c3

Observation bd3aabfe-71db-4b69-8496-acb03859ecd1 · inbound

SpatialFlow-GRPO: Where Spatial Credit Drives Image Editing cites this paper.

SpatialFlow-GRPO: Where Spatial Credit Drives Image Editing VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:33:51.270522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T05:03:56.872939Z digest=sha256:66dbb9c6cc06778eef86b622647380ac1d63f320fe294fc50c19a52de4c5dfec