Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Spatial Relationships in Text-to-Image Generation

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2212.10015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.10015 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:29:14.923026Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:29.849367Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c9468441-2afa-4be0-b5c9-cf182aa22214 · inbound

Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation cites this paper.

Align Beyond Prompts: Evaluating World Knowledge Alignment in Text-to-Image Generation Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:29:14.923026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:29:14.923026Z digest=sha256:75e93e849937b0b73ca5d0f0b115d1483dbb6f770c87a37913af0eebd8b33d2c

Observation fbaee488-b011-4168-9d0d-9bf5f8ac655b · inbound

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models cites this paper.

Re-Thinking the Automatic Evaluation of Image-Text Alignment in Text-to-Image Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:41.429534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:41.429534Z digest=sha256:4acea0874ddc355bb44d76c7e65db8f33fd9139f9b40a7648d31a832cfbd6515

Observation d00ac191-626e-49e9-bd97-7f9e6761377d · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:46.511922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:46.511922Z digest=sha256:7b1abc09ae2b8fe1ded16f0a3cfad7af28d28fda5e743119b27955a46038a011

Observation 54e5f73d-cf88-4460-bc27-fa90af3275fe · inbound

Why Settle for Mid: A Probabilistic Viewpoint to Spatial Relationship Alignment in Text-to-image Models cites this paper.

Why Settle for Mid: A Probabilistic Viewpoint to Spatial Relationship Alignment in Text-to-image Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T21:48:52.391265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:48:52.391265Z digest=sha256:60e5b07c9fa3bc4696a1721091e09be291d8d89273e7374303120f71d56c9097

Observation 8dda9b9c-2267-4826-8203-8d08ffa3620f · inbound

Towards Evaluating Robustness of Prompt Adherence in Text to Image Models cites this paper.

Towards Evaluating Robustness of Prompt Adherence in Text to Image Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:50:42.353509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:50:42.353509Z digest=sha256:45420029a7c6e837e099b72e69d8cdb054ebe39df7950864c47320dfb103c883

Observation 9d0e6edc-0bf6-4db3-bc1a-0f0fca961dd1 · inbound

Canvas3D: Empowering Precise Spatial Control for Image Generation with Constraints from a 3D Virtual Canvas cites this paper.

Canvas3D: Empowering Precise Spatial Control for Image Generation with Constraints from a 3D Virtual Canvas Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T22:23:09.608412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:23:09.608412Z digest=sha256:07d9902191f4a61df02814471a717f01f543e6306e7b5aca1f10821a260d8c1d

Observation a5df3ae9-484b-47a3-9946-033742454da1 · inbound

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation cites this paper.

Human Preference-Aligned Concept Customization Benchmark via Decomposed Evaluation Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T10:59:25.229741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:59:25.229741Z digest=sha256:65e4788e6aa1e889e1509a3ddf785ee1aa874b70ea30d3b6acf6d54005125e29

Observation 4cba0c9c-9294-49d7-96a7-d81c09a42921 · inbound

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration cites this paper.

CARINOX: Inference-time Scaling with Category-Aware Reward-based Initial Noise Optimization and Exploration Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:26:33.618048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T15:25:39.116881Z digest=sha256:cb6a39b893709779b9ca44cb3a45d28125dd18ea4c4c90c6e968e68f5a55f984

Observation 4fbf225c-09f3-4334-a9f3-ca043d6b7fa5 · inbound

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs cites this paper.

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T09:35:27.378827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T09:35:27.378827Z digest=sha256:1bfeeaf68c5359b440e32663ec99f0db83a198cdbea0b2b9590961da63e7f8d0

Observation da2973fe-c940-41b8-a44c-8499cff71122 · inbound

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk? cites this paper.

Exploring the AI Obedience: Why is Generating a Pure Color Image Harder than CyberPunk? Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:16:31.387201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T19:15:11.575505Z digest=sha256:7015e60a2c1d6db0966fb9ea90d47d3d09e04d3c09619092bea2324548185286

Observation 63def0fd-170a-45bf-a37f-3b9949a7e061 · inbound

Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification cites this paper.

Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:56:47.397746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T06:55:33.464951Z digest=sha256:bf416c3013259068d1097902ecb64c483d9e85e77dd9929d68b8a56ac1749007

Observation 9a3cbfd8-8b21-4830-ad48-68a08784bdde · inbound

Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification cites this paper.

Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T19:10:41.883296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:10:41.883296Z digest=sha256:bd61749283d2c057ccc07c1d844855997515c87c2e76dd52268210135613ef90

Observation c5b2d4fb-eb60-40ee-838d-538e2cb0d628 · inbound

EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation cites this paper.

EPIC: Efficient Predicate-Guided Inference-Time Control for Compositional Text-to-Image Generation Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:12:22.668531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T06:10:52.567634Z digest=sha256:298ee405aec7370fa4b4de25bd16b3b282c9e3637fc277b768084b741d29ef3b

Observation 300a6ef3-f683-48d3-a286-9aa5671031eb · inbound

Boosting Text-to-Image Diffusion Models via Core Token Attention-Based Seed Selection cites this paper.

Boosting Text-to-Image Diffusion Models via Core Token Attention-Based Seed Selection Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:28:05.540179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T06:25:32.694239Z digest=sha256:d77785b94ad396263516d7cd9fc6604b50c86239b06b5fe474caacb0f9ee5417

Observation db924f41-522a-4d10-aa6d-5074d461b1f2 · inbound

A Systematic Study of Behavioral Cloning for Scientific Data Annotation cites this paper.

A Systematic Study of Behavioral Cloning for Scientific Data Annotation Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 268

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T16:23:39.188094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T16:23:08.402194Z digest=sha256:1e2b279426fab72fb6cca6ef7e1579b0f8892052bee7a2df9e7e41675254d7ca

Observation b3361d10-8a55-4354-b500-ec4d341334e8 · inbound

Compositionality Emerges in a Narrow Depth-Connectivity Regime: Architecture Constraints and Solution Manifolds cites this paper.

Compositionality Emerges in a Narrow Depth-Connectivity Regime: Architecture Constraints and Solution Manifolds Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:09:29.854512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T18:22:56.676469Z digest=sha256:a4e8eb04588c95837fcd5d90b2a65b9fe02aaa009ce48463d5dd7764ad984c94

Observation 3e74535b-719e-49d1-a2a2-c3d88c0dfca8 · inbound

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion cites this paper.

ELDiff: When Evidential Learning Meets Text-to-Image Diffusion Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:49:29.850942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:40:14.587143Z digest=sha256:5611df55234e023c1e53e6a1605b3c0e91053dda424d3376cd0db11d0e773504

Observation 36b8fd0f-250a-427f-9c62-25dba55a608d · inbound

Transferability Between Understanding and Generation in Unified Multimodal Models cites this paper.

Transferability Between Understanding and Generation in Unified Multimodal Models Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T19:17:19.634242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T19:17:19.634242Z digest=sha256:1da77fea7c11c7d125be7e5e7241c089dbeb215446646fba5bc2e7e1cdbd1c4a

Observation d397bc5e-c954-44b6-a872-c5f290657a1a · inbound

Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text cites this paper.

Show, Don't Tell: Evaluating Spatial Cognition in Generative Pixels Rather Than LLM Text Benchmarking Spatial Relationships in Text-to-Image Generation

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T08:39:37.518805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:39:37.518805Z digest=sha256:430a7c7032b73f4189823d5f62dbb8b2569af20a8ce7bca00ce924bd44cff342