Pith. sign in

Paper Citation Record · LEDGER

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation

As of 8 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2607.14962.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.14962 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T00:41:12.605274Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 569ee3a7-2bf9-46e0-b837-c33b7b552229 · outbound

This paper cites Consistency-diversity-realism Pareto fronts of conditional image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Consistency-diversity-realism Pareto fronts of conditional image generative models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.330364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.330364Z digest=sha256:dd7cffa15302d2c6f6906ab47ff3c335e75003f62c8698c97bade20d4b2c74c3

Observation 7d26e924-7fdd-4fd7-946e-d04ebfb37e24 · outbound

This paper cites The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.396092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.396092Z digest=sha256:7be9e46fcd634ae6f574e7c3d37cfb2602ff70afd2eabda5521180318a5a8234

Observation 6dbbaf3f-abfe-4fd6-8cd8-dd72242ce49b · outbound

This paper cites Vector Policy Optimization: Training for Diversity Improves Test-Time Search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Vector Policy Optimization: Training for Diversity Improves Test-Time Search

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.402933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.402933Z digest=sha256:fca2d51f3f7d9e2bf8aa41aef713fe1ea17ba0ab19e73d7d04bfe6f93cb996d4

Observation b7311ce4-cb29-4b31-8e50-63955e6b841a · outbound

This paper cites Qwen2.5-VL Technical Report.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.436642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.436642Z digest=sha256:10a0841e24ddda99b36718664cf50909df25dc94436838053f983e2f28a42449

Observation badcd5a8-bc19-48dd-9688-7b7344c73393 · outbound

This paper cites Easily accessible text-to-image generation amplifies demographic stereotypes at large scale.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Easily accessible text-to-image generation amplifies demographic stereotypes at large scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.451300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.451300Z digest=sha256:1da74f96af494e1195605fbafbedcbd7ef97eef5d0218c10499fd1e77e3277ca

Observation 62f855ec-16f5-4735-a3ce-57e646de96c9 · outbound

This paper cites Training diffusion models with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models with reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.465729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.465729Z digest=sha256:7a4468632b0fd518de6d7add10f4b37c88bf20b0d7215c8a43a09e4d0a34a4d8

Observation 0c20ee32-9ca1-490f-8f1b-0322f29db8f7 · outbound

This paper cites SEGA: Instructing text-to-image models using semantic guidance.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SEGA: Instructing text-to-image models using semantic guidance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.468740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.468740Z digest=sha256:04e8dcd253a22bba78a1148b0c49b2b286f25fa980dc1b39b69f7cdd557a27db

Observation 3cbf7845-bfda-463a-9ac2-a873e4e1218e · outbound

This paper cites HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.471192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.471192Z digest=sha256:8338afff8d3596ede744831d6dc23698042dd8f4a34fcc399142344e8544efc9

Observation eed9be49-730e-4f4f-8e26-a1da0f51c7d8 · outbound

This paper cites TIBET: Identifying and evaluating biases in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TIBET: Identifying and evaluating biases in text-to-image generative models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.473944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.473944Z digest=sha256:81456ce0f785c820f4cd26b800509d508eef6563ddd36ea78ce2b0f73ce6a96f

Observation fd0cc5f1-d2aa-4b60-a01b-b2eb0b9f9444 · outbound

This paper cites Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.476286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.476286Z digest=sha256:68187fed19eddfaaefbb0317ba2bad325da86e6e91a7a654c278f0167ae2d942

Observation c7165b64-9106-4070-9251-e8a30c3b829e · outbound

This paper cites Debiasing Vision-Language Models via Biased Prompts.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Vision-Language Models via Biased Prompts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.478856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.478856Z digest=sha256:b427f31fdb4e31e991779482979bb002050b7fd0bf4198ab7f2afce1735c0d80

Observation 38b569cd-0078-44c2-a6e6-4a6d8abe8474 · outbound

This paper cites OpenBias: Open-set bias detection in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OpenBias: Open-set bias detection in text-to-image generative models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.481670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.481670Z digest=sha256:a7c5fcfeeca99e5d5000dfb5c95ae6201efbb7b7089729b46321158ba1d2539f

Observation 90c5914b-b939-4393-bd7c-e4b684633e51 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.483879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.483879Z digest=sha256:ba3c32c8862f39961c85761a8b0223b141821b795bcaa529bf1b3842d25e2c32

Observation ddb27611-5b72-4a2b-bac2-897d4009c88d · outbound

This paper cites DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.486268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.486268Z digest=sha256:d654d713e535abbf1b29ca94843f106f811ab72e1f87bd7698ce1fc7b04c17a8

Observation 3c9651ac-24b0-4253-bc3a-94c29f1f306b · outbound

This paper cites Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.488625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.488625Z digest=sha256:bf4d91c89c4eba798792811323853ba896b43cb9afea772cb1f6d80158c2fd20

Observation da91c3c7-2659-4af8-91f4-b6c97dc0cf8b · outbound

This paper cites FairImagen: Post- processing for bias mitigation in text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairImagen: Post- processing for bias mitigation in text-to-image models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.491290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.491290Z digest=sha256:29dff0bcbb4a968b0c7ae840d9e7818ac2107798c09d9a11c72ba634937e5264

Observation f7b880c9-7001-4b37-bcca-39b6c3193b86 · outbound

This paper cites Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.493826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.493826Z digest=sha256:a86c2ba7cbfec93003d35335a15c9a0de81d03425a480e797ede6498ee909c45

Observation 3149b138-104b-4a52-b9f1-cbb0718bc99b · outbound

This paper cites Unified concept editing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unified concept editing in diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.496320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.496320Z digest=sha256:203ecf399c4c87ccccf021d0368608584885e6bae85d1291f212f311cda57b70

Observation f00ee3cd-ecd5-47ae-b844-71a294de6562 · outbound

This paper cites Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.498785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.498785Z digest=sha256:a8c48425483cd4c47ad528e4b5aaf902b9f483f6d96e1fb32aea4705c068f187

Observation 020db60b-35ba-45bd-9967-7daf77b9f646 · outbound

This paper cites Polychromic objectives for reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Polychromic objectives for reinforcement learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.501274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.501274Z digest=sha256:8bfaf1f52597fb2bfcdb565f6d4cd1df3808719101f306d7d1a3b333428e57e6

Observation 5aa290b9-9b70-42dd-8d0d-7e8220f21cef · outbound

This paper cites TempFlow-GRPO: When Timing Matters for GRPO in Flow Models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.503578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.503578Z digest=sha256:8d0d07c1fb31db0b8fe7652112ae3f7a047044a59d66620ea66f800761d35e77

Observation 49fafe22-6b62-49ed-a2cc-2da025729424 · outbound

This paper cites Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.506040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.506040Z digest=sha256:93d4b1b2adecd79a6fdea16dd81c0df5ff87910ef50747f33b447b3ae00d7ae0

Observation 1e78e774-80aa-4405-836c-8059598e90bd · outbound

This paper cites FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.508514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.508514Z digest=sha256:c2d4ba0ada09a1f4f72c04c85bd58330c6122c2af5a7143c86b91cd33078e3f8

Observation 78838bea-cd1e-4f7d-aa32-9eb7400ea55b · outbound

This paper cites FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.510955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.510955Z digest=sha256:423801145c441f3b5872464e90679a9b63bbd03007c8361260855fb6f74bb695

Observation 1b98a6db-d775-4a7b-b363-37f8757dd418 · outbound

This paper cites Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.513219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.513219Z digest=sha256:e71fd1b17112d48d66d78843601f0d12abdc79f1c9797adeb99552d5db39119c

Observation ecae609f-eeff-4e44-98df-fb10d3eaec18 · outbound

This paper cites Pick-a-Pic: An open dataset of user preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pick-a-Pic: An open dataset of user preferences for text-to-image generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.515488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.515488Z digest=sha256:112f41acc622abd54ca95d0d5118b3d702a67fa28a5779856eea6a0a09594f23

Observation 95189d68-7248-40db-8a92-200f1f05fc54 · outbound

This paper cites Emergence of exploration in policy gradient reinforcement learning via resetting.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of exploration in policy gradient reinforcement learning via resetting

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.517916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.517916Z digest=sha256:606343e9bd0991de04d8be95b82677e62fa223c3721c5691cf305ea88e952ed9

Observation 5e72bcb7-1640-4ad8-a53a-87498f3da14f · outbound

This paper cites Improved precision and recall metric for assessing generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improved precision and recall metric for assessing generative models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.520460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.520460Z digest=sha256:be5caadf53ad368da20d6f2005713bfdfeb26f595b6d821be70e0b52e167adfb

Observation 1ab7d50b-32f1-4b0f-8792-8d377c397003 · outbound

This paper cites Holistic evaluation of text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Holistic evaluation of text-to-image models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.522754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.522754Z digest=sha256:36054140bcac60de74c12752d605fea4210ec2c2c2ae41aee2195ce1f1e878a1

Observation d410701a-f8da-4728-aa45-a184037b381a · outbound

This paper cites SetPO: Set-level policy optimization for diversity-preserving LLM reasoning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SetPO: Set-level policy optimization for diversity-preserving LLM reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.525144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.525144Z digest=sha256:a7df7cca7a58dd8fb31445cc2f96ca7019a54a920aacb13304f7e6a604fdce8a

Observation 4b7add92-3928-48b7-890d-161f66c19520 · outbound

This paper cites Fair text-to-image diffusion via fair mapping.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair text-to-image diffusion via fair mapping

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.527336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.527336Z digest=sha256:b87d7959df63a38211949e86e53c42abcf69fa3d9250e604b1e8c525bdd2d88b

Observation a964ee23-6e0a-4e5b-be4b-63a9f1fa8859 · outbound

This paper cites DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.529558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.529558Z digest=sha256:3b9df4ff4e011ca5e9cedcae2f51a6979cd0f5a74fa3e0197d810e34d9f1e9de

Observation 8b575eda-5613-4e9b-9eb6-c45d16613aca · outbound

This paper cites Flow-GRPO: Training flow matching models via online RL.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Flow-GRPO: Training flow matching models via online RL

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.531859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.531859Z digest=sha256:abc623c601fc78f3aa8f0896eec71b81b7bc8b3b587d38423228f56c04a7b1bb

Observation c4d889dd-e52b-4b4e-82db-d6bc56dbf0f7 · outbound

This paper cites Improving video generation with human feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving video generation with human feedback

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.534200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.534200Z digest=sha256:aec63a2cde7e9459062c8dcb38660642fc7be4c1e6137427021ece70311bf5b7

Observation b613b1bf-43e1-47fc-92aa-7b2f4673c058 · outbound

This paper cites Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.536646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.536646Z digest=sha256:3cc329e50661d956863d415d7a2ce0d661b3104645291e4dde6c46f1e6870496

Observation 8c8ec63c-b297-425d-a09e-59593ac35a11 · outbound

This paper cites Stable bias: Evaluating societal representations in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Stable bias: Evaluating societal representations in diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.538932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.538932Z digest=sha256:cc77b3449540ad4c61794e2508aed6380ed51ab0c2da6416ae0e3d8381afa5a0

Observation 5628b4ff-29c0-46f2-a9ea-7e830cefaa8b · outbound

This paper cites Training diffusion models towards diverse image generation with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models towards diverse image generation with reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.541267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.541267Z digest=sha256:ccad5b696b3de2b393fea6d9c435703e4d7d437fc62956695d514887de68bfe3

Observation 956c240d-ac93-40ca-9bfb-6c9d3bcb2811 · outbound

This paper cites Reliable fidelity and diversity metrics for generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Reliable fidelity and diversity metrics for generative models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.543666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.543666Z digest=sha256:49806264243a4c4368dc078b6e015f4a312d16d545a0527a2c98f1936cca502d

Observation 98fbfd37-543b-4703-916f-ca483b3488e8 · outbound

This paper cites Retry Policy Gradients in Continuous Action Spaces.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Retry Policy Gradients in Continuous Action Spaces

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.545963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.545963Z digest=sha256:8de0116f9789ef3041f92bb62d873b6fbdbb1c2626e48a44df2ec501f3a381f2

Observation 3f06b6f3-348d-4473-9602-e638dd50a9ab · outbound

This paper cites Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.548554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.548554Z digest=sha256:740fc8bdcd8df50f807c7a1c599070142458a35ffbc0fb4d6114c55c60dc4cc6

Observation ce1d8d97-b08d-4084-b613-8308d7161789 · outbound

This paper cites Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese).

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese)

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.551264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.551264Z digest=sha256:d497a2af9533813c05649b92763f11d9f272aafaa7fdd947eeb253958716bba8

Observation bd6a6272-880d-4332-a414-dc436b678a0d · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-time text-to-video alignment with diffusion latent beam search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.553646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.553646Z digest=sha256:5fecdf5744cf56ede48d504ac88bf0f248802c7ce46b78fba6301c1b10e92ab4

Observation 52efdac7-f5e6-4d8e-aa48-766f4ba67354 · outbound

This paper cites MultiBanana: A challenging benchmark for multi-reference text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation MultiBanana: A challenging benchmark for multi-reference text-to-image generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.556171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.556171Z digest=sha256:fdf2706ddabe799d6e1f51a5747be8f3de2cbd88ed53a16613f480ca7ba7b5f2

Observation 5596e674-6524-4b44-8d7b-15cc3fd2caea · outbound

This paper cites Balancing act: Distribution-guided debiasing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Balancing act: Distribution-guided debiasing in diffusion models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.558556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.558556Z digest=sha256:5120a5b51eb6569ae8ef090080576d9b568c4b15fc5b53fdf66d3427ee10dbe2

Observation 61c207e2-3e47-44ae-a1b7-dae906663094 · outbound

This paper cites OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.561057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.561057Z digest=sha256:98d9387987e67710544a8b893daa1c08200e7c02067643ad63c508f72044be56

Observation 2ee8dd5a-d3e2-41e9-9c6d-493c763a4fd0 · outbound

This paper cites Escaping the mode: Multi-answer reinforcement learning in LMs.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Escaping the mode: Multi-answer reinforcement learning in LMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.563557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.563557Z digest=sha256:7b16e907a4273ad69040962d20e887187cd003278bf935a1f0b954e655394b26

Observation a407128f-6f22-4bdb-9148-3ceeb21bc320 · outbound

This paper cites From scale to speed: Adaptive test-time scaling for image editing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation From scale to speed: Adaptive test-time scaling for image editing

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.565970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.565970Z digest=sha256:6fe57b00a227d84b7e80bc867c99c1e364141aa34c31df9c0c6bc29ec8f3cd2e

Observation 29f0f96d-26a1-4413-840b-acc6c83eb4ff · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Learning transferable visual models from natural language supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.568234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.568234Z digest=sha256:64184dc72dfb3017177078547ef75341980d54a3f04fb4c499d146a9286ee3f7

Observation 21b08dd3-e390-4521-b26c-6613d2fe8f62 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.570492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.570492Z digest=sha256:c67d8fa1ed455fd347031278fdbd5d800818f76787f9dfa758f82acbd0f885a4

Observation 9b3a6574-b01c-40cc-8c6b-79122fbcf827 · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.572904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.572904Z digest=sha256:0ce4a23be7930bbba369a1dc21fd7243280c41abf0f86eeb180d328bf7ba652b

Observation 9f10100a-1356-40e4-9979-57777235438c · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.575402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.575402Z digest=sha256:fb8411ec6887aa643bc638353c835e5481cc514dd0d7d1e681f70eda81dd3cfc

Observation f4a61204-a395-43f6-9f89-da15f0ea4b5f · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.577708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.577708Z digest=sha256:24dd623008d16eea6b9b4bbcadaaca5ec3532f710ec00a416220074490c1bb1c

Observation 16d91b2e-c927-4d36-8ddc-dc6809a22eff · outbound

This paper cites Finetuning text-to-image diffusion models for fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finetuning text-to-image diffusion models for fairness

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.579983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.579983Z digest=sha256:984824ae2e5ce0cf1be185c69fa07680081267984c5d8043f27c7de2507c9bcd

Observation d3a013a1-75d5-4f88-a672-5f78cae46762 · outbound

This paper cites On Advantage Estimates for Max@K Policy Gradients.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation On Advantage Estimates for Max@K Policy Gradients

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.582293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.582293Z digest=sha256:0e6d049fa57d917c06aa9492f9ed43fc41c57f52902106bed1c4dd957765c028

Observation 90ddcec2-26b5-47db-b4da-ea82bb118978 · outbound

This paper cites Finite-Time Regret Analysis of Retry-Aware Bandits.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finite-Time Regret Analysis of Retry-Aware Bandits

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.584759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.584759Z digest=sha256:76eeaae661118f78ea388e022f90fa24c5e85ecad88baeb6f3cb8f209656e471

Observation 18167a3c-1c8f-4e50-9ed5-81008fded9d4 · outbound

This paper cites Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.587281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.587281Z digest=sha256:0add74d535fb4d3a750abe30c20714f8bf10d32a4ffdb3019825615b47801923

Observation 1a402d90-111e-4717-95c6-8f29f4a440f4 · outbound

This paper cites Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.589526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.589526Z digest=sha256:d961e5ab5135161346c6bfd1883e2632dfab88bee3d5ec4ce5aa9663fb452309

Observation 39af284b-761a-4af8-84d1-f12370095706 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diffusion model alignment using direct preference optimization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.591916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.591916Z digest=sha256:95504b111a68195026fab865a37cbced76592902f8c8564e2040454a6afb60e6

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · outbound

This paper cites RewardDance: Reward Scaling in Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:1bff51c6a011fbfaef4d827ec4c98330963b52ab983ae83273cc71325585fffa

Observation e68a2686-f2c6-4ee4-b14c-b9d9d13acff1 · outbound

This paper cites ImageReward: Learning and evaluating human preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ImageReward: Learning and evaluating human preferences for text-to-image generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.597785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.597785Z digest=sha256:97fa43c347fe486290cda31f1e8d3026087f4b34718622e2bd8aa1208aebe6c3

Observation 7a91fd58-a628-4ebf-ba52-da95886e125f · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.600146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.600146Z digest=sha256:7bd2ac684e27c0a5a62850093f7a764bb1883d8d2c527d7cf6c69566ad398ec2

Observation 1391ca56-142d-400f-a972-89dd5b4d7b44 · outbound

This paper cites ITI-GEN: Inclusive text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ITI-GEN: Inclusive text-to-image generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.602630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.602630Z digest=sha256:e7dbbab13928ee0d2c44d3b68a55a91a2e4a543cbcdaf7764557f02c773286f6

Observation 08f4c5d8-6bb3-4181-94f6-4d654eb75c57 · outbound

This paper cites a photo of the face of a person.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation a photo of the face of a person

Reference 63

Resolution
malformed identifier
no resolver link, observed 2026-08-02T00:41:12.605274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.605274Z digest=sha256:0056b551eb265ab79c927643ced2477f3d4ae0efe5865b422597a9e7c06366e8

Pith citing papers

No inbound Pith citation observations are available.