Pith. sign in

Paper Citation Record · LEDGER

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation

As of 20 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2607.14962.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.14962 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T00:41:12.605274Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved62
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 569ee3a7-2bf9-46e0-b837-c33b7b552229 · outbound

This paper cites Consistency-diversity-realism Pareto fronts of conditional image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Consistency-diversity-realism Pareto fronts of conditional image generative models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.330364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.330364Z digest=sha256:27998cb2da9c3ea5bd1b18a9aa8b4d5a813d7cc8847b5d484ec5f5f1a46fdf3b

Observation 7d26e924-7fdd-4fd7-946e-d04ebfb37e24 · outbound

This paper cites The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation The best of N worlds: Aligning reinforcement learning with best-of-N sampling via max@k optimisation.arXiv preprint arXiv:2510.23393, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.396092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.396092Z digest=sha256:4493d3055a9fa0919662e7356b5068e7e4802a812e01d8ced4c20a4c5e432b0c

Observation 6dbbaf3f-abfe-4fd6-8cd8-dd72242ce49b · outbound

This paper cites Vector Policy Optimization: Training for Diversity Improves Test-Time Search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Vector Policy Optimization: Training for Diversity Improves Test-Time Search

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.402933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.402933Z digest=sha256:d84849c1f1f821105e62eb04f82a689c8428090c6e5afbf11da38e9e6d8e24f8

Observation b7311ce4-cb29-4b31-8e50-63955e6b841a · outbound

This paper cites Qwen2.5-VL Technical Report.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Qwen2.5-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.436642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.436642Z digest=sha256:5b6ff5415badd4bc18dcc240be94364f1d007ad7a08328a5aba7936ae9e12589

Observation badcd5a8-bc19-48dd-9688-7b7344c73393 · outbound

This paper cites Easily accessible text-to-image generation amplifies demographic stereotypes at large scale.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Easily accessible text-to-image generation amplifies demographic stereotypes at large scale

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.451300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.451300Z digest=sha256:4f3ac3bc284093fad190e230ce77e9d8518e535f1e52fca983b133f21bd3720b

Observation 62f855ec-16f5-4735-a3ce-57e646de96c9 · outbound

This paper cites Training diffusion models with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models with reinforcement learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.465729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.465729Z digest=sha256:7c3b4d7485774cdd9ca9b07a052d7db7d2b07219495888c1069b5a944fd43b23

Observation 0c20ee32-9ca1-490f-8f1b-0322f29db8f7 · outbound

This paper cites SEGA: Instructing text-to-image models using semantic guidance.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SEGA: Instructing text-to-image models using semantic guidance

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.468740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.468740Z digest=sha256:0d9f66b6f64562699f121a105df79166793ec72964b80b796736f53af88cd649

Observation 3cbf7845-bfda-463a-9ac2-a873e4e1218e · outbound

This paper cites HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation HoloFair: Unified T2I Fairness Evaluation and Fair-GRPO Debiasing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.471192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.471192Z digest=sha256:f9839d753e270772c14b33cc2378f2d41d170b4e76443cdae92b994d9502d33f

Observation eed9be49-730e-4f4f-8e26-a1da0f51c7d8 · outbound

This paper cites TIBET: Identifying and evaluating biases in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TIBET: Identifying and evaluating biases in text-to-image generative models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.473944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.473944Z digest=sha256:323fe6522596bbdd7bba71becb4d6441cc6b4262ae57d5a18ee5cab11915528b

Observation fd0cc5f1-d2aa-4b60-a01b-b2eb0b9f9444 · outbound

This paper cites Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-aware fine-tuning for best-of-n sampling in large language models.arXiv preprint arXiv:2412.15287, 2024

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.476286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.476286Z digest=sha256:2409b0b420024da36e44ea08acf3e46da06c8f3edc06cf7f5e22bb93a518b704

Observation c7165b64-9106-4070-9251-e8a30c3b829e · outbound

This paper cites Debiasing Vision-Language Models via Biased Prompts.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Vision-Language Models via Biased Prompts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.478856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.478856Z digest=sha256:e780d2a2df94566a854d9d2721ade72cee4065076e299ba4045a8295eadee6e2

Observation 38b569cd-0078-44c2-a6e6-4a6d8abe8474 · outbound

This paper cites OpenBias: Open-set bias detection in text-to-image generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OpenBias: Open-set bias detection in text-to-image generative models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.481670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.481670Z digest=sha256:adc311247c3eacd8966fd45f88bce36d45cb25ba79a6452e4644dfd68ee10a43

Observation 90c5914b-b939-4393-bd7c-e4b684633e51 · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.483879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.483879Z digest=sha256:af46bbe9a286a98acd2884fb64d546f445d5e614717f2264ea388515c14ffdd4

Observation ddb27611-5b72-4a2b-bac2-897d4009c88d · outbound

This paper cites DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DPOK: Reinforcement learning for fine-tuning text-to-image diffusion models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.486268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.486268Z digest=sha256:861f04c4709729c237625a345a5bd1d466f53212a8646ed26905f6a0ed951fd5

Observation 3c9651ac-24b0-4253-bc3a-94c29f1f306b · outbound

This paper cites Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair Diffusion: Instructing Text-to-Image Generation Models on Fairness

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.488625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.488625Z digest=sha256:cc02fc8ab2adf013a1ef2aae8ccf92f2520cceea13b9ca44286b2a09b0f8125c

Observation da91c3c7-2659-4af8-91f4-b6c97dc0cf8b · outbound

This paper cites FairImagen: Post- processing for bias mitigation in text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairImagen: Post- processing for bias mitigation in text-to-image models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.491290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.491290Z digest=sha256:f52e96c284a52a91c0043cdb80b209c04998bec1eda6b578ed3e51c91ce62483

Observation f7b880c9-7001-4b37-bcca-39b6c3193b86 · outbound

This paper cites Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving Dynamic Object Interactions in Text-to-Video Generation with AI Feedback

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.493826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.493826Z digest=sha256:2e9b82858ba2c9b3d2dc208f5a4f90e6b3690a743a2529742fa2883611616338

Observation 3149b138-104b-4a52-b9f1-cbb0718bc99b · outbound

This paper cites Unified concept editing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unified concept editing in diffusion models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.496320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.496320Z digest=sha256:2c8368ba716c34120ba4f3e1b1cc71af8bcabed5d555f93c0ce55bd26a62eb50

Observation f00ee3cd-ecd5-47ae-b844-71a294de6562 · outbound

This paper cites Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Using Reward Uncertainty to Induce Diverse Behaviour in Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.498785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.498785Z digest=sha256:c58559953e09a7265ebcb007fb85521fe8c048ce74422ef9318a12503b1fc796

Observation 020db60b-35ba-45bd-9967-7daf77b9f646 · outbound

This paper cites Polychromic objectives for reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Polychromic objectives for reinforcement learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.501274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.501274Z digest=sha256:915d84875e1b1ae892b57fa1d529ba5d2582e18f1e9fd01b1769e2fd1773a4f9

Observation 5aa290b9-9b70-42dd-8d0d-7e8220f21cef · outbound

This paper cites TempFlow-GRPO: When Timing Matters for GRPO in Flow Models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation TempFlow-GRPO: When Timing Matters for GRPO in Flow Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.503578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.503578Z digest=sha256:e877cc50e624d1cc8b43602af15a56bbace68fbc2a99e33ff840c51a58c91379

Observation 49fafe22-6b62-49ed-a2cc-2da025729424 · outbound

This paper cites Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Debiasing Diffusion Model: Enhancing Fairness through Latent Representation Learning in Stable Diffusion Model

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.506040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.506040Z digest=sha256:9ac9f2759c61988dbb65a0fcbf2e2f14e32d2e6a3e84e577f31312fbf4bd3095

Observation 1e78e774-80aa-4405-836c-8059598e90bd · outbound

This paper cites FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairGen: Enhancing fairness in text-to-image diffusion models via self-discovering latent directions.arXiv preprint arXiv:2412.18810, 2024

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.508514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.508514Z digest=sha256:0799e1ba8f32194d689e9516514d2680a75cdd3f6a8d6ca4f3c175be2621622b

Observation 78838bea-cd1e-4f7d-aa32-9eb7400ea55b · outbound

This paper cites FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation FairFace: Face attribute dataset for balanced race, gender, and age for bias measurement and mitigation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.510955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.510955Z digest=sha256:419e2467e69879bf77f1fcfaf028f33084188570feccab7fbf24c95c8b26a0d2

Observation 1b98a6db-d775-4a7b-b363-37f8757dd418 · outbound

This paper cites Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Rethinking training for de-biasing text-to-image generation: Unlocking the potential of stable diffusion

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.513219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.513219Z digest=sha256:3bea1ddc57124d108626b04862828f6c49a76766c26b4c96bd3cc05c8469cba5

Observation ecae609f-eeff-4e44-98df-fb10d3eaec18 · outbound

This paper cites Pick-a-Pic: An open dataset of user preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pick-a-Pic: An open dataset of user preferences for text-to-image generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.515488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.515488Z digest=sha256:f5a738bd81e49d2d3f314697a693586a4a52f35d90fa7d03a492ffa0a188a9db

Observation 95189d68-7248-40db-8a92-200f1f05fc54 · outbound

This paper cites Emergence of exploration in policy gradient reinforcement learning via resetting.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of exploration in policy gradient reinforcement learning via resetting

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.517916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.517916Z digest=sha256:b870070735e9ed6072bd680a7ba86bb73d5d16bbd0a7f6caf854f100d9d68ead

Observation 5e72bcb7-1640-4ad8-a53a-87498f3da14f · outbound

This paper cites Improved precision and recall metric for assessing generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improved precision and recall metric for assessing generative models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.520460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.520460Z digest=sha256:bdbc4f7cf70c921fd60d44c0d9c70f27072bbf8332a7e7c8004d8fc13973b1eb

Observation 1ab7d50b-32f1-4b0f-8792-8d377c397003 · outbound

This paper cites Holistic evaluation of text-to-image models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Holistic evaluation of text-to-image models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.522754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.522754Z digest=sha256:7e5836dfc92b21bdf7a4f8e8cc77dd32b0d116234f97bcd31bbf96001c983f8a

Observation d410701a-f8da-4728-aa45-a184037b381a · outbound

This paper cites SetPO: Set-level policy optimization for diversity-preserving LLM reasoning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation SetPO: Set-level policy optimization for diversity-preserving LLM reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.525144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.525144Z digest=sha256:e19f6d0892b0496229d35099ed181d5b067455476eb16e61182da4baa1264c30

Observation 4b7add92-3928-48b7-890d-161f66c19520 · outbound

This paper cites Fair text-to-image diffusion via fair mapping.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Fair text-to-image diffusion via fair mapping

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.527336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.527336Z digest=sha256:1e7972ef658c82c06e47e48720edcbd3c6d23634cc00fbcce184afbe11cc6d1a

Observation a964ee23-6e0a-4e5b-be4b-63a9f1fa8859 · outbound

This paper cites DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DiverseGRPO: Mitigating mode collapse in image generation via diversity-aware GRPO.arXiv preprint arXiv:2512.21514, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.529558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.529558Z digest=sha256:c31c8b1701e5754315d1e29d11ff09f632d90a211b55a1aaab34b2ac24a695f0

Observation 8b575eda-5613-4e9b-9eb6-c45d16613aca · outbound

This paper cites Flow-GRPO: Training flow matching models via online RL.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Flow-GRPO: Training flow matching models via online RL

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.531859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.531859Z digest=sha256:fa60c05788627edb09a226634bf5d277452ca271d70cf9039a9fb1318a2c1ee9

Observation c4d889dd-e52b-4b4e-82db-d6bc56dbf0f7 · outbound

This paper cites Improving video generation with human feedback.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Improving video generation with human feedback

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.534200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.534200Z digest=sha256:0079fbaba21496593128734ca53c1da06e22fc066545e6c6e10fb6b0bfcc1e45

Observation b613b1bf-43e1-47fc-92aa-7b2f4673c058 · outbound

This paper cites Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the Dirac Delta: Mitigating diversity collapse in reinforcement fine-tuning for versatile image generation.arXiv preprint arXiv:2601.12401, 2026

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.536646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.536646Z digest=sha256:a26bf3b8162f6cfe55ce8de26ff4e206c87fb65b936af087ddd4c371a2a8adf6

Observation 8c8ec63c-b297-425d-a09e-59593ac35a11 · outbound

This paper cites Stable bias: Evaluating societal representations in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Stable bias: Evaluating societal representations in diffusion models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.538932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.538932Z digest=sha256:fbd61788335cda4cc4c00778d34ce25d135befa1672cb57732e2062ea586bd3f

Observation 5628b4ff-29c0-46f2-a9ea-7e830cefaa8b · outbound

This paper cites Training diffusion models towards diverse image generation with reinforcement learning.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Training diffusion models towards diverse image generation with reinforcement learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.541267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.541267Z digest=sha256:fab59cf90df971fb27e8976fe5185210ebdfb5a1f67430df1369a1abeacf4ea2

Observation 956c240d-ac93-40ca-9bfb-6c9d3bcb2811 · outbound

This paper cites Reliable fidelity and diversity metrics for generative models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Reliable fidelity and diversity metrics for generative models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.543666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.543666Z digest=sha256:0b68fbef711c86615af15a4285628ee65e25b740afae617418bf1f5d53595d40

Observation 98fbfd37-543b-4703-916f-ca483b3488e8 · outbound

This paper cites Retry Policy Gradients in Continuous Action Spaces.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Retry Policy Gradients in Continuous Action Spaces

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.545963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.545963Z digest=sha256:1eab5dd3b1f4e4fc5dcfe37386027e1fab453c49b1e2f15f2df71649e2dcc2be

Observation 3f06b6f3-348d-4473-9602-e638dd50a9ab · outbound

This paper cites Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.548554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.548554Z digest=sha256:8598c3225f2670ce6f422773a1d77809e9a84c6f37572f50498790dab56dec0a

Observation ce1d8d97-b08d-4084-b613-8308d7161789 · outbound

This paper cites Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese).

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diversity-aware max@k optimization for improving best-of-N performance in image generation with diffusion models (in Japanese)

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.551264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.551264Z digest=sha256:6f17565032d3905d165e8726406304a1ae5852c46a9812fc3a4e30e27f1768c2

Observation bd6a6272-880d-4332-a414-dc436b678a0d · outbound

This paper cites Inference-time text-to-video alignment with diffusion latent beam search.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Inference-time text-to-video alignment with diffusion latent beam search

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.553646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.553646Z digest=sha256:dd3e9ef07de6d604a7b6d9c69109cb00cab5acbb4fc58f2e65f777d0ae23e82d

Observation 52efdac7-f5e6-4d8e-aa48-766f4ba67354 · outbound

This paper cites MultiBanana: A challenging benchmark for multi-reference text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation MultiBanana: A challenging benchmark for multi-reference text-to-image generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.556171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.556171Z digest=sha256:2524d538607b1482d3c58912b8f38dfdcd9dcdb10f6bec690682dab55692dcbc

Observation 5596e674-6524-4b44-8d7b-15cc3fd2caea · outbound

This paper cites Balancing act: Distribution-guided debiasing in diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Balancing act: Distribution-guided debiasing in diffusion models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.558556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.558556Z digest=sha256:022380efdec1866bc0847c345f2a492e2b9e60f6652cb96fa25c107be3e586fa

Observation 61c207e2-3e47-44ae-a1b7-dae906663094 · outbound

This paper cites OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation OrderGrad: Optimizing Beyond the Mean with Order-Statistic Policy Gradient Estimation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.561057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.561057Z digest=sha256:9edfb50b77aea565834e101feb25452382928781e6eab4b471f187f397100602

Observation 2ee8dd5a-d3e2-41e9-9c6d-493c763a4fd0 · outbound

This paper cites Escaping the mode: Multi-answer reinforcement learning in LMs.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Escaping the mode: Multi-answer reinforcement learning in LMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.563557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.563557Z digest=sha256:baab5d07adfbe45b633c4685623b2d1c4dce1f744d4059ffe4d971c5885670c5

Observation a407128f-6f22-4bdb-9148-3ceeb21bc320 · outbound

This paper cites From scale to speed: Adaptive test-time scaling for image editing.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation From scale to speed: Adaptive test-time scaling for image editing

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.565970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.565970Z digest=sha256:fffaef8d04733ef14e005b3cb927ed205f91b9af1509ab86a5d5a42c2776fcc5

Observation 29f0f96d-26a1-4413-840b-acc6c83eb4ff · outbound

This paper cites Learning transferable visual models from natural language supervision.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Learning transferable visual models from natural language supervision

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.568234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.568234Z digest=sha256:eb2056dfef26d1432c2c6feface90fe39fd9b9246edc7a15a7657f1191536b95

Observation 21b08dd3-e390-4521-b26c-6613d2fe8f62 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.570492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.570492Z digest=sha256:20ee8029c3d357608e348f8ff32a9ef8893e9ceaced7145c1d4239cb5af2f135

Observation 9b3a6574-b01c-40cc-8c6b-79122fbcf827 · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.572904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.572904Z digest=sha256:7d256c7cdb7871280ae2d1d5f35abc469467036a6bc172174a9af82998c60b6d

Observation 9f10100a-1356-40e4-9979-57777235438c · outbound

This paper cites an unresolved cited work.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.575402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.575402Z digest=sha256:8ea1c149ffb3be15a75d80fbf7dbbd8cb42ed0779dfc40d2ff1d5df1de30f854

Observation f4a61204-a395-43f6-9f89-da15f0ea4b5f · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.577708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.577708Z digest=sha256:896ef350af808f256f1927909d3e173360c19e54bffdb3aaff560a6c6c047205

Observation 16d91b2e-c927-4d36-8ddc-dc6809a22eff · outbound

This paper cites Finetuning text-to-image diffusion models for fairness.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finetuning text-to-image diffusion models for fairness

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.579983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.579983Z digest=sha256:810d14cb5855e3d1b0a16dd18389c4f2265b9d666f47e2f421b2bb2145efa513

Observation d3a013a1-75d5-4f88-a672-5f78cae46762 · outbound

This paper cites On Advantage Estimates for Max@K Policy Gradients.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation On Advantage Estimates for Max@K Policy Gradients

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.582293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.582293Z digest=sha256:796fc9fa94a9d33f472c8c9483895b37de521f3aadbb538b12d35cf2a429a110

Observation 90ddcec2-26b5-47db-b4da-ea82bb118978 · outbound

This paper cites Finite-Time Regret Analysis of Retry-Aware Bandits.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Finite-Time Regret Analysis of Retry-Aware Bandits

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.584759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.584759Z digest=sha256:1a8492b9cb32c6bd5a5280bd94e7dad9d6ca27935a2e0097b6eba62567a44b07

Observation 18167a3c-1c8f-4e50-9ed5-81008fded9d4 · outbound

This paper cites Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Beyond the prompt: Gender bias in text-to-image models, with a case study on hospital professions.arXiv preprint arXiv:2510.00045, 2025

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.587281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.587281Z digest=sha256:9044442d503242cdc1317fd00522061250377543fcade56a6c19c037542217ff

Observation 1a402d90-111e-4717-95c6-8f29f4a440f4 · outbound

This paper cites Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.589526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.589526Z digest=sha256:fe5e43a20313c717bf1230334478f974bd15b4201aa4d18bd534ae4ffcda5c77

Observation 39af284b-761a-4af8-84d1-f12370095706 · outbound

This paper cites Diffusion model alignment using direct preference optimization.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation Diffusion model alignment using direct preference optimization

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.591916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.591916Z digest=sha256:00716ee72321ee629f212b057162e56c12e4a9a78d64dab857fb7a172a517706

Observation 99a15030-df7d-4c09-b32d-139e4a24111f · outbound

This paper cites RewardDance: Reward Scaling in Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation RewardDance: Reward Scaling in Visual Generation

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.594851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.594851Z digest=sha256:7f9f9914bff641f9e1ab26616b9f1f2efc9d75ed212a8bfce63c1ca3929befbc

Observation e68a2686-f2c6-4ee4-b14c-b9d9d13acff1 · outbound

This paper cites ImageReward: Learning and evaluating human preferences for text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ImageReward: Learning and evaluating human preferences for text-to-image generation

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.597785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.597785Z digest=sha256:55ec582aa28c6dda1c39f94135248502c6b5888ff880868e26906a06f371380e

Observation 7a91fd58-a628-4ebf-ba52-da95886e125f · outbound

This paper cites DanceGRPO: Unleashing GRPO on Visual Generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation DanceGRPO: Unleashing GRPO on Visual Generation

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.600146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.600146Z digest=sha256:036b86de02b4779c5bad277bd424d5ca1fc7b633c34bcbf76332eec0bbce2f28

Observation 1391ca56-142d-400f-a972-89dd5b4d7b44 · outbound

This paper cites ITI-GEN: Inclusive text-to-image generation.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation ITI-GEN: Inclusive text-to-image generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T00:41:12.602630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.602630Z digest=sha256:647f22cfaf73662ee83323349bae9508a6c56d73c8f17de82cc5c0c2537ae383

Observation 08f4c5d8-6bb3-4181-94f6-4d654eb75c57 · outbound

This paper cites a photo of the face of a person.

Multi-Axis Max@K Reinforcement Learning for Representative Diversity in Text-to-Image Generation a photo of the face of a person

Reference 63

Resolution
malformed identifier
no resolver link, observed 2026-08-02T00:41:12.605274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:41:12.605274Z digest=sha256:f3c13308d23752d7bd431effebe803aceab543c0b13cf8bee13ebd665cfb6c27

Pith citing papers

No inbound Pith citation observations are available.