REVIEW 2 cited by
Rarity Score : A New Metric to Evaluate the Uncommonness of Synthesized Images
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Evaluation metrics in image synthesis play a key role to measure performances of generative models. However, most metrics mainly focus on image fidelity. Existing diversity metrics are derived by comparing distributions, and thus they cannot quantify the diversity or rarity degree of each generated image. In this work, we propose a new evaluation metric, called `rarity score', to measure the individual rarity of each image synthesized by generative models. We first show empirical observation that common samples are close to each other and rare samples are far from each other in nearest-neighbor distances of feature space. We then use our metric to demonstrate that the extent to which different generative models produce rare images can be effectively compared. We also propose a method to compare rarities between datasets that share the same concept such as CelebA-HQ and FFHQ. Finally, we analyze the use of metrics in different designs of feature spaces to better understand the relationship between feature spaces and resulting sparse images. Code will be publicly available online for the research community.
Forward citations
Cited by 2 Pith papers
-
CREward: A Type-Specific Creativity Reward Model
CREward, trained only on Gemma-3-generated preference labels, predicts geometry/material/texture creativity rankings that correlate moderately with human designer judgments on a five-object benchmark.
-
Boost-and-Skip: A Simple Guidance-Free Diffusion for Minority Generation
Starting diffusion sampling from variance-boosted noise and skipping early timesteps generates minority samples at guided-method quality with far less compute.
Discussion (0). Continue with ORCID to comment.