Pith. sign in

REVIEW 1 cited by

On Fairness of Unified Multimodal Large Language Model for Image Generation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.03429 v1 pith:SK2CPNOF submitted 2025-02-05 cs.CL cs.AI

classification cs.CLcs.AI
keywords biasu-mllmsmodeldemographicgenerationlanguageunifiedaffected
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Unified multimodal large language models (U-MLLMs) have demonstrated impressive performance in visual understanding and generation in an end-to-end pipeline. Compared with generation-only models (e.g., Stable Diffusion), U-MLLMs may raise new questions about bias in their outputs, which can be affected by their unified capabilities. This gap is particularly concerning given the under-explored risk of propagating harmful stereotypes. In this paper, we benchmark the latest U-MLLMs and find that most exhibit significant demographic biases, such as gender and race bias. To better understand and mitigate this issue, we propose a locate-then-fix strategy, where we audit and show how the individual model component is affected by bias. Our analysis shows that bias originates primarily from the language model. More interestingly, we observe a "partial alignment" phenomenon in U-MLLMs, where understanding bias appears minimal, but generation bias remains substantial. Thus, we propose a novel balanced preference model to balance the demographic distribution with synthetic data. Experiments demonstrate that our approach reduces demographic bias while preserving semantic fidelity. We hope our findings underscore the need for more holistic interpretation and debiasing strategies of U-MLLMs in the future.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Bias Analysis in Unconditional Image Generative Models

    cs.CV 2025-06 conditional novelty 6.0 of 10

    In unconditional image generators, measured attribute bias shifts are small and are strongly influenced by whether the attribute classifier's decision boundary falls in a dense or sparse region of the attribute's dist...

Pith tools