Pith. sign in

REVIEW 2 cited by

RealisHuman: A Two-Stage Approach for Refining Malformed Human Parts in Generated Images

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.03644 v2 pith:TUAFKSRC submitted 2024-09-05 cs.CV

classification cs.CV
keywords partsrealishumanhumanrealisticfacesframeworkgenerationhands
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In recent years, diffusion models have revolutionized visual generation, outperforming traditional frameworks like Generative Adversarial Networks (GANs). However, generating images of humans with realistic semantic parts, such as hands and faces, remains a significant challenge due to their intricate structural complexity. To address this issue, we propose a novel post-processing solution named RealisHuman. The RealisHuman framework operates in two stages. First, it generates realistic human parts, such as hands or faces, using the original malformed parts as references, ensuring consistent details with the original image. Second, it seamlessly integrates the rectified human parts back into their corresponding positions by repainting the surrounding areas to ensure smooth and realistic blending. The RealisHuman framework significantly enhances the realism of human generation, as demonstrated by notable improvements in both qualitative and quantitative metrics. Code is available at https://github.com/Wangbenzhi/RealisHuman.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. AnimateAnywhere: Rouse the Background in Human Image Animation

    cs.CV 2025-04 conditional novelty 6.0 of 10

    AnimateAnywhere learns to animate the background of human videos directly from human pose sequences, using an epipolar-constrained 3D attention mechanism, and reports state-of-the-art results without camera trajectories.

  2. ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping

    cs.CV 2024-12 conditional novelty 6.0 of 10

    ManiVideo generates bimanual hand-object manipulation videos conditioned on 3D motion sequences, using a multi-layer occlusion representation and Objaverse-based training to improve 3D consistency and object generalization.

Pith tools