CDPR integrates polarization priors into a diffusion-based monocular depth estimator via shared latent space and adaptive gating, outperforming RGB-only methods in challenging scenes.
Laion- 5b: An open large-scale dataset for training next generation image-text models
3 Pith papers cite this work. Polarity classification is still indexing.
years
2026 3representative citing papers
Hierarchical anti-aesthetic adversarial noise, guided by global and face-local preference reward models, degrades customized diffusion outputs and reduces facial identity leakage more than prior cloaking methods.
SwiftAudio performs caption-only distillation of a one-step TTA diffusion model by adapting VSD to audio with temporal smoothness regularization, achieving SOTA among one-step methods on AudioCaps and Clotho using ~45K captions.
citing papers explorer
-
CDPR: Cross-modal Diffusion with Polarization for Reliable Monocular Depth Estimation
CDPR integrates polarization priors into a diffusion-based monocular depth estimator via shared latent space and adaptive gating, outperforming RGB-only methods in challenging scenes.
-
Hierarchical Anti-Aesthetics: Protecting Facial Privacy against Customized Diffusion Models
Hierarchical anti-aesthetic adversarial noise, guided by global and face-local preference reward models, degrades customized diffusion outputs and reduces facial identity leakage more than prior cloaking methods.
-
SwiftAudio: Data-Efficient Caption-Only Distillation for One-Step Text-to-Audio Diffusion-based Generation
SwiftAudio performs caption-only distillation of a one-step TTA diffusion model by adapting VSD to audio with temporal smoothness regularization, achieving SOTA among one-step methods on AudioCaps and Clotho using ~45K captions.