REVIEW 4 cited by
A strong baseline for image and video quality assessment
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In this work, we present a simple yet effective unified model for perceptual quality assessment of image and video. In contrast to existing models which usually consist of complex network architecture, or rely on the concatenation of multiple branches of features, our model achieves a comparable performance by applying only one global feature derived from a backbone network (i.e. resnet18 in the presented work). Combined with some training tricks, the proposed model surpasses the current baselines of SOTA models on public and private datasets. Based on the architecture proposed, we release the models well trained for three common real-world scenarios: UGC videos in the wild, PGC videos with compression, Game videos with compression. These three pre-trained models can be directly applied for quality assessment, or be further fine-tuned for more customized usages. All the code, SDK, and the pre-trained weights of the proposed models are publicly available at https://github.com/Tencent/CenseoQoE.
Forward citations
Cited by 4 Pith papers
-
DIVA-VQA: Detecting Inter-frame Variations in UGC Video Quality
DIVA-VQA selects high-difference patches between consecutive frames and uses SlowFast plus SwinT features to predict UGC video quality, reporting competitive state-of-the-art correlations and low runtime.
-
GaussianVAE: Adaptive Learning Dynamics of 3D Gaussians for High-Fidelity Super-Resolution
A VAE with transformer attention and Hessian-guided sampling is proposed to extrapolate 3D Gaussian Splatting scenes beyond their training resolution, claiming 0.015s inference and improved Chamfer distance and Censeo...
-
MSPT: A Lightweight Face Image Quality Assessment Method with Multi-stage Progressive Training
A lightweight face quality assessment model trained with progressive data diversity and resolution scaling achieves second place on the VQualA 2025 benchmark.
-
VQualA 2025 Challenge on Face Image Quality Assessment: Methods and Results
An ICCV 2025 workshop challenge compared lightweight face image quality assessment models under strict compute limits, and this report surveys the winning methods.
Discussion (0). Continue with ORCID to comment.