REVIEW 3 cited by
Perceptual Visual Quality Assessment: Principles, Methods, and Future Directions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
As multimedia services such as video streaming, video conferencing, virtual reality (VR), and online gaming continue to expand, ensuring high perceptual visual quality becomes a priority to maintain user satisfaction and competitiveness. However, multimedia content undergoes various distortions during acquisition, compression, transmission, and storage, resulting in the degradation of experienced quality. Thus, perceptual visual quality assessment (PVQA), which focuses on evaluating the quality of multimedia content based on human perception, is essential for optimizing user experiences in advanced communication systems. Several challenges are involved in the PVQA process, including diverse characteristics of multimedia content such as image, video, VR, point cloud, mesh, multimodality, etc., and complex distortion scenarios as well as viewing conditions. In this paper, we first present an overview of PVQA principles and methods. This includes both subjective methods, where users directly rate their experiences, and objective methods, where algorithms predict human perception based on measurable factors such as bitrate, frame rate, and compression levels. Based on the basics of PVQA, quality predictors for different multimedia data are then introduced. In addition to traditional images and videos, immersive multimedia and generative artificial intelligence (GenAI) content are also discussed. Finally, the paper concludes with a discussion on the future directions of PVQA research.
Forward citations
Cited by 3 Pith papers
-
Elastic3D: Controllable Stereo Video Conversion with Guided Latent Decoding
Elastic3D converts monocular video to stereo by directly synthesizing the right-eye view with a one-step latent diffusion model conditioned on a user-set median disparity, using a guided decoder to preserve left-view details.
-
VQualA 2025 Challenge on Image Super-Resolution Generated Content Quality Assessment: Methods and Results
A new SR image quality dataset focused on modern GAN and diffusion super-resolution outputs, plus benchmark results from four teams achieving SRCC above 0.90, is presented.
-
Beyond VMAF: Towards Application-Specific Metrics for Teleoperation Video
Retraining VMAF on teleoperation-specific subjective ratings reduces RMSE from 10.36 to 8.83 and MAD from 8.71 to 6.38 compared to the standard model.
Discussion (0). Sign in to comment.