Pith. sign in

REVIEW 2 cited by

CaDM: Codec-aware Diffusion Modeling for Neural-enhanced Video Streaming

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2211.08428 v2 pith:WD2NBR22 submitted 2022-11-15 eess.IV cs.CVcs.LGcs.MM

classification eess.IVcs.CVcs.LGcs.MM
keywords videocadmdiffusionqualityencoderenhancementrestorationstreaming
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recent years have witnessed the dramatic growth of Internet video traffic, where the video bitstreams are often compressed and delivered in low quality to fit the streamer's uplink bandwidth. To alleviate the quality degradation, it comes the rise of Neural-enhanced Video Streaming (NVS), which shows great prospects for recovering low-quality videos by mostly deploying neural super-resolution (SR) on the media server. Despite its benefit, we reveal that current mainstream works with SR enhancement have not achieved the desired rate-distortion trade-off between bitrate saving and quality restoration, due to: (1) overemphasizing the enhancement on the decoder side while omitting the co-design of encoder, (2) limited generative capacity to recover high-fidelity perceptual details, and (3) optimizing the compression-and-restoration pipeline from the resolution perspective solely, without considering color bit-depth. Aiming at overcoming these limitations, we are the first to conduct an encoder-decoder (i.e., codec) synergy by leveraging the inherent visual-generative property of diffusion models. Specifically, we present the Codec-aware Diffusion Modeling (CaDM), a novel NVS paradigm to significantly reduce streaming delivery bitrates while holding pretty higher restoration capacity over existing methods. First, CaDM improves the encoder's compression efficiency by simultaneously reducing resolution and color bit-depth of video frames. Second, CaDM empowers the decoder with high-quality enhancement by making the denoising diffusion restoration aware of encoder's resolution-color conditions. Evaluation on public cloud services with OpenMMLab benchmarks shows that CaDM effectively saves up to 5.12 - 21.44 times bitrates based on common video standards and achieves much better recovery quality (e.g., FID of 0.61) over state-of-the-art neural-enhancing methods.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ScalablePromptus: Scalable and High-Fidelity Prompt-Based Video Streaming

    eess.IV 2026-07 conditional novelty 6.0 of 10

    Dropout-trained, rank-ordered prompt embeddings let a video receiver reconstruct useful frames from truncated prompts, cutting truncation-induced LPIPS degradation by 82–95% versus Promptus.

  2. Semantic-Aware Adaptive Video Streaming Using Latent Diffusion Models for Wireless Networks

    cs.MM 2025-02 reject novelty 4.0 of 10

    The paper proposes LD-ABS, an adaptive bitrate streaming framework that compresses I-frames with a latent diffusion model and reconstructs P and B frames at the receiver, claiming better QoE than existing ABR algorithms.

Pith tools