Pith. sign in

Robust Speech Recognition via Large-Scale Weak Supervision,

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.SD 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

unclear 1

representative citing papers

From Sound to Sight: Towards AI-authored Music Videos

cs.SD · 2025-08-20 · conditional · novelty 5.0

This paper presents two off-the-shelf model pipelines (CLAP or LALM, an LLM, and a text-to-video model) for generating music videos from arbitrary songs, validated by a preliminary five-participant user study with modest results.

citing papers explorer

Showing 1 of 1 citing paper.

  • From Sound to Sight: Towards AI-authored Music Videos cs.SD · 2025-08-20 · conditional · none · ref 50

    This paper presents two off-the-shelf model pipelines (CLAP or LALM, an LLM, and a text-to-video model) for generating music videos from arbitrary songs, validated by a preliminary five-participant user study with modest results.