← back to paper
arxiv: 2604.11283 · 2 revisions
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey