Signals 4 · free daily AI digest

What, When, and How: Audio Description as Constrained Global Optimization

Paper recorded by Signals 4 on 2026-09-24 in cs.CL. Abstract reproduced from arXiv; link to the original below.

Published 2026-09-24 on arXiv · recorded by Signals 4 on 2026-09-25

Category: cs.CL · 自然语言处理 · first seen 2026-09-25

Abstract

Audio Description (AD) makes movies accessible to blind and visually impaired audiences by narrating visual information in gaps between dialogue. Existing automatic AD systems largely treat generation as a local video-to-text problem, assuming that the content to describe and its temporal location are already provided. Realistic AD instead requires coupled decisions about what visual information i

Read on arXiv →

#7 most recent of 252 cs.CL papers we have recorded · ↑ newer: Multimodal Thinking with Renderable Programs · ↓ older: R-DEIM Net: An Efficient Rationale-Augmented Dual-Expert Interaction M
Cite this page: What, When, and How: Audio Description as Constrained Global Optimization: the #7 most recent of 252 cs.CL papers we have recorded (as of 2026-09-24). Source: Signals 4 (Signals API) — https://data.jiangzhang.ca/signals4/t/papers/what-when-and-how-audio-description-as-constrained-global-optimization.html
Free to quote with attribution to “Signals 4 (Signals API)”. Machine-readable: papers.json
Related: More cs.CL papers · arXiv signals · All papers · Today in AI
Get 4 AI signals a day by email — free.
Subscribe free → See all plans →
Get 4 AI signals a day by email — free
All models · All repos · By company · Daily editions