arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 2608.08349cs.HCcs.AIcs.MM

Dramarrator:面向书籍音频剧制作的基于对象的音频编辑工具

Dramarrator: Object-Based Audio Editing for Audio Drama Production from Books

Karim Benharrak, Oriol Nieto, Bryan Wang, Zeyu Jin, Amy Pavel

首次发表
浏览论文内容

中文总结 AI 辅助

Dramarrator是一款基于对象的音频编辑工具,可从书籍提取故事元素生成关联音频资产,编辑对象时自动同步依赖内容,能降低音频剧制作任务负荷,其输出质量接近专业工具,还可推广至其他音频创作领域。

中文摘要 AI 辅助

音频剧将对话、音效和音乐编织成沉浸式故事。创作者常将书籍改编为音频剧,但该过程仍需大量人工,需解读源材料、编写脚本、生成音频资产并在时间轴上组装。由于角色和场景等故事元素分布在多个相互依赖的资产中,单次修改可能引发整个项目的手动更新。我们提出Dramarrator,一款基于对象音频编辑的音频剧创作工具,其中这些故事元素被表示为可编辑对象。Dramarrator从书籍中提取这些对象,生成关联的音频资产(语音、音效和音乐),并合成多轨音频剧。对任意对象(如角色语音)的编辑会自动传播到所有依赖资产。在针对专业人士的用户研究(N=8)中,Dramarrator显著降低了制作音频剧的任务负荷;在听众研究(N=300)中,经创作者优化的Dramarrator输出效果接近现有专业工具制作的作品;一项探索性研究(N=3)表明,基于对象的编辑降低了入门门槛且可推广至音频剧之外的领域。

英文摘要

Audio dramas weave dialogue, sound effects, and music into immersive stories. Creators often adapt books into audio dramas, but this process remains labor-intensive, requiring them to interpret source material, author scripts, generate audio assets, and assemble them on a timeline. Because story elements like characters and scenes manifest across many interdependent assets, a single change can ripple into manual updates across the entire project. We present Dramarrator, an audio drama authoring tool built around object-based audio editing, where these story elements are represented as editable objects. Dramarrator extracts these objects from a book, generates linked audio assets (speech, sound effects, and music), and composes a multi-track audio drama. Edits to any object (e.g., a character's voice) automatically propagate to all dependent assets. In a user study with professionals (N=8), Dramarrator significantly lowered task load when creating audio dramas. A listener study (N=300) shows that creator-refined output from Dramarrator approaches the quality of productions made with existing professional tools, and an exploratory study (N=3) suggests object-based editing lowers entry barriers and generalizes beyond audio dramas.

发表机构

  • University of California, Berkeley(加州大学伯克利分校)
  • Adobe Research(奥多比研究院)

机构由 AI 辅助整理,请以论文原文为准。

补充信息

↑