AI 中文总结
本研究用冻结音频编码器比较四种声音库,发现声学事件累积与序列依赖性差异取决于测量属性与时间尺度,而非单一等级。
AI 中文摘要
声音库在声学事件类型累积和时间组织上可能不同,但由于语料库使用不同的原生事件和不等量的序列,直接比较较为困难。我们使用相同的冻结音频编码器程序,在匹配事件数量和局部序列机会的条件下,比较了抹香鲸的编码叫声、人类语音音素、孟加拉雀的鸣叫音节和普通狨猴的叫声。鲸鱼显示出最快的类型累积;雀类显示出最强的即时依赖性和重复子序列的复发。物理上可解释的声学特征恢复了该概况的互补部分,无聚类的连续分析支持鲸鱼广泛的声学覆盖,而保留源和位置的零模型保留了雀类的顺序效应。扩展预测上下文将比较转向鲸鱼。因此,声音库的差异取决于所测量的声学属性和时间尺度,而非形成单一等级。
英文摘要
Vocal repertoires can differ in acoustic-event type accumulation and temporal organization, yet direct comparison is difficult because corpora use different native events and unequal amounts of sequence. We compare sperm whale codas, human speech phones, Bengalese finch syllables, and common marmoset calls using the same frozen-audio-encoder procedure while matching event count and local sequence opportunity. Whale shows the fastest type accumulation; Finch shows the strongest immediate dependence and repeated-subsequence recurrence. Physically interpretable acoustics recover complementary parts of this profile, continuous analyses without clustering support broad Whale acoustic coverage, and source- and position-preserving nulls retain both Finch order effects. Extending predictive context shifts the comparison toward Whale. Thus repertoire differences depend on the acoustic property and temporal scale measured rather than forming a single hierarchy.
Comments12 pages, 4 figures. Preprint