arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

University of Cambridge(剑桥大学)

2026-03-18 至 2026-03-18 共收录 7
2603.16713 2026-03-18 cs.SD

Evaluating Latent Space Structure in Timbre VAEs: A Comparative Study of Unsupervised, Descriptor-Conditioned, and Perceptual Feature-Conditioned Models

评估在音色VAE中的潜在空间结构:对无监督、描述符条件和感知特征条件模型的比较研究

Joseph Cameron, Alan Blackwell

机构 * Department of Computer Science \& Technology, University of Cambridge\ , United Kingdom

AI总结 本文比较了三种音乐音色生成VAE的潜在空间结构,发现基于感知特征的条件模型更紧凑且具有鉴别力,优于无监督和离散描述符条件模型。

Comments 5 pages, 1 figure, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16682 2026-03-18 cs.SD

A Semantic Timbre Dataset for the Electric Guitar

用于电吉他的情感音色数据集

Joseph Cameron, Alan Blackwell

机构 * Department of Computer Science \& Technology, University of Cambridge\ , United Kingdom

AI总结 本文提出一个电吉他音色数据集,通过19个语义音色描述符和对应幅度标注,支持音色控制和语义音频生成,验证了数据集的有效性。

Comments 5 pages, 7 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16435 2026-03-18 cs.CL

VQKV: High-Fidelity and High-Ratio Cache Compression via Vector-Quantization

VQKV:通过向量量化实现高保真和高比的缓存压缩

Yixuan Wang, Qingyu Shi, Jiayu Zhou, Dianbo Liu, Ziwei He, Zhouhan Lin

机构 * LUMIA Lab(LUMIA实验室) School of Artificial Intelligence(人工智能学院) Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) University of Cambridge(剑桥大学) National University of Singapore(新加坡国立大学)

AI总结 本文提出VQKV方法,利用向量量化实现高压缩比和高重建保真度的KV缓存压缩,实现在有限资源环境下提升模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16086 2026-03-18 cs.RO cs.AI cs.CV cs.SD

Towards the Vision-Sound-Language-Action Paradigm: The HEAR Framework for Sound-Centric Manipulation

迈向视觉-声音-语言-动作范式:用于以声音为中心的操作的HEAR框架

Chang Nie, Tianchen Deng, Guangming Wang, Zhe Liu, Hesheng Wang

机构 * School of Automation and Intelligent Sensing, Shanghai Jiao Tong University and Shanghai Key Laboratory of Navigation and Location Based Services(自动化与智能感知学院,上海交通大学,导航与基于位置的服务重点实验室) Department of Engineering, Cambridge University(工程系,剑桥大学)

AI总结 本文提出HEAR框架,通过整合声音、视觉、语言和本体感知,解决实时声音中心操作中的关键声音遗漏问题,强调因果持续性和显式时间学习的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15994 2026-03-18 cs.AI

Selective Memory for Artificial Intelligence: Write-Time Gating with Hierarchical Archiving

选择性记忆用于人工智能:具有层次归档的写时门控

Oliver Zahn, Simran Chana

机构 * Independent Researcher(独立研究者) University of Cambridge(剑桥大学)

AI总结 本文提出写时门控机制,通过复合显著性评分筛选知识对象,保持版本链以保存先前状态,实验证明其在噪声环境下优于传统方法,尤其在干扰比例增加时表现更优。

Comments 20 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21637 2026-03-18 cs.CV

CARE: A Molecular-Guided Foundation Model with Adaptive Region Modeling for Whole Slide Image Analysis

CARE:一种具有自适应区域建模的分子引导基础模型用于全切片图像分析

Di Zhang, Zhangpeng Gong, Xiaobo Pang, Jiashuai Liu, Junbo Lu, Hao Cui, Jiusong Ge, Zhi Zeng, Kai Yi, Yinghua Li, Si Liu, Tingsong Yu, Haoran Wang, Mireia Crispin-Ortuzar, Weimiao Yu, Chen Li, Zeyu Gao

机构 * Xi’an Jiaotong University(西安交通大学) University of Cambridge(剑桥大学) KingMed(康方生物) BGI Research(贝登基因研究院) A ⋆ STAR

AI总结 CARE通过自适应区域建模和分子引导,提升全切片图像分析的性能,实现对病理区域的精准识别与分类,优于现有基础模型。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.12272 2026-03-18 cond-mat.mtrl-sci cs.CE cs.LG math.CO

From structure mining to unsupervised exploration of atomic octahedral networks

从结构挖掘到原子八面体网络的无监督探索

R. Patrick Xian, Ryan J. Morelock, Ido Hadar, Charles B. Musgrave, Christopher Sutton

机构 * Department of Engineering, University of Cambridge(剑桥大学工程系) Department of Chemical and Biological Engineering, University of Colorado Boulder(科罗拉多大学波尔得分校化学与生物工程系) The Institute of Chemistry, Casali Center for Applied Chemistry, and the Center for Nanoscience and Nanotechnology, The Hebrew University of Jerusalem(耶路撒冷希伯来大学化学系、应用化学Casali中心和纳米科学与技术中心) Department of Chemistry and Biochemistry, University of South Carolina(南卡罗来纳大学化学与生物化学系)

AI总结 本文提出利用无监督机器学习自动化分析原子八面体网络的几何解析与分类,通过两个数据集验证了其在发现氧化态变化和揭示八面体连接规则中的有效性。

Comments updated version, incl. three supporting information files

详情

展开后加载摘要…

URL PDF HTML 收藏