arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Harvard University(哈佛大学)

2026-02-05 至 2026-02-05 共收录 5
2602.04752 2026-02-05 cs.LG

Decomposing Query-Key Feature Interactions Using Contrastive Covariances

通过对比协方差分解查询-键特征交互

Andrew Lee, Yonatan Belinkov, Fernanda Viégas, Martin Wattenberg

机构 * Harvard University(哈佛大学) Technion - Israel Institute of Technology(技术ion-以色列理工学院) Kempner Institute, Harvard University(凯普纳研究所,哈佛大学) Google DeepMind(谷歌DeepMind)

AI总结 本文提出通过对比协方差分解查询-键特征交互,以揭示Transformer中注意力机制的可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22208 2026-02-05 cs.CL cs.CV cs.LG

Open-Source Multimodal Moxin Models with Moxin-VLM and Moxin-VLA

开源多模态Moxin模型:Moxin-VLM和Moxin-VLA

Pu Zhao, Arash Akbari, Xuan Shen, Zhenglun Kong, Yixin Shen, Sung-En Chang, Timothy Rupprecht, Lei Lu, Enfu Nan, Changdi Yang, Yumei He, Weiyan Shi, Xingchen Xu, Yu Huang, Wei Jiang, Wei Wang, Yue Chen, Yong He, Yanzhi Wang

机构 * Northeastern University(东北大学) Harvard University(哈佛大学) Cornell University(康奈尔大学) Tulane University(路易斯安那州立大学) University of Washington(华盛顿大学) Futurewei(未来通信) AIBAO LLC

AI总结 本文提出开源多模态Moxin模型,通过Moxin-VLM、Moxin-VLA和Moxin-Chinese三种变体,提升视觉-语言、视觉-语言-动作及中文任务的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04227 2026-02-05 cs.CV

An Intuitionistic Fuzzy Logic Driven UNet architecture: Application to Brain Image segmentation

直觉模糊逻辑驱动的UNet架构:应用于脑图像分割

Hanuman Verma, Kiho Im, Pranabesh Maji, Akshansh Gupta

机构 * Department of Mathematics, Bareilly College, Bareilly (MJP Rohilkhand University), Uttar Pradesh, India(印度乌塔兰普尔邦巴雷利学院数学系) Division of Newborn Medicine, Fetal Neonatal Neuroimaging and Developmental Science Center, Boston Children’s Hospital, Harvard Medical School, Boston, MA 02115, USA(美国哈佛医学院波士顿儿童医院新生儿医学部) Department of Pediatrics, Harvard Medical School, Boston, MA, USA(美国哈佛医学院儿科系) CSIR–Central Electronics Engineering Research Institute, Pilani 333031, Rajasthan, India(印度拉贾斯坦邦比拉尼中央电子工程研究机构) Academy of Scientific and Innovative Research (AcSIR), Ghaziabad 201002, India(印度德里阿齐亚德瓦学术与创新研究院) CSIR–National Institute of Science Communication and Policy Research, New Delhi, India(印度新德里国家科学传播与政策研究所)

AI总结 本文提出基于直觉模糊逻辑的UNet架构,用于提升脑图像分割中对不确定性的处理能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03978 2026-02-05 cs.AI cs.LG

Monitorability as a Free Gift: How RLVR Spontaneously Aligns Reasoning

监控性作为免费礼物:RLVR如何自发地对齐推理

Zidi Xiong, Shan Chen, Himabindu Lakkaraju

机构 * Harvard University(哈佛大学)

AI总结 RLVR通过增强响应分布和提示关注度自发提升推理透明性,但这种提升依赖于数据多样性与训练难度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05059 2026-02-05 cs.CV

SurgiATM: A Physics-Guided Plug-and-Play Model for Deep Learning-Based Smoke Removal in Laparoscopic Surgery

SurgiATM:一种基于物理的插件式模型,用于基于深度学习的腹腔镜手术烟雾去除

Mingyu Sheng, Jianan Fan, Dongnan Liu, Guoyan Zheng, Ron Kikinis, Weidong Cai

机构 * School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) Institute of Medical Robotics, School of Biomedical Engineering, Shanghai Jiao Tong University(上海交通大学生物医学工程学院医学机器人研究所) Department of Radiology, Brigham and Women's Hospital, and Harvard Medical School(哈佛医学院放射科、布里洛女性医院)

AI总结 SurgiATM通过结合物理模型和深度学习,提出了一种轻量级插件式模型,用于有效去除腹腔镜手术中的烟雾,提升手术精度和效率。

Comments 21 pages, 9 figures, 10 tables. Code available at https://github.com/MingyuShengSMY/SurgiATM

详情

展开后加载摘要…

URL PDF HTML 收藏