arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Johns Hopkins University(约翰斯·霍普金斯大学)

2025-11-27 至 2025-11-27 共收录 3
2511.21191 2025-11-27 cs.CV

Scenes as Tokens: Multi-Scale Normal Distributions Transform Tokenizer for General 3D Vision-Language Understanding

场景作为标记:多尺度正常分布变换标记器用于通用3D视觉-语言理解

Yutao Tang, Cheng Zhao, Gaurav Mittal, Rohith Kukkala, Rama Chellappa, Cheng Peng, Mei Chen

机构 * Johns Hopkins University(约翰霍普金斯大学) Microsoft(微软公司)

AI总结 NDTokenizer3D通过多尺度NDT表示和解码器实现通用3D视觉-语言理解,提升3D场景任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09573 2025-11-27 cs.LG stat.ML

Group Averaging for Physics Applications: Accuracy Improvements at Zero Training Cost

群平均用于物理应用:零训练成本下的精度提升

Valentino F. Foit, David W. Hogg, Soledad Villar

机构 * Center for Cosmology and Particle Physics(宇宙与粒子物理中心) Department of Physics(物理系) New York University(纽约大学) Department of Applied Mathematics & Statistics and Mathematical Institute for Data Science(应用数学与统计系及数据科学数学研究所) Johns Hopkins University(约翰霍普金斯大学)

AI总结 本文提出通过群平均技术在零训练成本下提升物理应用中模型的预测精度,实验表明该方法能有效降低评估损失并改善连续动态预测效果。

Comments 10 pages, 2 figures, 1 table, Machine Learning and the Physical Sciences Workshop, NeurIPS 2025

Journal ref Machine Learning and the Physical Sciences Workshop, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10772 2025-11-27 cs.CV

FlowTok: Flowing Seamlessly Across Text and Image Tokens

FlowTok: 在文本和图像标记之间无缝流动

Ju He, Qihang Yu, Qihao Liu, Liang-Chieh Chen

机构 * ByteDance Seed(字节跳动种子) Johns Hopkins University(约翰霍普金斯大学)

AI总结 FlowTok通过将图像编码为紧凑的1D标记,实现文本与图像模态间的无缝流动,提升效率与性能。

Comments Project page at https://tacju.github.io/projects/flowtok.html

详情

展开后加载摘要…

URL PDF HTML 收藏