VisPlay: Self-Evolving Vision-Language Models from Images
VisPlay: 从图像中自我进化视觉-语言模型
Yicheng He, Chengsong Huang, Zongxia Li, Jiaxin Huang, Yonghui Yang
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Washington University in St. Louis(华盛顿大学圣路易斯分校)
;
University of Maryland(马里兰大学)
;
National University of Singapore(新加坡国立大学)
CARE-RAG - Clinical Assessment and Reasoning in RAG
CARE-RAG - 临床评估与推理中的RAG
Deepthi Potluri, Aby Mammen Mathew, Jeffrey B DeWitt, Alexander L. Rasgon, Yide Hao, Junyuan Hong, Ying Ding
机构
*
Department of Computer Science(计算机科学系)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Behavioral Science and Psychiatry(行为科学与精神病学)
;
Department of Statistics(统计学系)
;
University of Michigan(密歇根大学)
;
School of Information(信息学院)
机构
*
State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
Sensetime Research(商汤科技研究院)
;
Beijing Institute of Technology(北京理工大学)
;
Shanghai AI Lab(上海人工智能实验室)
Pass@k Metric for RLVR: A Diagnostic Tool of Exploration, But Not an Objective
Pass@k指标用于RLVR:探索的诊断工具,而非目标
Yang Yu
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University, China(新型软件技术国家实验室,南京大学,中国)
;
School of Artificial Intelligence, Nanjing University, China(人工智能学院,南京大学,中国)
Conan: Progressive Learning to Reason Like a Detective over Multi-Scale Visual Evidence
Conan:基于多尺度视觉证据的逐步学习以像侦探一样推理
Kun Ouyang, Yuanxin Liu, Linli Yao, Yishuo Cai, Hao Zhou, Jie Zhou, Fandong Meng, Xu Sun
机构
*
State Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机学院,北京大学)
;
WeChat AI, Tencent Inc., China(微信AI,腾讯公司,中国)
Nemotron Elastic: Towards Efficient Many-in-One Reasoning LLMs
Nemotron弹性:迈向高效多任务推理LLM
Ali Taghibakhshi, Sharath Turuvekere Sreenivas, Saurav Muralidharan, Ruisi Cai, Marcin Chochowski, Ameya Sunil Mahabaleshwarkar, Yoshi Suhara, Oluwatobi Olabiyi, Daniel Korzekwa, Mostofa Patwary, Mohammad Shoeybi, Jan Kautz, Bryan Catanzaro, Ashwath Aithal, Nima Tajbakhsh, Pavlo Molchanov