Benchmarking Autonomous Vehicles: A Driver Foundation Model Framework
自动驾驶性能评估:一种驾驶员基础模型框架
Yuxin Zhang, Cheng Wang, Hubert P. H. Shum
机构
*
National Key Laboratory of Automotive Chassis Integration and Bionics, Jilin University(汽车底盘集成与仿生国家重点实验室,吉林大学)
;
Department of Computer Science, Durham University(杜伦大学计算机科学系)
;
School of Engineering and Physical Sciences, Heriot-Watt University(赫瑞斯泰大学工程与物理科学学院)
;
DRIVEResearch
A Survey of Behavior Foundation Model: Next-Generation Whole-Body Control System of Humanoid Robots
人形机器人行为基础模型综述:下一代全身体控系统
Mingqi Yuan, Tao Yu, Wenqi Ge, Xiuyong Yao, Huijiang Wang, Jiayu Chen, Bo Li, Wei Zhang, Wenjun Zeng, Hua Chen, Xin Jin
机构
*
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系)
;
LimX Dynamics
;
Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究院、东部技术研究所)
;
Department of Data and Systems Engineering, The University of Hong Kong(香港大学数据与系统工程系)
;
CREATE Lab, EPFL(EPFL CREATE 实验室)
;
School of System Design and Intelligent Manufacturing, Southern University of Science and Technology(南方科技大学系统设计与智能制造学院)
;
ZJU-UIUC Institute, Zhejiang University(浙江大学ZJU-UIUC研究院)
;
INFIFORCE Intelligent Technology Co., Ltd.(INFIFORCE智能科技有限公司)
Beyond Scalar Scores: Reinforcement Learning for Error-Aware Quality Estimation of Machine Translation
超越标量评分:用于机器翻译错误感知质量估计的强化学习
Archchana Sindhujan, Girish A. Koushik, Shenbin Qian, Diptesh Kanojia, Constantin Orăsan
机构
*
Institute for People-Centred AI(以人为本的人工智能研究所)
;
University of Surrey(塞弗尔大学)
;
NICE Research Group(NICE研究组)
;
University of Oslo(奥斯陆大学)
;
Centre for Translation Studies(翻译研究中心)
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
Interaction-Grounded Learning for Contextual Markov Decision Processes with Personalized Feedback
基于交互的 contextual Markov 决策过程学习与个性化反馈
Mengxiao Zhang, Yuheng Zhang, Haipeng Luo, Paul Mineiro
机构
*
University of Iowa(爱荷华大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Southern California(南加州大学)
;
Microsoft Research(微软研究院)
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG
Reproducible Benchmarking for Lung Nodule Detection and Malignancy Classification Across Multiple Low-Dose CT Datasets
多低剂量CT数据集上的肺结节检测与恶性分类可重复基准测试
Fakrul Islam Tushar, Avivah Wang, Lavsen Dahal, Ehsan Samei, Michael R. Harowicz, Jayashree Kalpathy-Cramer, Kyle J. Lafata, Tina D. Tailor, Cynthia Rudin, Joseph Y. Lo
机构
*
Engineering Research Center of Intelligent Theranostics Technology and Instruments, Ministry of Education, School of Biomedical Engineering and Informatics, Nanjing Medical University(智能诊疗技术与仪器工程研究中心、教育部、生物医学工程与信息学院、南京医科大学)
;
Institute of Biomedical Engineering, University of Oxford(生物医学工程研究所、牛津大学)
;
Institute of High Performance Computing, Agency for Science, Technology and Research (A*STAR)(高性能计算研究所、科技研究局(A*STAR))
;
Department of Neurology, National Neuroscience Institute, Singapore, 308433, Singapore(神经病学部、新加坡国家神经科学研究所)
;
Duke-NUS Medical School, Singapore, 169857, Singapore(新加坡国立大学医学院)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
ViHERMES: A Graph-Grounded Multihop Question Answering Benchmark and System for Vietnamese Healthcare Regulations
ViHERMES: 一个基于图的多跳问答基准和系统用于越南卫生法规
Long S. T. Nguyen, Quan M. Bui, Tin T. Ngo, Quynh T. N. Vo, Dung N. H. Le, Tho T. Quan
机构
*
URA Research Group, Faculty of Computer Science and Engineering, Ho Chi Minh City University of Technology (HCMUT), VNU-HCM, Vietnam(URA研究组,计算机科学与工程学院,胡志明市技术大学(HCMUT),VNU-HCM,越南)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
KV-CoRE: Benchmarking Data-Dependent Low-Rank Compressibility of KV-Caches in LLMs
KV-CoRE: 对LLMs中KV缓存数据依赖性低秩压缩性的基准测试
Jian Chen, Zhuoran Wang, Jiayu Qin, Ming Li, Meng Wang, Changyou Chen, Yin Chen, Qizhen Weng, Yirui Liu
机构
*
University at Buffalo(布法罗大学)
;
Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究所(TeleAI),中国电信)
;
Delft University of Technology(代尔夫特理工大学)
;
University of Maryland(马里兰大学)
;
ByteDance(字节跳动)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
机构
*
School of Engineering, Westlake University(西湖大学工程学院)
;
Global College, Shanghai Jiao Tong University(上海交通大学全球学院)
;
Academy of Mathematics and Systems Science, Chinese Academy of Sciences(中国科学院数学与系统科学研究院)
;
Department of Geotechnical Engineering, Tongji University(同济大学地质工程系)
;
School of Physics, Peking University(北京大学物理学院)
;
Key Laboratory for Power Machinery and Engineering of M. O. E., Shanghai Jiao Tong University(上海交通大学机械工程重点实验室)
Cognitive Linguistic Identity Fusion Score (CLIFS): A Scalable Cognition-Informed Approach to Quantifying Identity Fusion from Text
认知语言身份融合评分(CLIFS):一种可扩展的认知导向方法,用于从文本中量化身份融合
Devin R. Wright, Jisun An, Yong-Yeol Ahn
机构
*
Center for Complex Networks and Systems Research, Luddy School of Informatics, Computing, and Engineering, Indiana University Bloomington(复杂网络与系统研究中心,信息学、计算与工程学院,印第安纳大学布卢明顿分校)
;
Cognitive Science Program, Indiana University Bloomington(认知科学项目,印第安纳大学布卢明顿分校)
;
School of Data Science, University of Virginia(数据科学学院,弗吉尼亚大学)
;
CulturePulse, Inc.(CulturePulse公司)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
CommentsAuthors' accepted manuscript (postprint; camera-ready). To appear in the Proceedings of EMNLP 2025. Pagination/footer layout may differ from the Version of Record
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
VideoVeritas:通过感知预设强化学习实现AI生成视频检测
Hao Tan, Jun Lan, Senyuan Shi, Zichang Tan, Zijian Yu, Huijia Zhu, Weiqiang Wang, Jun Wan, Zhen Lei
机构
*
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS部)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Shenzhen Institute of Advanced Technology (SIAT), Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
专题命中
评测与基准
:large language model(abstract);language model(abstract)