EEG-FM-Bench: A Comprehensive Benchmark for the Systematic Evaluation and Diagnostic Analyses of EEG Foundation Models
EEG-FM-Bench:脑电图基础模型系统评估与诊断分析的综合基准
Wei Xiong, Jiangtong Li, Jie Li, Kun Zhu, Changjun Jiang
机构
*
School of Computer Science and Technology, Tongji University, Shanghai, China(同济大学计算机科学与技术学院,上海,中国)
;
Translational Research Center, Shanghai Yangzhi Rehabilitation Hospital (Shanghai Sunshine Rehabilitation Center), China(上海杨氏康复医院(上海阳光康复中心)转化研究中心,中国)
Dual-branch Prompting for Multimodal Machine Translation
双分支提示用于多模态机器翻译
Jie Wang, Zhendong Yang, Liansong Zong, Xiaobo Zhang, Dexian Wang, Ji Zhang
机构
*
School of Computing and Artificial Intelligence, Southwest Jiaotong University(西南交通大学计算机与人工智能学院)
;
School of Computer and Software Engineering, Xihua University(西华大学计算机与软件工程学院)
;
School of Intelligent Medicine, Chengdu University of Traditional Chinese Medicine(成都中医药大学针灸推拿学院)
机构
*
BenchFlow
;
OSU
;
Amazon
;
UC Berkeley
;
UC Santa Cruz
;
UC Davis
;
Dartmouth
;
RLWRLD
;
Independent
;
Princeton University
;
Oxford University
;
Stanford University
;
USC
;
CMU
;
Foxconn
;
Zenity
;
UNSW
;
UT Austin
;
MSU
;
Duke University
;
ByteDance
;
UT Dallas
;
UC San Diego
;
Columbia University
;
University of Rochester
;
Cornell Tech
;
Georgia Tech
;
Cornell University
;
NEU
;
UCLA
;
Snap Inc.
;
Fanshawe College
;
University of Science and Technology of China
;
HKUST(GZ)
;
Anyscale
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
Spectro-Temporal Interference Confounds Phase Encoding in Spatial Audio Foundation Models
频谱-时间干扰混淆空间音频基础模型中的相位编码
Yuxuan Chen, Haoyuan Yu, Peize He
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Jilin University(吉林大学)
;
Hunan University(湖南大学)
;
University of Electronic Science and Technology of China(电子科技大学)
IMPACTeen: Intentions, Manipulation, Persuasion, Annotations, and Consequences in Teen Communication Dataset
IMPACTeen:青少年沟通数据集中的意图、操纵、说服、标注与后果
Aleksander Szczęsny, Wiktoria Mieleszczenko-Kowszewicz, Maciej Markiewicz, Beata Bajcar, Tomasz Adamczyk, Jolanta Babiak, Grzegorz Chodak, Przemysław Kazienko
机构
*
Wrocław University of Science and Technology(弗罗茨瓦夫理工大学)
The Faithfulness Gap: Certifying Semantic Equivalence Between Natural-Language and Formal Mathematical Statements
忠实性差距:认证自然语言与形式数学语句之间的语义等价性
Noor Islam S. Mohammad, Tamim Sheikh
机构
*
Department of Computer Science, Informatics Institute, Istanbul Technical University, İstanbul, Türkiye(信息学院计算机科学系,伊斯坦布尔技术大学,伊斯坦布尔,土耳其)
;
Department of Computer Science(计算机科学系)
;
Engineering, Jashore University of Science(工程系,贾沙尔大学科学学院)
PACUTE: Phonology-, Affix-, and Character-level Understanding of Tokens for Filipino
PACUTE: 面向菲律宾语的音韵、词缀和字符级词元理解
Jann Railey Montalan, David Demitri Africa, Jimson Paulo Layacan, Richell Isaiah Flores, Ivan Yuri De Leon, Lance Calvin Gamboa
机构
*
AI Singapore(AI新加坡)
;
Nanyang Technological University(南洋理工大学)
;
UK AI Security Institute(英国人工智能安全研究所)
;
Ateneo de Manila University(马尼拉雅典耀大学)
;
University of Birmingham(伯明翰大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
AgentBeats: Agentifying Agent Assessment for Openness, Standardization, and Reproducibility
AgentBeats:面向开放性、标准化和可复现性的智能体评估代理化
Xiaoyuan Liu, Jianhong Tu, Yuqi Chen, Siyuan Xie, Sihan Ren, Tianneng Shi, Gal Gantar, Evan Sandoval, Donghyun Lee, Daniel Miao, Peter J. Gilbert, Nick Hynes, Mauro Staver, Warren He, David Marn, Andrew Low, Xi Zhang, Elron Bandel, Michal Shmueli-Scheuer, Siva Reddy, Alexandre Drouin, Alexandre Lacoste, Ramayya Krishnan, Elham Tabassi, Yu Su, Victor Barres, Chenguang Wang, Wenbo Guo, Dawn Song
机构
*
University of California, Berkeley(加州大学伯克利分校)
;
Purdue University(普渡大学)
;
University of Ljubljana(卢布尔雅那大学)
;
University of Washington(华盛顿大学)
;
Oasis Labs
;
University of Maryland(马里兰大学)
;
IBM Research(IBM研究院)
;
Mila
;
McGill University(麦吉尔大学)
;
ServiceNow Research(ServiceNow研究院)
;
Carnegie Mellon University(卡内基梅隆大学)
;
National Institute of Standards and Technology(美国国家标准与技术研究院)
;
The Ohio State University(俄亥俄州立大学)
;
University of Cambridge(剑桥大学)
;
University of California, Santa Barbara(加州大学圣塔芭芭拉分校)
SorryDB: Can AI Provers Complete Real-World Lean Theorems?
SorryDB: AI证明者能完成现实世界的Lean定理吗?
Austin Letson, Leopoldo Sarra, Auguste Poiroux, Oliver Dressler, Paul Lezeau, Dhyan Aranha, Frederick Pu, Aaron Hill, Miguel Corredera Hidalgo, Julian Berman, George Tsoukalas, Lenny Taelman
机构
*
University of California, Berkeley(加州大学伯克利分校)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Multi-Agent Framework for Audit Risk Assessment with Explicit Uncertainty and Evidence Conflict Modeling
具有显式不确定性和证据冲突建模的审计风险评估多智能体框架
Yuhan Wang, Manqing Wang, Yixuan Lu, Zhaoyue Peng, Shengda Lin
机构
*
Columbia University(哥伦比亚大学)
;
Trine University(特林大学)
;
University of Sofia(索菲亚大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Westcliff University(韦斯特克莱夫大学)
FORTIS: Benchmarking Over-Privilege in Agent Skills
FORTIS:评估代理技能中的过度特权
Shawn Li, Chenxiao Yu, Han Wang, Wei Yang, Ryan Rossi, Franck Dernoncourt, Xiyang Hu, Philip Yu, Chaowei Xiao, Huan Zhang, Yue Zhao
机构
*
University of Southern California(南加州大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Adobe Research(Adobe研究)
;
Arizona State University(亚利桑那州立大学)
;
University of Illinois Chicago(伊利诺伊大学芝加哥分校)
;
Johns Hopkins University(约翰霍普金斯大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI