机构
*
Fudan University(复旦大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Zhejiang University(浙江大学)
;
The Chinese University of Hong Kong(香港中文大学)
专题命中
评测与基准
:large language model(title);language model(title);foundation model(abstract)
CommentsAccepted at Datasets and Benchmarks Track, ACM Knowledge Discovery and Data Mining (KDD) 2026. Project page: https://realize-lab.github.io/RELIANCE/
CommentsAccepted at the Third Conference on Parsimony and Learning (CPAL 2026). 36 pages, 12 figures. (Equal contribution: Yasaman Amou Jafari and Mahdi Noori.)
Journal refConference on Parsimony and Learning, Proceedings of Machine Learning Research, 328:989-1024, 2026
机构
*
Fudan University(复旦大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳分校)
;
National University of Singapore(新加坡国立大学)
专题命中
评测与基准
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
DREAM: Extending Vision-Language Models with Dual-Objective Encoding for Cross-Modal Retrieval
DREAM: 通过双目标编码扩展视觉-语言模型用于跨模态检索
Kaleem Ullah, Altaf Hussain, Muhammad Munsif, Sung Wook Baik
机构
*
Sejong University(世宗大学)
;
Korea Advanced Institute of Science and Technology(韩国科学技术院)
;
Ulsan National Institute of Science and Technology(乌山国立科学研究院)
机构
*
School of Computer Science and Engineering, Southeast University, China(东南大学计算机科学与工程学院,中国)
;
Alibaba Group(阿里巴巴集团)
;
School of Computer Science and Engineering, School of Intelligence Science and Engineering, and Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications, Southeast University, China(东南大学计算机科学与工程学院、智能科学与工程学院以及新一代人工智能技术及其交叉应用关键实验室,中国)
;
Wangxuan Institute of Computer Technology, National Key Laboratory for Multimedia Information Processing, Peking University, China(北京大学王轩计算机技术研究所、多媒体信息处理国家重点实验室,中国)
;
University of Copenhagen, Denmark(丹麦哥本哈根大学)
LALM-as-a-Judge: Benchmarking Large Audio-Language Models for Safety Evaluation in Multi-Turn Spoken Dialogues
LALM-as-a-Judge:用于多轮口语对话安全评估的大型音频语言模型基准测试
Amir Ivry, Shinji Watanabe
机构
*
Computer Engineering, Technion--Israel Institute of Technology, Haifa, Israel(技术学院电子工程系,技术离子技术研究所,以色列海法)
;
Language Technologies Institute, Carnegie Mellon University, Pittsburgh, PA, USA(语言技术研究所,卡内基梅隆大学,美国匹兹堡)
RECOM: A Validity Discrimination Tradeoff in Automatic Metrics for Open Ended Reddit Question Answering
RECOM:开放式 Reddit 问答中自动评估指标的有效性与区分性权衡
Pushwitha Krishnappa, Amit Das, Vinija Jain, Aman Chadha, Tathagata Mukherjee
机构
*
University of Alabama Huntsville(阿拉巴马大学亨茨维尔分校)
;
University of North Alabama(北阿拉巴马大学)
;
Stanford University(斯坦福大学)
;
Meta AI
;
Amazon GenAI(亚马逊生成人工智能)
机构
*
The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳))
;
School of Data Science, School of Artificial Intelligence, The Chinese University of Hong Kong, Shenzhen, China(数据科学学院、人工智能学院、香港中文大学(深圳))
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.AI