CMDAR: A Chinese Multi-scene Dynamic Audio Reasoning Benchmark with Diverse Challenges
CMDAR:一个包含多样挑战的中文多场景动态音频推理基准
Hui Li, Changhao Jiang, Hongyu Wang, Ming Zhang, Jiajun Sun, Zhixiong Yang, Yifei Cao, Shihan Dou, Xiaoran Fan, Baoyu Fan, Tao Ji, Tao Gui, Qi Zhang, Xuanjing Huang
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(计算机科学与人工智能学院,复旦大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
IEIT Systems Co Ltd(IEIT系统有限公司)
Less is more: Probabilistic reduction is best explained by small-scale predictability measures
少即是多:概率性缩减最好由小规模可预测性度量解释
Cassandra L. Jacobs, Andrés Buxó-Lugo, Anna K. Taylor, Marie Leopold-Hooke
机构
*
Department of Linguistics, University at Buffalo(语言学系,布法罗大学)
;
Department of Psychology, University at Buffalo(心理学系,布法罗大学)
;
Department of Computer Science and Engineering, University at Buffalo(计算机科学与工程系,布法罗大学)
Comments8 pages for main paper (exclude citation pages), 6 pages for appendix, totally 10 figures 7 tables and 2 algorithms. The paper is accepted by WACV 2026
Journal refIEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026
Uni-FinLLM: A Unified Multimodal Large Language Model with Modular Task Heads for Micro-Level Stock Prediction and Macro-Level Systemic Risk Assessment
机构
*
Department of Computer Science, Hong Kong Baptist University(香港 Baptist 大学计算机科学系)
;
Salesforce AI Research(Salesforce AI 研究院)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
National University of Singapore(新加坡国立大学)
;
The Education University of Hong Kong(香港教育大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL
机构
*
Postgraduate Institute of Medical Education and Research(医学教育与研究研究生院)
;
Tata Medical Center(塔塔医学中心)
;
All India Institute of Medical Sciences(全印度医学科学研究所)
;
National Cancer Institute(国家癌症研究所)
;
Banaras Hindu University(巴纳尔斯·胡斯·印度大学)
;
Aster Malabar Institute of Medical Sciences(Aster马拉巴医学科学研究所)
专题命中
评测与基准
:language model(title,abstract);small language model(abstract);prompting(abstract);分类 cs.CL、cs.AI
X-MuTeST: A Multilingual Benchmark for Explainable Hate Speech Detection and A Novel LLM-consulted Explanation Framework
X-MuTeST:一个用于可解释仇恨言论检测的多语言基准及一种新颖的LLM咨询解释框架
Mohammad Zia Ur Rehman, Sai Kartheek Reddy Kasu, Shashivardhan Reddy Koppula, Sai Rithwik Reddy Chirra, Shwetank Shekhar Singh, Nagendra Kumar
机构
*
Indian Institute of Technology Indore(印度理工学院Indore)
;
Indian Institute of Information Technology Dharwad(印度信息科技学院Dharwad)
;
Arizona State University(亚利桑那州立大学)
;
Indian Institute of Technology Mandi(印度理工学院Mandi)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL
Trust in LLM-controlled Robotics: a Survey of Security Threats, Defenses and Challenges
对LLM控制的机器人信任:安全威胁、防御和挑战的综述
Xinyu Huang, Shyam Karthick V B, Taozhao Chen, Mitch Bryson, Thomas Chaffey, Huaming Chen, Kim-Kwang Raymond Choo, Ian R. Manchester
机构
*
School of Electrical and Computer Engineering, The University of Sydney(悉尼大学电气与计算机工程学院)
;
Australian Centre for Robotics and School of Aerospace, Mechanical and Mechatronic Engineering, The University of Sydney(悉尼大学机器人中心及航空航天、机械与机电工程学院)
;
Department of Information Systems and Cybersecurity, University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校信息系统与网络安全系)
;
School of Engineering and Natural Sciences, University of Iceland(冰岛大学工程与自然科学学院)
专题命中
评测与基准
:LLM(title,abstract);large language model(abstract);language model(abstract)
机构
*
The State Key Laboratory of Complex and Critical Software Environment, Beihang University(复杂与关键软件环境国家重点实验室,北航)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
The Key Laboratory of Data Science and Intelligent Computing, International Innovation Institute, Beihang University(数据科学与智能计算重点实验室,国际创新研究院,北航)
Textual Explanations and Their Evaluations for Reinforcement Learning Policy
强化学习策略的文本解释及其评估
Ahmad Terra, Mohit Ahmed, Rafia Inam, Elena Fersman, Martin Törngren
机构
*
Ericsson Research(爱立信研究)
;
Ericsson AB(爱立信公司)
;
KTH Royal Institute of Technology(皇家理工学院)
;
Uppsala University(乌普萨拉大学)
;
Global AI Accelerator(全球人工智能加速器)
;
Ericsson Inc.(爱立信公司)
专题命中
评测与基准
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
CommentsWithdrawn by the authors due to an inconsistency in the reported base model: Section 4 (Experiments) states "Llama-2-7B" while Fig. 3 labels "Llama-2-7B-Chat". Because this affects the experimental configuration, parts of the results must be re-verified by rerunning experiments; we withdraw to avoid misleading readers