BLUFF: Benchmarking the Detection of False and Synthetic Content across 58 Low-Resource Languages
BLUFF:跨58种低资源语言的虚假和合成内容检测基准测试
Jason Lucas, Matt Murtagh-White, Adaku Uchendu, Ali Al-Lawati, Michiharu Yamashita, Dominik Macko, Ivan Srba, Robert Moro, Dongwon Lee
机构
*
Penn State University(宾夕法尼亚州立大学)
;
Trinity College Dublin(都柏林圣三一学院)
;
MIT Lincoln Lab(麻省理工学院林肯实验室)
;
Visa Research(Visa研究)
;
Kempelen Institute of Intelligent Technologies(凯普勒智能技术研究所)
CommentsStrengthening SWE-Bench Verified and SWE-Bench Pro through adversarial test augmentation to improve the semantic reliability of LLM-based code agents
GroundedSurg: A Multi-Procedure Benchmark for Language-Conditioned Surgical Tool Segmentation
GroundedSurg: 一种多手术流程的语言条件手术工具分割基准
Tajamul Ashraf, Abrar Ul Riyaz, Wasif Tak, Tavaheed Tariq, Sonia Yadav, Moloud Abdar, Janibul Bashir
机构
*
King Abdullah University of Science and Technology (KAUST)(卡奥尔大学科学与技术学院)
;
Thapar Institute of Engineering and Technology(塔帕尔工程与技术学院)
;
The University of Queensland(昆士兰大学)
;
Gaash Research Lab, National Institute of Technology Srinagar(加什研究实验室,锡纳加尔国家理工学院)
CommentsWe present a framework for evaluation of Multi-modal Agents consisting of Voice-to-voice model components viz. Text to Speech (TTS), Retrieval Augmented Generation (RAG) and Speech-to-text (STT)
Validation of Space Robotics in Underwater Environments via Disturbance Robustness Equivalency
通过扰动鲁棒性等效性验证水下环境中的空间机器人
Joris Verhagen, Elias Krantz, Chelsea Sidrane, David Dörner, Nicola De Carli, Pedro Roque, Huina Mao, Gunnar Tibert, Ivan Stenius, Christer Fuglesang, Dimos Dimarogonas, Jana Tumova
机构
*
Division of Robotics, Perception and Learning, KTH Royal Institute of Technology(机器人、感知与学习 division,皇家理工学院)
;
School of Engineering Sciences, KTH Royal Institute of Technology(工程科学学院,皇家理工学院)
;
Division of Decision and Control Systems, KTH Royal Institute of Technology(决策与控制系统 division,皇家理工学院)
;
Department of Mechanical and Civil Engineering, California Institute of Technology(机械与土木工程 department,加州理工学院)
Mental Models of Autonomy and Sentience Shape Reactions to AI
自主性与意识的内心模型影响对AI的反应
Janet V. T. Pauketat, Daniel B. Shank, Aikaterina Manoli, Jacy Reese Anthis
机构
*
Sentience Institute(意识研究所)
;
Missouri University of Science and Technology(密苏里科技大学)
;
Max Planck Institute for Human Cognitive and Brain Sciences(人类认知与脑科学Max Planck研究所)
;
Stanford University(斯坦福大学)
;
University of Chicago(芝加哥大学)
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(计算机科学与人工智能学院,复旦大学)
;
School of Information and Intelligent Science, Donghua University(信息与智能科学学院,东华大学)