DiscoverPhysics: Benchmarking LLMs for Out-of-the-Box Scientific Thinking
DiscoverPhysics: 基准测试LLMs的即用型科学思维
Matt L. Wiemann, Lindsay M. Smith, Peter Melchior, Siddharth Mishra-Sharma, Andrew Gordon Wilson, Pavel Izmailov, Carolina Cuesta-Lázaro
机构
*
Princeton University(普林斯顿大学)
;
Boston University(波士顿大学)
;
New York University(纽约大学)
;
Flatiron Institute(Flatiron研究所)
;
Institute for Advanced Studies(高级研究研究所)
Dennis Frauen, Marie Brockschmidt, Konstantin Hess, Haorui Ma, Yuchen Ma, Abdurahman Maarouf, Maresa Schröder, Jonas Schweisthal, Yuxin Wang, Athiya Deviyani, Sonali Parbhoo, Rahul G. Krishnan, Stefan Feuerriegel
机构
*
Imperial College London(帝国理工学院伦敦分校)
;
University of Toronto(多伦多大学)
Emission-Aware Reinforcement Learning for Sustainable Electric Vehicle Charging and Carbon Dioxide Reduction Under Varying Renewable Penetration
面向可持续电动汽车充电与二氧化碳减排的排放感知强化学习:在不同可再生能源渗透率下
Ninglin Ou, Mohammad A. Razzaque, Iftekher Islam Shovon, Shafkat Khan Siam, Shafiuzzaman K Khadem, Krishnendu Guha, Mayeen U Khandaker, Md. Noor-A-Rahim
机构
*
organization= nasc Research, School of Computer Science \& IT, University College Cork , country = IE
;
organization= School of Computing, Engineering
;
Digital Technologies, Teesside University , country= UK
;
organization= International Energy Research Centre, Tyndall National Institute, Cork , country = IE
;
Radiation Technologies Group, CCDCU, Faculty of Engineering
;
Technology, Sunway University , country = Malaysia
;
organization= Department of Physics, College of Science, Korea University , country = Republic of Korea
机构
*
Independent Researcher(独立研究者)
;
Department of Computing Science, University of Alberta, Canada(阿尔伯塔大学计算机科学系)
;
Alberta Machine Intelligence Institute (Amii)(阿尔伯塔机器智能研究所)
机构
*
MOS Intelligent Connectivity Technology Co. Ltd.(MOS智能连接技术有限公司)
;
Sichuan Vocational College of Post and Telecom(四川邮电职业技术学院)
;
East China Normal University(华东师范大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shanghai Jiao Tong University(上海交通大学)
Disentangling Interaction and Bias Effects in Opinion Dynamics of Large Language Models
大型语言模型中意见动态的交互与偏差效应的分离
Vincent C. Brockers, David A. Ehrlich, Viola Priesemann
机构
*
Max-Planck-Institute for Dynamics and Self-Organization(马克斯·普朗克动态与自组织研究所)
;
Institute for the Dynamics of Complex Systems(复杂系统动力学研究所)
;
University of Göttingen(哥廷根大学)
;
Campus Institute for Dynamics of Biological Networks(校园生物网络动力学研究所)
Insulin Resistance Prediction From Wearables and Routine Blood Biomarkers
从可穿戴设备和常规血液生物标志物预测胰岛素抵抗
Ahmed A. Metwally, A. Ali Heydari, Daniel McDuff, Alexandru Solot, Zeinab Esmaeilpour, Anthony Z Faranesh, Menglian Zhou, David B. Savage, Conor Heneghan, Shwetak Patel, Cathy Speed, Javier L. Prieto
机构
*
Google Research(谷歌研究)
;
Institute of Metabolic Science, University of Cambridge(剑桥大学代谢科学研究所)
Measuring Security Without Fooling Ourselves: Why Benchmarking Agents Is Hard
在不欺骗自己的情况下衡量安全:为什么基准测试智能体是困难的
Sahar Abdelnabi, Chris Hicks, Konrad Rieck, Ahmad-Reza Sadeghi
机构
*
ELLIS Institute Tübingen & MPI-IS & Tübingen AI Center(图宾根ELLIS研究所及MPI-IS与图宾根人工智能中心)
;
The Alan Turing Institute(艾伦·图灵研究所)
;
BIFOLD & Technische Universität Berlin(BIFOLD与柏林技术大学)
;
Technische Universität Darmstadt(达姆施塔特技术大学)
Houxuan Zhou, Sriram Prasad, Chenghao Huang, Jiajie Feng, Hao Wang
机构
*
Department of Data Science and AI, Faculty of IT, Monash University, Australia(数据科学与人工智能系,IT学院,墨尔本大学,澳大利亚)
;
School of Electrical Engineering and Computer Science, University of Queensland, Australia(电气工程与计算机科学学院,昆士兰大学,澳大利亚)
;
Monash Energy Institute, Monash University, Australia(墨尔本能源研究所,墨尔本大学,澳大利亚)
stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation
stable-worldmodel: 一个用于可重复世界建模研究和评估的平台
Lucas Maes, Quentin Le Lidec, Luiz Facury, Nassim Massaudi, Ayush Chaurasia, Francesco Capuano, Richard Gao, Taj Gillin, Dan Haramati, Damien Scieur, Yann LeCun, Randall Balestriero
机构
*
Mila & Université de Montréal(Mila与蒙特利尔大学)
;
New York University(纽约大学)
;
Universidade Federal de Minas Gerais(巴西联邦大学矿务学院)
;
Independent Researcher(独立研究者)
;
LanceDB
;
University of Oxford(牛津大学)
;
Brown University(布朗大学)
Compositional Transduction with Latent Analogies for Offline Goal-Conditioned Reinforcement Learning
基于潜在类比的组合转导用于离线目标条件强化学习
Junseok Kim, Dohyeong Kim, Mineui Hong, Songhwai Oh
机构
*
Department of Electrical and Computer Engineering and ASRI, Seoul National University(电气与计算机工程系和首尔国立大学ASRI)
;
Independent researcher(独立研究者)
;
Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)
AI-Powered Facial Mask Removal Is Not Suitable For Identification
基于AI的面部遮挡去除并不适合识别
Emily A Cooper, Hany Farid
机构
*
Herbert Wertheim School of Optometry & Vision Science University of California, Berkeley(赫伯特·韦瑟姆视觉科学学院,加州大学伯克利分校)
;
School of Information University of California, Berkeley(信息学院,加州大学伯克利分校)
ShadeBench: A Benchmark Dataset for Building Shade Simulation in Sustainable Society
ShadeBench: 一个用于可持续社会建筑阴影模拟的基准数据集
Longchao Da, Mithun Shivakoti, Xiangrui Liu, T Pranav Kutralingam, Yezhou Yang, Hua Wei
机构
*
School of Computing and Augmented Intelligence, Arizona State University(计算与增强智能学院,亚利桑那州立大学)
;
Global Futures Laboratory, Arizona State University(全球未来实验室,亚利桑那州立大学)