机构
*
Brain-inspired Cognitive Intelligence Lab, Institute of Automation, Chinese Academy of Sciences, Beijing, China(脑启发认知智能实验室,自动化研究所,中国科学院,北京,中国)
;
School of Future Technology, University of Chinese Academy of Sciences, China(未来技术学院,中国科学院大学,中国)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences, China(人工智能学院,中国科学院大学,中国)
;
Zhongguancun Academy, China(中关村学院,中国)
;
Beijing Key Laboratory of Safe AI and Superalignment(北京安全人工智能与超对齐重点实验室)
;
Gaoling School of AI, Renmin University of China(甘露人工智能学院,中国人民大学)
;
Beijing Institute of AI Safety and Governance (Beijing-AISI)(北京人工智能安全与治理研究院(北京-AISI))
;
School of Humanities, University of Chinese Academy of Sciences, China(人文学院,中国科学院大学,中国)
Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows
可能但错误:天体物理工作流中代理失败的案例研究
Shivam Rawat, Lucie Flek
机构
*
Bonn-Aachen International Center for Information Technology, University of Bonn, Germany(波恩-阿森国际信息科技中心,波恩大学,德国)
;
Lamarr Institute for Machine Learning and Artificial Intelligence, Germany(拉马尔机器学习与人工智能研究所,德国)
机构
*
Department of Computer and Network Engineering, United Arab Emirates University, UAE(计算机与网络工程系,阿联酋大学)
;
Research Institute for Digital Future, Khalifa University, UAE(未来数字研究院,哈利法大学)
CausalRAG2: Hierarchical Causal Knowledge Graph Design for RAG
CausalRAG2: 面向RAG的分层因果知识图谱设计
Nengbo Wang, Tuo Liang, Vikash Singh, Chaoda Song, Van Yang, Yu Yin, Jing Ma, Jagdip Singh, Vipin Chaudhary
机构
*
Department of Computer and Data Sciences, Case Western Reserve University, Cleveland, OH, USA(计算机与数据科学系,凯斯西储大学,克利夫兰,OH,USA)
;
Design and Innovation Department, Case Western Reserve University, Cleveland, OH, USA(设计与创新部门,凯斯西储大学,克利夫兰,OH,USA)
机构
*
Centennial High School, Frisco, Texas, USA(Centennial High School, Texas, USA)
;
Lebanon Trail High School, Frisco, Texas, USA(Lebanon Trail High School, Texas, USA)
;
West Windsor-Plainsboro High School, Princeton Junction, New Jersey, USA(West Windsor-Plainsboro High School, New Jersey, USA)
;
Algoverse AI Research, Palo Alto, California, USA(Algoververse AI Research, California, USA)
Comments14 pages, 4 figures, 8 tables. Presented at the 39th Conference on Neural Information Processing Systems Workshop: VLM4RWD. Presented at the 43th International Conference on Machine Learning Workshops: ICML 2026 CTB, ICML 2026 FAGEN, ICML 2026 EMM-QA. Authors Aahana Basappa and Pranay Goel contributed equally. Code: https://github.com/AahanaB24/AMVICC, Data: https://doi.org/10.5281/zenodo.17646068
SHERLOC: Structured Diagnostic Localization for Code Repair Agents
SHERLOC: 代码修复智能体的结构化诊断定位
Hovhannes Tamoyan, Sean Narenthiran, Erik Arakelyan, Mira Mezini, Boris Ginsburg
机构
*
NVIDIA(英伟达)
;
Santa Clara, CA 95051, USA(美国加利福尼亚州圣克拉拉)
;
TU Darmstadt(达姆施塔特工业大学)
;
National Research Center for Applied Cybersecurity ATHENE(国家应用网络安全研究中心 ATHENE)
PETRA: Transforming Web Text for Petroleum-Engineering Domain Adaptation
PETRA: 将网络文本转化为石油工程领域适应的数据集
Kirill Dubovikov, Omar El Mansouri, Hachem Madmoun, Yanda Li, Sandeep Kumar, Aya El Mir, Supriyo Ghosh, Writabrata Bhattacharya, Adrian Garcia-Garcia, Onkar Pandit, Sunil Kumar Sahu, Federico Castanedo, Larry Murray, Martin Takac, Salem Lahlou
机构
*
Mohamed bin Zayed University of Artificial Intelligence(莫扎伊德大学人工智能学院)
;
Inception AI
SICI: A Semantic-Pragmatic Complexity Index Reveals Regime Shifts in LLM Stance Detection
SICI:一种揭示LLM立场检测中相变的语义-语用复杂度指数
Fuqiang Niu, Bowen Zhang
机构
*
School of Cyber Science and Technology, University of Science and Technology of China(中国科学技术大学网络空间安全学院)
;
School of Artificial Intelligence, Shenzhen Technology University(深圳技术大学人工智能学院)
Business as Rulesual: A Benchmark and Framework for Business Rule Flow Modeling with LLMs
业务即规则:面向LLM的业务规则流建模基准与框架
Chen Yang, Ruping Xu, Ruizhe Li, Bin Cao, Jing Fan
机构
*
Zhejiang University of Technology(浙江工业大学)
;
Zhejiang Key Laboratory of Visual Information Intelligent Processing(浙江省视觉信息智能处理重点实验室)
;
University of Aberdeen(阿伯丁大学)
;
University of Birmingham(伯明翰大学)
The Correct Answer Trap: Pedagogically-Grounded Detection and Feedback for Hidden Misconceptions
正确答案陷阱:基于教学法的隐藏误解检测与反馈
Moiz Imran, Sahan Bulathwela
机构
*
Department of Computer Science, University College London, The United Kingdom(伦敦大学学院计算机科学系,英国)
;
Centre for Artificial Intelligence, University College London, The United Kingdom(伦敦大学学院人工智能中心,英国)