BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent
Zijian Chen, Xueguang Ma, Shengyao Zhuang, Ping Nie, Kai Zou, Andrew Liu, Joshua Green, Kshama Patel, Ruoxi Meng, Mingyi Su, Sahel Sharifymoghaddam, Yanxi Li, Haoran Hong, Xinyu Shi, Xuye Liu, Nandan Thakur, Crystina Zhang, Luyu Gao, Wenhu Chen, Jimmy Lin
机构
*
University of Waterloo(滑铁卢大学)
;
CSIRO(澳大利亚联邦科学与工业研究组织)
;
Independent(独立研究者)
;
Carnegie Mellon University(卡内基梅隆大学)
;
The University of Queensland(昆士兰大学)
Positional Biases Shift as Inputs Approach Context Window Limits
Blerta Veseli, Julian Chibane, Mariya Toneva, Alexander Koller
机构
*
Saarland Informatics Campus, Saarland University, Germany(萨尔兰大学信息学校区)
;
Max Planck Institute for Informatics, Saarland Informatics Campus, Germany(马克斯·普朗克信息研究所)
;
Max Planck Institute for Software Systems, Saarland Informatics Campus, Germany(马克斯·普朗克软件系统研究所)
专题命中
推理评测
:reasoning(abstract);分类 cs.CL
Journal refConference on Language Modeling (COLM) 2025
RSVLM-QA: A Benchmark Dataset for Remote Sensing Vision Language Model-based Question Answering
Xing Zi, Jinghao Xiao, Yunxiao Shi, Xian Tao, Jun Li, Ali Braytee, Mukesh Prasad
机构
*
School of Computer Science, University of Technology Sydney(技术悉尼大学计算机科学学院)
;
SEDE, University of Technology Sydney(技术悉尼大学SEDE)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
专题命中
推理评测
:reasoning(abstract)
CommentsThis paper has been accepted to the proceedings of the 33rd ACM International Multimedia Conference (ACM Multimedia 2025)
From Prediction to Explanation: Multimodal, Explainable, and Interactive Deepfake Detection Framework for Non-Expert Users
Shahroz Tariq, Simon S. Woo, Priyanka Singh, Irena Irmalasari, Saakshi Gupta, Dev Gupta
机构
*
Sungkyunkwan University, S. Korea(顺天大学)
;
University of Queensland, Australia(昆士兰大学)
专题命中
推理评测
:reasoning(abstract)
Comments11 pages, 3 tables, 5 figures, accepted for publicaiton in the 33rd ACM International Conference on Multimedia (MM '25), October 27-31, 2025, Dublin, Ireland
机构
*
Independent Researcher in AI and Statistics(人工智能与统计学独立研究者)
;
Shahrood University of Technology(沙霍罗德大学)
;
University of Pittsburgh(匹兹堡大学)
;
Duquesne University(杜克森大学)
PrLM: Learning Explicit Reasoning for Personalized RAG via Contrastive Reward Optimization
Kepu Zhang, Teng Shi, Weijie Yu, Jun Xu
机构
*
Gaoling School of Artificial Intelligence\ University of China Beijing China
;
School of Information Technology
;
Gaoling School of Artificial Intelligence\ University of China
Keyword-Centric Prompting for One-Shot Event Detection with Self-Generated Rationale Enhancements
Ziheng Li, Zhi-Hong Deng
机构
*
State Key Laboratory of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University(一般人工智能国家重点实验室,智能科学与技术学院,北京大学)