Hybrid Panels: Toward Human-AI Collaboration in Survey Research
混合面板:面向调查研究中的人机协作
Julia Romberg, Tobias Gummer, Gabriella Lapesa, Tanja Kunz, Claudia Wagner
机构
*
GESIS - Leibniz Institute for the Social Sciences(莱布尼茨社会科学研究所(GESIS))
;
Heidelberg University(海德堡大学)
;
Heinrich Heine University of Düsseldorf(杜塞尔多夫海因里希·海涅大学)
;
RWTH Aachen University(亚琛工业大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Shandong University(山东大学)
;
National University of Singapore(新加坡国立大学)
;
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Xiaomi Corporation(小米公司)
机构
*
Shenzhen Technology University(深圳技术大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Zhejiang Lab(之江实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
HKUST (Guangzhou)(香港科技大学(广州))
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation
M P V S Gopinadh
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Comments3 pages. Accepted at ACL 2026 Workshop on Evaluation in Practice: Methodological Rigor, Sociotechnical Perspectives, & Community Collaboration (EvalEval)
Comments25 pages, 4 figures, 16 tables, 6 appendices. Code, task suite, released per-run verdicts, and a one-command reproduction of every reported number: https://github.com/shivenkk/agentrelbench
TAF-MED: Multi-Turn Safety Refusal Collapse in LLMs Under Declared Self-Treatment Intent
TAF-MED:声明自我治疗意图下大语言模型的多轮安全拒绝崩溃
Waleed Jamil, Raphael Schmitt
机构
*
Independent Researcher(独立研究者)
;
School of Computation, Information and Technology, Technical University of Munich(慕尼黑工业大学计算、信息与技术学院)
;
Institute of General Practice, Faculty of Medicine and Medical Center, University of Freiburg(弗莱堡大学医学中心医学院普通实践研究所)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
Macau University of Science and Technology(澳门科技大学)
;
Tsinghua University(清华大学)
;
Southeast University(东南大学)
;
FellouAI
;
ARGUS Lab(ARGUS实验室)
;
The University of Edinburgh(爱丁堡大学)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
Comments18 pages, 12 figures, 2 tables. This manuscript has been accepted for publication in Artificial Life and Robotics following peer review
Journal refShota Miyazaki, Takaya Arita and Reiji Suzuki: An evolutionary model of animats with VLM-based subjective evaluation, Artificial Life and Robotics (2026). https://link.springer.com/article/10.1007/s10015-026-01135-4
From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models
从经济智能体到智能体经济:经济世界模型的系统蓝图
Jiale Han, Xiang Li, Jing Qian, Wenyuan Gu, Pin Gao, Ye Luo, Hongyuan Zha, Dacheng Tao, Benyou Wang, Lin William Cong
机构
*
Shenzhen Loop Area Institute(深圳河套学院)
;
School of Data Science, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)数据科学学院)
;
University of Hong Kong(香港大学)
;
Nanyang Technological University(南洋理工大学)