Comments46 pages, 3 figures + This paper proposes an LLM-based multi-agent system (MAS) for automated evaluation of new product concepts, incorporating retrieval-augmented generation (RAG) and cross-functional virtual agents to assess technical and market feasibility
AnchorBench: A Multi-Pathway Benchmark for the Anchoring Effect in LLMs
AnchorBench:针对大语言模型中锚定效应的多路径基准测试
Yiderigun Borjigin, Alexander Hermann, Christian Cyron, Roland Aydin
机构
*
Saarland University(萨尔大学)
;
Hamburg University of Technology(汉堡工业大学)
;
Helmholtz-Zentrum Hereon(亥姆霍兹中心赫伦)
;
German Research Centre for Artificial Intelligence (DFKI)(德国人工智能研究中心(DFKI))
Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles
告知、指导、共情、倾听:审计LLM护理支持角色
Drishti Goel, Agam Goyal, Veda Duddu, Olivia Pal, Jeongah Lee, Qiuyue Joy Zhong, Violeta J. Rodriguez, Daniel S. Brown, Dong Whi Yoo, Ravi Karkar, Koustuv Saha
机构
*
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
;
OSF HealthCare(OSF医疗集团)
;
Indiana University Indianapolis(印第安纳大学印第安纳波利斯分校)