Multi-Dimensional Behavioral Evaluation of Agentic Stock Prediction Systems Using Large Language Model Judges with Closed-Loop Reinforcement Learning Feedback
基于大语言模型判官的多维行为评估:用于代理股票预测系统的闭环强化学习反馈
Mohammad Al Ridhawi, Mahtab Haj Ali, Hussein Al Osman
机构
*
School of Electrical Engineering and Computer Science(电气工程与计算机科学学院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI、cs.LG
Can Small-Scale Data Poisoning Exacerbate Dialect-Linked Biases in Large Language Models?
Chaymaa Abbas, Mariette Awad, Razane Tajeddine
机构
*
Department of Electrical and Computer Engineering, Maroun Semaan Faculty of Engineering and Architecture(电气与计算机工程系,马鲁恩·塞马安工程与建筑学院)
;
American University of Beirut(贝鲁特美国大学)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);instruction tuning(abstract)
机构
*
Institute of Biomedical Engineering, Department of Engineering Science, University of Oxford(生物医学工程研究所,工程科学系,牛津大学)
;
Imperial College London(伦敦帝国学院)
;
University of Sheffield(谢菲尔德大学)
;
Novo Nordisk Research Centre Oxford (NNRCO)(牛津诺和硕研究中心(NNRCO))
;
Royal Society(皇家学会)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);pretraining(abstract)
Can Large Language Models Understand Symbolic Graphics Programs?
Zeju Qiu, Weiyang Liu, Haiwen Feng, Zhen Liu, Tim Z. Xiao, Katherine M. Collins, Joshua B. Tenenbaum, Adrian Weller, Michael J. Black, Bernhard Schölkopf
机构
*
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
;
University of Cambridge(剑桥大学)
;
MIT(麻省理工学院)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);instruction tuning(abstract)
A dataset and benchmark for hospital course summarization with adapted large language models
Asad Aali, Dave Van Veen, Yamin Ishraq Arefeen, Jason Hom, Christian Bluethgen, Eduardo Pontes Reis, Sergios Gatidis, Namuun Clifford, Joseph Daws, Arash S. Tehrani, Jangwon Kim, Akshay S. Chaudhari
机构
*
Stanford University(斯坦福大学)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
专题命中
评测与基准
:large language model(title,abstract);language model(title,abstract);LLM(abstract);prompting(abstract)