CommentsInitial controlled diagnostic study on 23 natural drawing sets and three VLMs; broader model, building, repeated-inference, and human coverage is planned for a subsequent version
Rong Fu, Yongtai Liu, Xiaowen Ma, Haoyu Zhao, Shuo Yin, Yiqing Lyu, Long Zhang, Wangyu Wu
机构
*
University of Macau(澳门大学)
;
Hanyang University(汉阳大学)
;
Zhejiang University(浙江大学)
;
Wuhan University(武汉大学)
;
Tsinghua University(清华大学)
;
South China University of Technology(华南理工大学)
;
University of Liverpool(利物浦大学)
ST-LoRA: Single Trajectory LoRA Ensemble for Uncertainty Aware Agricultural Segmentation
ST-LoRA:用于不确定性感知农业分割的单轨迹LoRA集成
Mohamed Farag, Genc Hoxha, Yahia Maleki, Chris McCool, Ribana Roscher
机构
*
University of Bonn(波恩大学)
;
Lamarr Institute for Machine Learning and Artificial Intelligence, University of Bonn(波恩大学拉马尔机器学习与人工智能研究所)
;
Commonwealth Scientific and Industrial Research Organisation (CSIRO)(英联邦科学与工业研究组织(CSIRO))
机构
*
Department of Data Science and Hong Kong Institute of AI for Science, City University of Hong Kong(数据科学系和香港人工智能科学研究所,香港城市大学)
;
Li Auto Inc., China(中国利汽车公司)
;
Department of Statistics, University of Oxford(统计系,牛津大学)
专题命中
后训练与偏好优化
:large language model(title,abstract_cn);language model(title,abstract_cn);LLM(abstract,abstract_cn);分类 cs.CL
Tail-Aware Information-Theoretic Bounds for LLM Alignment under Heavy-Tailed Rewards
尾感知信息论泛化用于RLHF和SGLD
Huiming Zhang, Binghan Li, Wan Tian, Qiang Sun
机构
*
Institute of Artificial Intelligence, Beihang University(北京航空航天大学人工智能研究院)
;
Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing(北京未来区块链与隐私计算高精尖创新中心)
;
Advanced Institute of Information Technology, Peking University(北京大学信息技术高等研究院)
;
Wangxuan Institute of Computer Technology, Peking University(北京大学王选计算机技术研究所)
;
Computer and Mathematical Sciences, Computer Science, and Statistics, University of Toronto(多伦多大学计算机与数学科学、计算机科学和统计学系)
;
MBZUAI(穆罕默德·本·扎耶德人工智能大学)
专题命中
后训练与偏好优化
:RLHF(title_cn,abstract);LLM(title);large language model(abstract);language model(abstract)
TLA-Prover: Verifiable TLA+ Specification Synthesis via Preference-Optimized Low-Rank Adaptation
TLA-Prover: 通过偏好优化低秩适配实现可验证的 TLA+ 规范合成
Eric Spencer, Arslan Bisharat, Brian Ortiz, Khushboo Bhadauria, Mujtaba Nazari, TaiNing Wang, George K. Thiruvathukal, Konstantin Laufer, Mohammed Abuhamad
机构
*
Department of Computer Science, Loyola University Chicago(洛约拉芝加哥大学计算机科学系)
专题命中
后训练与偏好优化
:SFT(abstract,abstract_cn);LLM(abstract_cn);large language model(abstract);language model(abstract)
机构
*
Institute of Science Tokyo(东京科学研究所)
;
MBZUAI
;
National Institute of Advanced Industrial Science and Technology(国家先进工业科学与技术研究院)
;
NII LLMC(日本信息基础设施中心LLMC)