机构
*
Institute of Automation, Chinese Academy of Sciences (CAS)(中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
CASI Vision Technology Co., Ltd.(中科慧远视觉技术有限公司)
;
Shandong Laboratory of Aluminum Advanced Manufacturing in Binzhou (SLAAMB), Binzhou Institute of Technology, Weiqiao-UCAS Science and Technology Park(山东省滨州市铝先进制造实验室(SLAAMB),滨州技术学院,魏桥国科科技园)
;
Space Information Research Institute, Hangzhou Dianzi University(杭州电子科技大学空间信息研究院)
;
School of Software, Tsinghua University(清华大学软件学院)
专题命中
后训练与偏好优化
:language model(title,abstract);large language model(abstract)
IRPO: Boosting Image Restoration via Post-training GRPO
IRPO:通过后训练GRPO提升图像恢复
Haoxuan Xu, Yi Liu, Tianfu Li, Ruolin Shen, Boyuan Jiang, Jinlong Peng, Donghao Luo, Xiaobin Hu, Shuicheng Yan, Haoang Li
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Tsinghua University(清华大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Zhejiang University(浙江大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Fudan University(复旦大学)
;
National University of Singapore(新加坡国立大学)
Contextualized Visual Personalization in Vision-Language Models
基于上下文的视觉个性化在视觉-语言模型中
Yeongtak Oh, Sangwon Yu, Junsung Park, Han Cheol Moon, Jisoo Mok, Sungroh Yoon
机构
*
Department of Electrical and Computer Engineering, Seoul National University, Seoul, South Korea(电气电子工程系,首尔国立大学,首尔,韩国)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University, Seoul, Korea(人工智能交叉学科项目,首尔国立大学,首尔,韩国)
Offline Preference Optimization for Rectified Flow with Noise-Tracked Pairs
离线偏好优化用于具有噪声跟踪配对的校正流
Yunhong Lu, Qichao Wang, Hengyuan Cao, Xiaoyin Xu, Min Zhang
机构
*
Zhejiang University(浙江大学)
;
Shanghai Institute for Advanced Study-Zhejiang University(上海先进研究院-浙江大学)
;
Shanghai Institute for Mathematics and Interdisciplinary Sciences(上海数学与交叉科学研究院)
机构
*
School of Computer Science and Technology, Dalian University of Technology(大连理工大学计算机科学与技术学院)
;
Key Laboratory of Social Computing and Cognitive Intelligence (Dalian University of Technology), Ministry of Education(社会计算与认知智能重点实验室(大连理工大学),教育部)
;
Institute of Image Processing and Understanding, North Minzu University(北华大学图像处理与理解研究所)
World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training
World-Env: 利用世界模型作为虚拟环境进行VLA后训练
Junjin Xiao, Yandan Yang, Xinyuan Chang, Ronghan Chen, Feng Xiong, Mu Xu, Wei-Shi Zheng, Qing Zhang
机构
*
School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院)
;
AMap, Alibaba Group(阿里巴巴集团高德地图)
;
Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部机器智能与先进计算重点实验室)
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Tencent Hunyuan(腾讯文元)
;
Hong Kong University of Science and Technology(香港科技大学)
MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
MARE: 多模态对齐与强化学习用于通过视觉-语言模型的可解释深度伪造检测
Wenbo Xu, Wei Lu, Xiangyang Luo, Jiantao Zhou
机构
*
School of Computer Science and Engineering, MoE Key Laboratory of Information Technology, Guangdong Province Key Laboratory of Information Security Technology, Sun Yat-sen University, Guangzhou 510006, China(计算机科学与工程学院,信息技术MOE实验室,广东省信息安全技术重点实验室,中山大学,广州510006,中国)
;
State Key Laboratory of Mathematical Engineering and Advanced Computing(数学工程与先进计算国家重点实验室)
;
Department of Computer and Information Science, University of Macau.(计算机与信息科学系,澳门大学)