arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Beihang University(北京航空航天大学)

2025-11-25 至 2025-11-25 共收录 8
2511.17405 2025-11-25 cs.CL cs.AI

Beyond Multiple Choice: Verifiable OpenQA for Robust Vision-Language RFT

超越多项选择:可验证的开放问答用于鲁棒的视觉-语言 RFT

Yesheng Liu, Hao Li, Haiyu Xu, Baoqi Pei, Jiahao Wang, Mingxuan Zhao, Jingshu Zheng, Zheqi He, JG Yao, Bowen Qin, Xi Yang, Jiajun Zhang

机构 * Institute of Automation, CAS(中国科学院自动化研究所) School of Artificial Intelligence, UCAS(中国科学技术大学人工智能学院) BAAI FlagEval Team(百度AI旗评团队) BUAA(北京航空航天大学) PKU(北京大学) ZJU(浙江大学)

AI总结 ReVeL通过重写和验证多项选择问题为开放式问题,提升视觉-语言模型在鲁棒性与数据效率上的表现。

Comments Project url: https://flageval-baai.github.io/ReVeL/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01421 2025-11-25 cs.CV

InfoScale: Unleashing Training-free Variable-scaled Image Generation via Effective Utilization of Information

InfoScale: 通过有效利用信息实现免训练可变尺度图像生成

Guohui Zhang, Jiangtong Tan, Linjiang Huang, Zhonghang Yuan, Mingde Yao, Jie Huang, Feng Zhao

机构 * USTC(中国科学技术大学) Beihang University(北京航空航天大学) CUHK MMLab(香港中文大学多媒体实验室) Kuaishou Technology(快手科技)

AI总结 InfoScale通过有效利用信息解决扩散模型在可变尺度图像生成中的信息丢失、聚合不灵活和分布不匹配问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17265 2025-11-25 cs.CL cs.CV

Systematic Reward Gap Optimization for Mitigating VLM Hallucinations

系统性奖励缺口优化以缓解视觉语言模型的幻觉

Lehan He, Zeren Chen, Zhelun Shi, Tianyu Yu, Jing Shao, Lu Sheng

机构 * School of Software, Beihang University(北京航空航天大学软件学院) Shanghai Innovation Institute(上海创新研究院) Shanghai AI Laboratory(上海人工智能实验室) Tsinghua University(清华大学)

AI总结 TPR通过主题级偏好重写系统优化奖励缺口配置,显著减少VLM幻觉并提升对齐效果。

Comments 34 pages, 12 figures, Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18385 2025-11-25 cs.CV cs.AI

Can a Second-View Image Be a Language? Geometric and Semantic Cross-Modal Reasoning for X-ray Prohibited Item Detection

第二视角图像可以成为一种语言吗?用于X射线禁止物品检测的几何和语义跨模态推理

Chuang Peng, Renshuai Tao, Zhongwei Ren, Xianglong Liu, Yunchao Wei

机构 * Institute of Information Science, Beijing Jiaotong University(北京交通大学信息科学学院) State Key Laboratory of Complex & Critical Software Environment, Beihang University(北航复杂与关键软件环境国家重点实验室)

AI总结 本文提出GSR模型,通过几何与语义跨模态推理,利用双视角图像作为语言模态,提升X射线禁止物品检测的准确性。

Comments 10 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02354 2025-11-25 cs.LG

Evolving Graph Learning for Out-of-Distribution Generalization in Non-stationary Environments

演化图学习用于非平稳环境中的分布外泛化

Qingyun Sun, Jiayi Luo, Haonan Yuan, Xingcheng Fu, Hao Peng, Jianxin Li, Philip S. Yu

机构 * Beijing Advanced Innovation Center for Big Data and Brain Computing, School of Computer Science and Engineering, Beihang University(北京大数据与脑计算先进创新中心,计算机科学与工程学院,北京航空航天大学) Key Lab of Education Blockchain and Intelligent Technology, Ministry of Education, Guangxi Normal University(教育区块链与智能技术重点实验室,教育部,广西师范大学) Department of Computer Science, University of Illinois at Chicago(计算机科学系,伊利诺伊大学芝加哥分校)

AI总结 本文提出EvoOOD框架,通过环境感知的不变模式识别提升动态图在非平稳环境中的分布外泛化能力。

Comments Accepted by TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26451 2025-11-25 cs.LG cs.AI

Robust Graph Condensation via Classification Complexity Mitigation

通过分类复杂性缓解实现鲁棒图压缩

Jiayi Luo, Qingyun Sun, Beining Yang, Haonan Yuan, Xingcheng Fu, Yanbiao Ma, Jianxin Li, Philip S. Yu

机构 * SKLCCSE, School of Computer Science and Engineering, Beihang University(北京航空航天大学信息与电子技术学院) Laboratory for Foundations of Computer Science, University of Edinburgh(爱丁堡大学计算机科学基础实验室) Key Lab of Education Blockchain and Intelligent Technology, Guangxi Normal University(广西师范大学教育区块链与智能技术重点实验室) Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学光荣人工智能学院) Department of Computer Science, University of Illinois, Chicago(伊利诺伊大学芝加哥分校计算机科学系)

AI总结 本文提出MRGC框架,通过引入流形约束模块,提升图压缩在对抗攻击下的鲁棒性。

Comments Accepted by Neurips 2025 (Spotlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23035 2025-11-25 cs.CV

FreeInv: Free Lunch for Improving DDIM Inversion

FreeInv: 为改进DDIM反向过程提供免费午餐

Yuxiang Bao, Huijie Liu, Xun Gao, Huan Fu, Guoliang Kang

机构 * Beihang University(北航大学) HUJING Digital Media & Entertainment Group(HUJING数字媒体与娱乐集团)

AI总结 FreeInv通过随机变换潜变量表示,有效解决DDIM反向过程中的轨迹偏移问题,提升视频反向的保真度和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22688 2025-11-25 cs.SE cs.AI cs.PL

CodeIF-Bench: Evaluating Instruction-Following Capabilities of Large Language Models in Interactive Code Generation

CodeIF-Bench: 评估大型语言模型在交互式代码生成中的指令遵循能力

Peiding Wang, Li Zhang, Fang Liu, Lin Shi, Minxiao Li, Bo Shen, An Fu

机构 * School of Computer Science & Engineering, State Key Laboratory of Complex & Critical Software Environment, Beihang University(计算机科学与工程学院、复杂与关键软件环境国家重点实验室、北京航空航天大学) School of Software, Beihang University(软件学院、北京航空航天大学) Huawei Cloud Computing Technologies Co., Ltd. China(华为云计算技术有限公司)

AI总结 CodeIF-Bench通过九种可验证指令评估LLMs在多轮交互中的指令遵循能力,揭示了上下文管理对性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏