arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2604.15701 2026-04-20 cs.CL 57%

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

通过关键信息的分步注意力混合层蒸馏提升小模型的推理能力

Yao Chen, Jiawei Sheng, Wenyuan Zhang, Tingwen Liu

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出一种基于分步注意力的混合层蒸馏方法,通过引导学生模型逐步聚焦关键信息,提升小模型的推理能力,并在多个数学和常识推理数据集上取得一致性能提升。

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02263 2026-04-20 cs.CV cs.AI 57%

Social-JEPA: Emergent Geometric Isomorphism

Social-JEPA:涌现的几何同构

Haoran Zhang, Youjin Wang, Yi Duan, Rong Fu, Dianyu Zhao, Sicheng Fan, Shuaishuai Cao, Wentao Guo, Xiao Zhou

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 Social-JEPA通过让不同视角的独立代理学习环境模型,发现其潜在空间近似线性同构,从而实现跨代理的透明转换与高效迁移学习。

Comments This preprint is withdrawn due to significant errors in the emergent geometric isomorphism results that necessitate full rewriting, coupled with unresolved author disagreement on authorship. A corrected and revised manuscript will be released separately

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14808 2026-04-17 cs.CL 57%

Modeling LLM Unlearning as an Asymmetric Two-Task Learning Problem

将LLM去学习视为一个非对称的双任务学习问题

Zeguan Xiao, Siqing Li, Yong Wang, Xuetao Wei, Jian Yang, Yun Chen, Guanhua Chen

机构 * Shanghai University of Finance and Economics(上海金融学院) Alibaba Group(阿里巴巴集团) Southern University of Science and Technology(南方科技大学) Beihang University(北航) MoE Key Laboratory of Interdisciplinary Research of Computation and Economics(计算与经济交叉学科研究教育部重点实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文将LLM去学习视为非对称双任务问题,提出优先保留的梯度合成框架,通过分离任务特定梯度提取与冲突感知组合,改进梯度冲突解决方法,实验证明通过重塑梯度几何而非重新平衡损失,有效缓解去学习-保留的权衡。

Comments ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14723 2026-04-17 cs.SE cs.AI 57%

Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution

企业AI的有界自主性:带类型动作合同的消费者端执行

Sarmad Sohail, Ghufran Haider

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文提出一种有界自主性架构,通过类型动作合同和消费者端执行限制大型语言模型的自主性,确保企业系统安全可靠。

Comments 37 pages, 5 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14477 2026-04-17 cs.AI 57%

Seeing Through Circuits: Faithful Mechanistic Interpretability for Vision Transformers

通过电路看见:面向视觉变换器的忠实机制可解释性

Nina Żukowska, Wolfgang Stammer, Bernt Schiele, Jonas Fischer

机构 * Max Planck Institute for Informatics(马克斯·普朗克信息研究所)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文研究了通过计算图在视觉变换器中识别有用机制电路的可能性,提出自动视觉电路发现方法,发现分类特定电路、CLIP中的文字攻击电路以及可引导纠正有害行为的电路。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14459 2026-04-17 cs.CL 57%

Filling in the Mechanisms: How do LMs Learn Filler-Gap Dependencies under Developmental Constraints?

填补机制:语言模型如何在发育约束下学习填充-缺口依赖?

Atrey Desai, Sathvik Nair

机构 * University of Maryland(马里兰大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 研究通过DAS分析语言模型在不同数据量下的填充-缺口依赖表示,发现有限数据下存在共享但敏感的机制,但语言模型仍需更多数据才能达到人类水平,凸显语言特异性偏见的重要性。

Comments To be published in the 64th Annual Meeting of the Association for Computational Linguistics

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14223 2026-04-17 cs.IR cs.AI 57%

TRACE: A Conversational Framework for Sustainable Tourism Recommendation with Agentic Counterfactual Explanations

TRACE:一种用于可持续旅游推荐的对话框架,配备代理反事实解释

Ashmi Banerjee, Adithi Satish, Wolfgang Wörndl, Yashar Deldjoo

机构 * Technical University of Munich(慕尼黑技术大学) Polytechnic University of Bari(巴里理工学院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 TRACE通过多代理架构促进可持续旅游,利用反事实解释和LLM生成问题,提升用户环保意识,实验证明其在推荐质量与交互响应上的有效性。

Journal ref Proceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR '26), July 20--24, 2026, Melbourne, VIC, Australia

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17512 2026-04-17 cs.CL 57%

Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality

按需语言,知识为核心:通过编码器-解码器翻译模型组成LLM以实现可扩展的多语言性

Mengyu Bu, Yang Feng

机构 * Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences(智能信息处理重点实验室,计算技术研究所,中国科学院) State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出XBridge架构,利用预训练翻译模型实现多语言理解和生成,同时保留LLM作为英语核心处理通用知识,通过轻量级跨模型映射层和最优传输对齐目标,提升多语言生成性能。

Comments ACL 2026 Main Conference. Code: https://github.com/ictnlp/XBridge | Models: https://huggingface.co/collections/ICTNLP/xbridge

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23853 2026-04-17 cs.CL 57%

Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception

你的LLM代理是时间盲的:工具使用决策与人类时间感知之间的不一致

Yize Cheng, Arshia Soltani Moakhar, Chenrui Fan, Parsa Hosseini, Kazem Faghih, Zahra Sodagar, Wenxiao Wang, Soheil Feizi

机构 * University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 研究揭示LLM代理在动态环境中因时间感知不足导致的工具调用偏差,通过TicToc数据集分析发现现有模型与人类时间感知对齐率低,提出通过后训练对齐提升多轮对话中工具使用与人类时间感知的一致性。

Comments ACL 2026 (findings), Camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13715 2026-04-16 cs.SD cs.AI 57%

Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt

迈向细粒度时间感知:通过音频侧时间提示进行后训练的大音频-语言模型

Yanfeng Shi, Pengfei Cai, Jun Liu, Qing Gu, Nan Jiang, Lirong Dai, Ian McLoughlin, Yan Song

机构 * National Engineering Research Center of Speech and Language Information Processing, University of Science and Technology of China, Hefei, China(语音与语言信息处理国家级工程研究中心,中国科学技术大学,合肥,中国) ICT Cluster, Singapore Institute of Technology, Singapore(新加坡理工学院ICT集群,新加坡)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出Audio-Side Time Prompt方法,结合强化学习改进大音频-语言模型的时间感知能力,在音频定位、事件检测等任务中取得显著提升。

Comments Submitted to Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13695 2026-04-16 cs.CV cs.AI 57%

Med-CAM: Minimal Evidence for Explaining Medical Decision Making

Med-CAM:为解释医学决策制定提供最小证据

Pirzada Suhail, Aditya Anand, Amit Sethi

机构 * IIT Bombay(印度理工学院班加罗尔)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出Med-CAM框架,通过分类器激活匹配生成最小且清晰的证据图,为医学决策提供解释性证据,优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.13203 2026-04-16 cs.HC cs.AI 57%

Inclusive Kitchen Design for Older Adults: Generative AI Visualizations to Support Mild Cognitive Impairment

包容性厨房设计用于老年人:生成式AI可视化以支持轻度认知障碍

Ibrahim Bilau, Nicole Li, Terrence Malayvong, Eunhwa Yang

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本研究利用生成式AI创建MCI友好的厨房设计,通过训练Stable Diffusion模型提升可视化效果,帮助老年人更易独立生活。

Comments 19 pages, 7 figures, 5 tables, IAFOR Agen2026 Conference Proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12776 2026-04-15 cs.CL 57%

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

EvoSpark:内生交互代理社会的统一长周期叙述进化

Shiyu He, Minchi Kuang, Mengxian Wang, Bin Hu, Tingxiang Gu

机构 * School of Computer Science and Technology, Xinjiang University(新疆大学计算机科学与技术学院) Department of Precision Instrument, Tsinghua University(清华大学精密仪器系)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 EvoSpark通过内生交互代理社会框架,解决LLM多代理系统中长周期叙述演化的矛盾,通过分层叙述记忆和生成场景机制实现逻辑一致的持续叙事。

Comments Accepted to the Main Conference of ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12721 2026-04-15 cs.CL 57%

InsightFlow: LLM-Driven Synthesis of Patient Narratives for Mental Health into Causal Models

InsightFlow: 基于LLM的患者心理健康叙事合成至因果模型

Shreya Gupta, Prottay Kumar Adhikary, Bhavyaa Dave, Salam Michael Singh, Aniket Deroy, Tanmoy Chakraborty

机构 * Department of Mathematics, IIT Delhi(印度德里理工学院数学系) Department of Electrical Engineering, IIT Delhi(印度德里理工学院电气工程系) Department of Computer Science and Engineering, IIIT Manipur(曼尼普尔理工学院计算机科学与工程系)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 InsightFlow利用LLM自动从患者-治疗师对话生成与5P框架对齐的因果图,通过结构、语义和专家评估验证,证明其在临床案例构建中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12663 2026-04-15 cs.AI 57%

Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport

以人为中心的主题建模:基于目标提示对比学习与最优传输

Rui Wang, Yi Zheng, Dongxin Wang, Haiping Huang, Yuanzhi Yao, Yuxiang Zhou, Jialin Yu, Philip Torr

机构 * School of Computer Science, Nanjing University of Posts and Telecommunications(南京邮电大学计算机科学学院) School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) School of Electronic Engineering and Computer Science, Queen Mary University of London(伦敦玛丽女王大学电子工程与计算机科学学院) Department of Engineering Science, University of Oxford(牛津大学工程科学系)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出Human-TM任务框架,结合人类提供的目标进行主题建模,通过GCTM-OT方法提升主题的可解释性、多样性和目标导向性,实验表明其在主题连贯性和多样性上优于现有方法。

Comments 11 Pages, 6 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12537 2026-04-15 cs.CV cs.AI 57%

MODIX: A Training-Free Multimodal Information-Driven Positional Index Scaling for Vision-Language Models

MODIX:一种无需训练的多模态信息驱动的位置索引缩放

Ruoxiang Huang, Zhen Yuan

机构 * Peking University(北京大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 MODIX通过动态调整位置步长,基于模态特定贡献优化多模态模型的位置编码,提升多模态推理能力。

Comments Accepted by CVPR 2026 (Highlight). 10 pages, 2 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12506 2026-04-15 cs.CL cs.SD 57%

Beyond Transcription: Unified Audio Schema for Perception-Aware AudioLLMs

超越转录:面向感知意识的统一音频架构

Linhao Zhang, Yuhan Song, Aiwei Liu, Chuhan Wu, Sijun Zhang, Wei Jia, Yuan Liu, Houfeng Wang, Xiao Zhou

机构 * Basic Model Technology Center, WeChat AI, Tencent Inc.(腾讯基本模型技术中心、微信AI、腾讯公司) State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室、计算机学院、北京大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出统一音频架构(UAS),通过将音频信息划分为转录、语调和非语言事件三个组件,提升音频细粒度感知性能,同时保持推理能力。

Comments Accepted to ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12113 2026-04-15 cs.CV cs.AI 57%

PR-MaGIC: Prompt Refinement Via Mask Decoder Gradient Flow For In-Context Segmentation

PR-MaGIC:通过掩码解码器梯度流进行上下文分割的提示精炼

Minjae Lee, Sungwoo Hur, Soojin Hwang, Won Hwa Kim

机构 * Pohang University of Science and Technology(浦项科学技术大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 PR-MaGIC通过掩码解码器梯度流精炼提示,提升上下文分割性能,无需额外训练,有效缓解提示不足问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11948 2026-04-15 cs.LG cs.AR 57%

Active Imitation Learning for Thermal- and Kernel-Aware LFM Inference on 3D S-NUCA Many-Cores

主动模仿学习用于热感知和核意识的LFM推理在3D S-NUCA多核系统

Yixian Shen, Chaoyao Shen, Jan Deen, George Floros, Andy Pimentel, Anuj Pathania

机构 * University of Amsterdam, Netherlands(阿姆斯特丹大学) Southeast University, China(东南大学) University of Thessaly, Greece(塞萨洛尼基大学)

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出AILFM框架,通过主动模仿学习实现热感知调度,考虑核心性能异质性和LFM核行为,提升性能并保障热安全。

Comments Accepted for publication at the 63rd ACM/IEEE Design Automation Conference (DAC 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14004 2026-04-15 cs.CL 57%

Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models

定位、引导与改进:大型语言模型中可操作机制可解释性的一项实用调查

Hengyuan Zhang, Zhihao Zhang, Mingyang Wang, Zunhai Su, Yiwei Wang, Qianli Wang, Shuzhou Yuan, Ercong Nie, Xufeng Duan, Feijiang Han, Qibo Xue, Zeping Yu, Chenming Shang, Xiao Liang, Jing Xiong, Hui Shen, Chaofan Tao, Zhengwu Liu, Senjie Jin, Zhiheng Xi, Dongdong Zhang, Sophia Ananiadou, Tao Gui, Ruobing Xie, Hayden Kwok-Hay So, Hinrich Schütze, Xuanjing Huang, Qi Zhang, Ngai Wong

机构 * The University of Hong Kong(香港大学) Fudan University(复旦大学) LMU Munich(慕尼黑大学) Tsinghua University(清华大学) Technische Universität Darmstadt(达姆施塔特技术大学) Technische Universität Berlin(柏林技术大学) Technische Universität Dresden(德累斯顿技术大学) The Chinese University of Hong Kong(香港中文大学) University of Pennsylvania(宾夕法尼亚大学) Nanjing University(南京大学) University of Manchester(曼彻斯特大学) Dartmouth College(达特茅斯学院) University of California Los Angeles(加州大学洛杉矶分校) University of Michigan(密歇根大学) Microsoft(微软) Tencent(腾讯)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出一个实用调查,围绕'定位、引导与改进'流程,系统分类定位和引导方法,展示如何通过该框架提升模型对齐、能力和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.09087 2026-04-15 cs.AI 57%

The Stackelberg Speaker: Optimizing Persuasive Communication in Social Deduction Games

Stackelberg发言者:优化社会推断游戏中的说服性沟通

Zhang Zheng, Deheng Ye, Peilin Zhao, Hao Wang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Tencent(腾讯) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出基于Stackelberg竞争的强化学习框架,用于优化社会推断游戏中说服性沟通,通过实验展示其在三种不同游戏中的优越性。

Comments Accepted by ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11773 2026-04-14 cs.LG cond-mat.mtrl-sci cs.CV 57%

Autonomous Diffractometry Enabled by Visual Reinforcement Learning

由视觉强化学习实现的自主衍射测量

J. Oppliger, M. Stifter, A. Rüegg, I. Biało, L. Martinelli, P. G. Freeman, D. Prabhakaran, J. Zhao, Q. Wang, J. Chang

机构 * Jeremiah Horrocks Institute for Mathematics, Physics and Astronomy, University of Central Lancashire(中央兰开夏大学杰里迈亚·霍罗克斯数学、物理与天文研究所) Department of Physics, Clarendon Laboratory, University of Oxford(牛津大学克拉伦登实验室物理系) State Key Laboratory of Surface Physics and Department of Physics, Fudan University(复旦大学表面物理国家重点实验室和物理系) Department of Physics, The Chinese University of Hong Kong(香港中文大学物理系) State Key Laboratory of Quantum Information Technologies and Materials, The Chinese University of Hong Kong(香港中文大学量子信息与材料国家重点实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本文提出一种无需晶体学知识的自主系统,利用强化学习从劳厄衍射图中自动对齐单晶体,提升材料科学自动化实验流程的能力。

Comments 20 pages, 16 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11709 2026-04-14 cs.AI 57%

A Mamba-Based Multimodal Network for Multiscale Blast-Induced Rapid Structural Damage Assessment

基于Mamba的多模态网络用于多尺度爆炸诱导快速结构损伤评估

Wanli Ma, Sivasakthy Selvakumaran, Dain G. Farrimond, Adam A. Dennis, Samuel E. Rigby

机构 * School of Mechanical, Aerospace

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文提出基于Mamba的多模态网络,整合多尺度爆炸载荷信息与光学遥感图像,提升爆炸后结构损伤评估的准确性和速度,通过2020年贝鲁特爆炸案例验证其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11483 2026-04-14 cs.LG q-bio.QM 57%

CAGenMol: Condition-Aware Diffusion Language Model for Goal-Directed Molecular Generation

CAGenMol:面向目标的扩散语言模型用于定向分子生成

Yanting Li, Zhuoyang Jiang, Enyan Dai, Lei Wang, Wen-Cai Ye, Li Liu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Jinan University, Guangzhou(暨南大学(广州))

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出CAGenMol,一种条件感知的离散扩散框架,用于解决分子生成中异构约束冲突问题,通过结合扩散与强化学习提升生成效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11287 2026-04-14 cs.AI q-bio.OT 57%

Consistency of AI-Generated Exercise Prescriptions: A Repeated Generation Study Using a Large Language Model

人工智能生成锻炼处方的一致性:使用大型语言模型的重复生成研究

Kihyuk Lee

机构 * Seoul National University Bundang Hospital(首尔大学盆唐医院)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本研究评估了大型语言模型生成锻炼处方在相同条件下的内在一致性,发现语义一致性高,但关键定量成分存在差异,需进一步结构约束和专家验证。

Comments 15 pages, 5 tables, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.02369 2026-04-14 cs.NI cs.AI 57%

Beyond Message Passing: A Semantic View of Agent Communication Protocols

超越消息传递:代理通信协议的语义视角

Dun Yuan, Fuyuan Lyu, Ye Yuan, Weixu Zhang, Bowei He, Jiayi Geng, Linfeng Du, Zipeng Sun, Yankai Chen, Changjiang Han, Jikun Kang, Xi Chen, Haolun Wu, Xue Liu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文从人类视角出发,将代理通信分为通信、语法和语义三层,分析18种协议在可靠传输、结构交互和语义协调中的不足,提出改进方向,旨在构建更安全、语义稳固的代理生态系统。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10721 2026-04-14 cs.CV cs.AI 57%

Turning Generators into Retrievers: Unlocking MLLMs for Natural Language-Guided Geo-Localization

将生成器转化为检索器:解锁多模态大语言模型用于自然语言引导的地理定位

Yuqi Chen, Xiaohan Zhang, Ahmad Arrabi, Waqas Sultani, Chen Chen, Safwan Wshah

机构 * Vermont Artificial Intelligence Lab, Department of Computer Science, University of Vermont(佛蒙特大学计算机科学系佛蒙特人工智能实验室) Intelligent Machines Lab, Department of Artificial Intelligence, Information Technology University(信息技术大学人工智能系智能机器实验室) Institute of Artificial Intelligence, University of Central Florida(中佛罗里达大学人工智能研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文提出一种简单有效的框架,通过参数高效微调将多模态大语言模型应用于自然语言引导的地理定位,实现强大的跨模态对齐,取得GeoText-1652的SOTA结果,并在CVG-Text的5个子任务中表现优异。

Comments CVPRF

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.10633 2026-04-14 cs.CL 57%

ProUIE: A Macro-to-Micro Progressive Learning Method for LLM-based Universal Information Extraction

ProUIE:一种面向大规模信息提取的宏到微渐进学习方法

Wenda Liu, Zhigang Song, Shuai Nie, Guangyao Liu, Lisung Chen, Binyu Yang, Yaran Chen, Peng Zhou, Hongzhen Wang, Yuchen Liu, Wenyue Hu, Jiaming Xu, Runyu Shi, Ying Huang

机构 * Xiaomi Corporation(小米公司) Xi’an Jiaotong-Liverpool University(西交利物浦大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本文提出ProUIE方法,通过宏、中、微三级阶段提升LLM基于统一信息提取性能,实验表明其在多个数据集上优于基线模型,尤其在大规模生产场景中表现更优。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08339 2026-04-14 cs.CL 57%

What Factors Affect LLMs and RLLMs in Financial Question Answering?

影响LLMs和RLLMs在金融问答中的因素是什么?

Peng Wang, Xuesi Hu, Jiageng Wu, Yuntao Zou, Qiancheng Zhang, Dagang Li

机构 * School of Computer Science and Engineering, Macau University of Science and Technology, China(澳门科技大学计算机科学与工程学院) SKLPlanets, Macau University of Science and Technology, China(澳门科技大学月球与行星科学国家重点实验室) School of Economics, Anhui University, China(安徽大学经济学院) School of Energy and Power Engineering, Huazhong University of Science and Technology, China(华中科技大学能源与动力工程学院)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 研究探讨了提示方法、代理框架和多语言对齐方法对LLMs和RLLMs在金融问答任务中的影响,发现提示方法和代理框架能提升LLMs性能,而RLLMs自身具备Long CoT能力,传统方法对其提升有限。

Comments Accepted by ACL 2026 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13331 2026-04-14 cs.LG 57%

Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization

认知推理的混合:类脑的专业化模块推理

Badr AlKhamissi, C. Nicolò De Sabbata, Greta Tuckute, Zeming Chen, Martin Schrimpf, Antoine Bosselut

机构 * EPFL(瑞士联邦理工学院洛桑) Brain and Cognitive Sciences at MIT(麻省理工学院脑与认知科学系) Kempner Institute at Harvard University(哈佛大学肯普纳研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本文提出MiCRo模型,通过类脑专业化模块实现认知行为,提供可解释的推理模块,支持动态路由并优于基线模型。

Comments ICLR 2026. Project Page at https://cognitive-reasoners.epfl.ch

详情

展开后加载摘要…

URL PDF HTML 收藏