arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8057 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8057 篇

2606.23724 2026-06-24 cs.IR cs.CL cs.HC 新提交 57%

EvidenceLens: A Claim-Evidence Matrix for Auditing Financial Question Answering

EvidenceLens: 用于审计金融问答的声明-证据矩阵

Fengchen Gu, Xiaotian Ren, Zhengyong Jiang, Zhilu Zhang, Ángel F. García-Fernández, Angelos Stefanidis, Mian Zhou, Huakang Li, Jionglong Su

机构 * 1 School of AI Advanced Computing, XJTLU Entrepreneur College (Taicang), Xi’an Jiaotong-Liverpool University, Suzhou, Jiangsu, China 2 ETSI de Telecomunicaci\' o n, Universidad Polit\' e cnica de Madrid, Madrid, Spain IEEE VIS 2026 conditionally accepted version

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出EvidenceLens,一种将金融问答视为声明-证据对齐问题的可视化分析工具,通过多模态声明-证据矩阵揭示覆盖、矛盾与模态不平衡,帮助分析师区分有依据的声明与过度自信的综合。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23566 2026-06-24 cs.CL 新提交 57%

LangMAP: A Language-Adaptive Approach to Tokenization

LangMAP:一种语言自适应的分词方法

Clara Meister, Suchir Salhan, Andrzej Szablewski, Pietro Lesci, Paula Buttery, Tiago Pimentel

机构 * EPFL(瑞士联邦理工学院洛桑) University of Cambridge(剑桥大学) ETH Zürich(苏黎世联邦理工学院)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出LangMAP分词方案,扩展UnigramLM到多语言设置,从共享词汇表生成语言特定分词,无需改变词汇表即可适配预训练模型,在14个分词器、9种自然语言和9种编程语言上提升了形态边界对齐和AST叶边界对齐。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20189 2026-06-24 cs.CV cs.AI cs.RO 新提交 57%

HilDA: Hierarchical Distillation with Diffusion for Advancing Self-Supervised LiDAR Pre-training

HilDA:利用扩散的分层蒸馏推进自监督LiDAR预训练

Maciej Wozniak, Jesper Ericsson, Hariprasath Govindarajan, Truls Nyberg, Thomas Gustafsson, Patric Jensfelt, Olov Andersson

机构 * KTH Royal Institute of Technology(瑞典皇家理工学院) Linköping University(林雪平大学) TRATON AB(TRATON公司) Qualcomm Auto Ltd Sweden Filial(高通汽车有限公司瑞典分公司)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 提出HilDA框架,通过分层蒸馏(多层蒸馏和全局上下文蒸馏)结合时间占用扩散目标,自监督预训练LiDAR骨干网络,在3D检测、场景流和语义占用预测任务上达到最先进水平。

Comments Accepted to ECCV 2026. Maciej and Jesper contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05174 2026-06-24 cs.MA cs.AI 57%

Emergent Coordination in Multi-Agent Language Models

多智能体语言模型中的涌现协调

Christoph Riedl

机构 * D’Amore-McKim School of Business(达莫-麦克金商学院) Khoury College of Computer Sciences(科里学院计算机科学学院) Network Science Institute(网络科学研究所) Northeastern University(东北大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 研究探讨多智能体系统是否形成更高层次结构,提出信息论框架通过数据驱动方法检测动态涌现,验证身份差异化与目标互补性,揭示交互模式与集体智慧原理的关联。

Journal ref International Conference on Learning Representations (ICLR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00208 2026-06-24 cs.LG 版本更新 57%

Similarity of Neural Network Representations in Superposition

神经网络表示在叠加中的相似性

Sunny Liu, Habon Issa, André Longon, Liv Gorton, Meenakshi Khosla, Alex Williams, David Klindt

机构 * Cold Spring Harbor Laboratory(冷泉港实验室) UC San Diego(加州大学圣地亚哥分校) Anthropic(Anthropic公司) New York University Flatiron Institute(纽约大学Flatiron研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 研究线性对齐度量在神经网络叠加表示中的失效问题,通过理论推导和稀疏自编码器实验证明对齐度量受投影Gram矩阵影响,并展示基于恢复潜在特征的度量能正确反映特征共享。

Comments 17 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22382 2026-06-23 eess.IV cs.AI cs.CV 新提交 57%

Large Language Model-Assisted Cleaning of Report-Derived Labels in a Large-Scale Chest CT Dataset

大型胸部CT数据集中基于大语言模型的报告衍生标签清洗

Yosuke Yamagishi, Atsushi Takamatsu, Mototsugu Sato, Tomohiro Kikuchi, Shouhei Hanaoka, Takeharu Yoshikawa, Osamu Abe

机构 * Division of Radiology and Biomedical Engineering, Graduate School of Medicine, The University of Tokyo(放射医学与生物医学工程系,东京大学医学研究生院) Department of Computational Diagnostic Radiology and Preventive Medicine, The University of Tokyo Hospital(计算诊断放射学与预防医学系,东京大学医院) Department of Radiology, Kanazawa University Hospital(金泽大学医院放射科) Faculty of Medicine, The University of Tokyo(东京大学医学系) Department of Radiology, School of Medicine, Jichi Medical University(立命馆大学医学系放射科) Department of Radiology, The University of Tokyo Hospital(东京大学医院放射科)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本研究利用大语言模型(GPT-5.4)清洗CT-RATE数据集中的标签-报告不一致性,通过放射科医生裁决验证,发现LLM辅助清洗能有效识别临床相关错误,并提升数据集质量。

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.23394 2026-06-23 cs.CL 新提交 57%

Do LLM Embedding Spaces Recover Expert Structure?

LLM嵌入空间是否恢复了专家结构?

Yixuan Zhu, Zhenke Duan, Fanghen Li

机构 * Zhongnan University of Economics and Law(中南财经政法大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 研究LLM嵌入空间能否恢复心理健康领域的专家定义类别结构,通过对比预训练和微调Qwen3嵌入,发现对齐程度依赖于粒度级别且需控制混杂因素。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22495 2026-06-23 cs.AI 新提交 57%

Grounded Scaling: Why Agentic AI Needs Deterministic Environments

Grounded Scaling: 为什么代理型AI需要确定性环境

Liang Ding, Xintong Wang

机构 * Alibaba Group(阿里巴巴集团)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 本文论证环境确定性是影响长链代理任务成功的关键因素,提出确定性-效率界限、验证者-古德哈特下限等理论,并构建供应确定性指数和确定性成熟度模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21882 2026-06-23 cs.SD cs.AI 新提交 57%

Streaming T5-based Text-to-Speech Synthesis with Limited Lookahead

基于有限前瞻的流式T5文本到语音合成

Muyang Du, Jason Roche, Junjie Lai

机构 * NVIDIAChina(英伟达中国) NVIDIAUSA(英伟达美国)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 提出S5-TTS,一种基于T5的流式变体,通过编码器-解码器语言模型和单调对齐学习实现逐词增量语音合成,并引入前瞻因果掩码和交错多源蒸馏以在低延迟下保持质量。

Comments 6 pages, 1 figure, 4 tables, Interspeech 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20761 2026-06-23 cs.SE cs.AI cs.ET cs.MA cs.SY eess.SY 新提交 57%

Integrating Large Language Model Agents with Digital Twins for Industrial Autonomous Systems

将大语言模型代理与数字孪生集成用于工业自主系统

Yuchen Xia

机构 * Institut für Automatisierungs- und Softwaresysteme (IAS) der Universität Stuttgart(自动化与软件系统研究所(IAS)的斯图加特大学)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 提出三层框架集成大语言模型、数字孪生与自动化系统,通过TPSR模型和四种LLM角色实现自适应任务规划与执行,提升工业自动化适应性。

Comments Doctoral Dissertation, University of Stuttgart. Doctoral Exam Video Recording: https://youtu.be/Mhd9LiV5TKE

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20707 2026-06-23 cs.CV cs.AI 新提交 57%

GEOPHYS: The Geometry of Physical Plausibility

GEOPHYS: 物理合理性的几何学

Christian Internò, Alexander Pondaven, Habon Issa, Fabio Pizzati, Francesco Pinto, Markus Olhofer, Ivan Laptev, Philip Torr, Eero P. Simoncelli, Barbara Hammer, David Klindt

机构 * Bielefeld University(比勒费尔德大学) University of Oxford(牛津大学) Cold Spring Harbor Laboratory(冷泉港实验室) MBZUAI(穆罕默德·本·扎耶德人工智能大学) Independent(独立作者) Honda Research Institute EU(本田欧洲研究院) New York University(纽约大学) Flatiron Institute, Simons Foundation(西蒙斯基金会熨斗研究所)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 提出GEOPHYS方法,利用冻结图像编码器的每帧嵌入的五个几何特征检测视频中的物理不合理性,在物理违反检测上达到SOTA,并作为验证器提升视频生成模型的物理一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20598 2026-06-23 cs.HC cs.AI 新提交 57%

Using Biometrics to Understand AI-Assisted Coding Performance and its Perception

使用生物特征理解AI辅助编程性能及其感知

Paolo Burelli, Fabio Calefato, Daniela Grassi, Mihaela Yurieva Hristova, Nicole Novielli, Alberto Antonio, Romano, Paolo Tell

机构 * IT University of Copenhagen(丹麦技术大学) University of Bari(巴里大学)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 通过脑电图、眼动追踪等生物特征数据,发现AI辅助编程与无辅助编程在认知过程上存在显著差异,且主观感知与客观测量不一致。

Comments Stage 2 RR under review at EMSE. The accepted Stage 1 protocol is publicly archived on OSF (DOI: 10.31219/osf.io/mex39_v2)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.22859 2026-06-23 cs.AI astro-ph.IM physics.soc-ph 新提交 57%

AI Scientists as Engines of Discovery: A Case for Development within Reformed Institutions

作为发现引擎的AI科学家:在改革机构内发展的案例

Raul Jimenez, Boris Bolliet, Francisco Villaescusa-Navarro, Rabih Zbib, Benjamin Wandelt, David N. Spergel, Thomas Meier, Jessica Montgomery, Hana Aliee, Licia Verde

机构 * Institute of Cosmos Sciences (ICCUB), University of Barcelona(巴塞罗那大学宇宙科学研究所) ICREA Cavendish Astrophysics, University of Cambridge(剑桥大学卡文迪什天体物理学) Kavli Institute for Cosmology, University of Cambridge(剑桥大学卡维里宇宙学研究所) Center for Computational Astrophysics, Flatiron Institute(弗拉蒂隆研究所计算天体物理中心) Department of Astrophysical Sciences, Princeton University(普林斯顿大学天体物理科学系) Avature Department of Physics and Astronomy, Johns Hopkins University(约翰霍普金斯大学物理与天文学系) Department of Applied Mathematics and Statistics, Johns Hopkins University(约翰霍普金斯大学应用数学与统计学系) Flatiron Institute(弗拉蒂隆研究所) Munich Center for Machine Learning, LMU Munich(慕尼黑大学慕尼黑机器学习中心) Department of Computer Science and Technology, University of Cambridge(剑桥大学计算机科学与技术系) School of Clinical Medicine, University of Cambridge(剑桥大学临床医学院)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文论证多智能体AI系统将从被动工具进化为“AI科学家”,通过原型框架Denario加速发现周期,并提出机构需为验证、问责、可解释性和双重用途安全进行改革。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.19309 2026-06-23 cs.LG eess.SP 57%

Time-Vertex Machine Learning for Optimal Sensor Placement in Temporal Graph Signals: Applications in Structural Health Monitoring

时间顶点机器学习用于时间图信号中的最优传感器布置:在结构健康监测中的应用

Keivan Faghih Niresi, Jun Qing, Mengjie Zhao, Olga Fink

机构 * Intelligent Maintenance and Operations Systems Lab., EPFL, Lausanne, Switzerland(智能维护与运营系统实验室,瑞士联邦理工学院,洛桑)

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 本文提出时间顶点机器学习框架,结合图信号处理、时域分析和机器学习,实现可解释且高效的传感器布置,通过识别代表性节点减少冗余并保留关键信息。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03818 2026-06-23 cs.CL 版本更新 57%

Empirical Prompt Engineering for Construct Identification with Large Language Models

改善人机编码对齐:心理学构念识别中提示工程的实证评估

Kylie L. Anglin, Stephanie Milan, Brittney Hernandez, Claudia Ventura

机构 * Department of Educational Psychology, Neag School of Education, University of Connecticut(教育心理学系,教育学院,康涅狄格大学) Department of Psychological Sciences, College of Liberal Arts and Sciences, University of Connecticut(心理学系,文理学院,康涅狄格大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 本研究提出一个实证框架,通过提示工程优化大语言模型在心理学文本中识别构念的性能。实验评估五种提示策略,发现构念定义和任务框架最关键,结合代码簿引导和自动提示工程的少样本方法最接近专家判断。

Comments 22 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00116 2026-06-23 q-bio.NC cs.AI 版本更新 57%

Meta-learning ecological priors from large language models explains human learning and decision making

从大型语言模型中元学习生态先验解释人类学习与决策

Akshay K. Jagadish, Mirko Thalmann, Julian Coda-Forno, Marcel Binz, Eric Schulz

机构 * Institute for Human-Centered AI, Helmholtz Computational Health Center(以人为本的人工智能研究所,海德堡计算健康中心) Computational Principles of Intelligence, Max Planck Institute for Biological Cybernetics(计算智能原理,马克斯·普朗克生物 cybernetics 研究院) Eberhard Karls University of Tübingen(图宾根大学) Princeton AI Lab, Princeton University(普林斯顿大学人工智能实验室)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 提出生态理性分析框架,利用大语言模型生成生态有效任务,通过元学习得到ERMI算法,该算法内化自然问题空间的统计规律,灵活适应新情境,在15个实验中优于多个认知模型,表明人类认知可能反映对日常问题生态结构的适应性对齐。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19989 2026-06-19 cs.DC cs.LG 新提交 57%

Online Dynamic Batching with Formal Guarantees for LLM Training

面向LLM训练的具有形式保证的在线动态批处理

Dian Li, Zekun Wang, Yaoru Wang, Jiahong Yan

机构 * Tencent(腾讯)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 提出在线动态批处理(ODB)系统,在数据加载器侧将批构建延迟到样本真实成本可观测时,解决离线批采样中预处理成本不可见问题,实现1.58-4.43x吞吐量提升,并提供无死锁有界终止的形式化保证。

Comments 29 pages, 3 figures, 21 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02885 2026-06-19 cs.CL 版本更新 57%

Med-R2: Perception and Reflection-driven Complex Reasoning for Medical Report Generation

Med-R2:面向医学报告生成的感知与反思驱动复杂推理

Hao Wang, Shuchang Ye, Jinghao Lin, Usman Naseem, Jinman Kim

机构 * The School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) The School of Computing, Macquarie University(麦考瑞大学计算机学院) Doubao Medical Group, ByteDance(字节跳动 doubao 医疗集团)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出Med-R2微调策略,通过引入感知驱动的长推理过程和放射学知识指导,并加入反思机制修正感知错误,提升LVLMs在医学报告生成中的病理特征感知和诊断准确性。

Comments 28 pages, 3 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17386 2026-06-19 eess.SY cs.LG cs.SY 版本更新 57%

A graph neural network surrogate model for mesh-based crashworthiness prediction of vehicle panel components

基于图神经网络的网格级车辆面板部件耐撞性预测代理模型

Haoran Li, Yingxue Zhao, Haosu Zhou, Tobias Pfaff, Nan Li

机构 * Dyson School of Design Engineering, Imperial College London(迪森设计工程学院,帝国理工学院伦敦分校) NVIDIA

专题命中 其他安全 :safety(abstract);分类 cs.LG

AI总结 提出递归图U-Net (ReGUNet) 代理模型,通过图表示有限元网格,结合层次架构和递归机制,高效准确预测车辆B柱等面板部件的动态变形和耐撞性指标。

Comments Accepted manuscript version. Final published version available in Results in Engineering via DOI: 10.1016/j.rineng.2026.110925

Journal ref Results in Engineering 30 (2026) 110925

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18263 2026-06-18 cs.HC cs.AI 新提交 57%

How Well Do Large Language Models Capture Human Personality?

大型语言模型在多大程度上捕捉人类个性?

Aanisha Bhattacharyya, Yaman Kumar Singla, Rajiv Ratn Shah, Changyou Chen, Jitendra Ajmera

机构 * Adobe Media and Data Science Research (MDSR)(Adobe媒体与数据科学研究院)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 研究通过形式化假设并系统评估,发现增加角色描述复杂性会导致表征和行为多样性收缩(角色流形坍缩),简单年龄-性别角色比丰富描述更准确。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.15504 2026-06-18 cs.AI 新提交 57%

Toward Vibe Medicine: A Self-Evolving Multi-Agent Framework for Clinical Decision Support

迈向振动医学:一种用于临床决策支持的自演化多智能体框架

Qianxue Zhang, Yiming Ren, Shihuan Qin, Xiao Zhang, Liao Zhang, Jinyang Huang, Zhengliang Liu, Chenbin Liu, Hongying Feng, Jingyuan Chen, Yuzhen Ding, Weihang You, Hanqi Jiang, Yi Pan, Yifan Zhou, Junhao Chen, Lifeng Chen, Wei Liu, Tianming Liu, Zengren Zhao, Lian Zhang

机构 * Medical AI Lab, The First Hospital of Hebei Medical University(河北医科大学第一医院医学人工智能实验室) Hebei Provincial Engineering Research Center for AI-Based Cancer Treatment Decision-Making, The First Hospital of Hebei Medical University(河北省人工智能癌症治疗决策工程研究中心,河北医科大学第一医院) State Key Laboratory of Neurology and Oncology Drug Development(神经与肿瘤药物研发国家重点实验室) School of Computing, University of Georgia(佐治亚大学计算学院) Department of Radiation Oncology, National Cancer Center/National Clinical Research Center for Cancer/Cancer Hospital and Shenzhen Hospital, Chinese Academy of Medical Sciences and Peking Union Medical College(中国医学科学院北京协和医学院国家癌症中心/国家肿瘤临床医学研究中心/肿瘤医院深圳医院放射治疗科) Department of Radiation Oncology, Mayo Clinic(梅奥诊所放射肿瘤科) College of Mechanical and Power Engineering, China Three Gorges University(三峡大学机械与动力工程学院) Department of Radiation Oncology, Guangzhou Concord Cancer Center(广州康华肿瘤中心放射治疗科) Gastrointestinal Disease Diagnosis and Treatment Center, The First Hospital of Hebei Medical University(河北医科大学第一医院胃肠疾病诊疗中心) Department of General Surgery, The First Hospital of Hebei Medical University(河北医科大学第一医院普通外科)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 提出VIBEMed多智能体框架,通过自演化机制和架构级安全沙箱,从交互历史中动态学习,实现个性化临床决策支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01697 2026-06-18 cs.CL 版本更新 57%

RCEM: Robust Conversational Search EMbedder in Distributional Shift

RCEM:配备查询重写技能的嵌入器,用于分布偏移下的鲁棒对话搜索

Kilho Son, Paul Hsu, Cha Zhang, Dinei Florencio

机构 * Microsoft(微软)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出RCEM模型,通过将LLM的查询重写能力蒸馏到嵌入模型中,实现无需显式重写的上下文感知检索,在分布偏移下提升鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25929 2026-06-18 cs.MA cs.LG 版本更新 57%

Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?

多智能体系统是专家混合:谁成为影响者?

Franka Bause, Jonas Niederle, Martin Pawelczyk, Rebekka Burkholz

机构 * CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全中心) Faculty of Computer Science, University of Vienna(维也纳大学计算机科学系)

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 本文通过Friedkin-Johnsen意见动力学模型分析多智能体LLM协商机制,揭示输入依赖的FJ参数使系统成为专家混合,并探讨基于自信度、感知自信度和初始观点对齐的影响者形成机制。

Comments Accepted at the 2nd Workshop on Compositional Learning at ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07375 2026-06-18 cs.CL cs.SD eess.AS 版本更新 57%

TurnGuide: Enhancing Meaningful Full Duplex Spoken Interactions via Dynamic Turn-Level Text-Speech Interleaving

TurnGuide: 通过动态轮次级文本-语音交错增强有意义的全双工口语交互

Wenqian Cui, Lei Zhu, Xiaohui Li, Zhihan Guo, Haoli Bai, Lu Hou, Irwin King

机构 * The Chinese University of Hong Kong(香港中文大学) Huawei Technologies(华为技术)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出TurnGuide方法,通过动态分割助手语音为对话轮次并交错生成轮次级文本和语音,解决全双工语音语言模型在连续双通道音频中集成离散文本令牌导致的时间对齐问题,显著提升语义连贯性和轮次交互性能。

Comments Interspeech 2026 Long Paper Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00567 2026-06-18 cs.IR cs.AI 版本更新 57%

Improving Scientific Document Retrieval with Academic Concept Index

利用学术概念索引改进科学文献检索

Jeyun Lee, Junhyoung Lee, Wonbin Kweon, Bowen Jin, Yu Zhang, Susik Yoon, Dongha Lee, Hwanjo Yu, Jiawei Han, Seongku Kang

机构 * Korea University Seoul South Korea University of Illinois Urbana-Champaign Champaign United States Texas A\&M University College Station United States Yonsei University Seoul South Korea Pohang University of Science Korea University University of Illinois Urbana-Champaign Texas A\&M University Yonsei University

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 针对通用检索器在科学领域因词汇和需求不匹配而表现不佳的问题,提出基于学术概念索引的方法,通过概念覆盖查询生成和概念聚焦上下文扩展,提升查询质量和检索性能。

Comments Accepted for publication in ACM TIST, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18222 2026-06-17 cs.CL cs.DL 新提交 57%

Darshana Graph: A Parallel Commentary Corpus for Comparative Indian Philosophy, with Stylometric and Exploratory Graph Analyses

Darshana Graph:用于比较印度哲学的平行注释语料库,附文体计量与探索性图分析

Joy Bose

机构 * Independent Researcher(独立研究者) Bangalore, India(印度班加罗尔)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 构建包含超12.5万条记录的印度哲学平行注释语料库,其中约8500条记录实现跨18位注释者的根颂对齐,通过文体计量和约束大语言模型管道分析论证风格与概念关系,揭示学派间分歧模式。

Comments 12 pages, 1 figure. Open Source Code available at https://github.com/joyboseroy/darshana-graph and dataset at https://huggingface.co/datasets/joyboseroy/darshana-graph

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18098 2026-06-17 cs.AI 新提交 57%

IsabeLLM: Automated Theorem Proving Applied to Formally Verifying Consensus

IsabeLLM: 自动化定理证明应用于共识的形式化验证

Elliot Jones, William Knottenbelt

机构 * Imperial College London(伦敦帝国学院)

专题命中 其他安全 :safety(abstract);分类 cs.AI

AI总结 本文改进IsabeLLM自动化定理证明工具,通过检索增强生成、错误追踪和反例生成提升大语言模型上下文,并兼容最新Isabelle和Sledgehammer,用于验证比特币工作量证明共识。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.17340 2026-06-17 cs.CV cs.AI 新提交 57%

Geometry-Consistent Endoscopic Representations for Image-Guided Navigation via Structured Foundation Model Adaptation

几何一致的内窥镜表示用于图像引导导航:基于结构化基础模型适配

Hongchao Shu, Roger D. Soberanis-Mukul, Hao Ding, Morgan Ringel, Mali Shen, Saif Iftekar Sayed, Hedyeh Rafii-Tari, Mathias Unberath

机构 * Department of Computer Science, Johns Hopkins University(约翰霍普金斯大学计算机科学系) Semaphor Surgical Johnson & Johnson MedTech(强生医疗科技)

专题命中 其他安全 :alignment(abstract);分类 cs.AI

AI总结 提出统一框架,结合合成数据管道与层级感知几何语义适配,学习几何一致且领域鲁棒的图像表示,提升单目内窥镜中的位姿估计与深度预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.17162 2026-06-17 cs.CL cs.HC cs.MA 新提交 57%

MemSlides: A Hierarchical Memory Driven Agent Framework for Personalized Slide Generation with Multi-turn Local Revision

MemSlides:一种用于个性化幻灯片生成与多轮局部修订的层次化记忆驱动智能体框架

Ye Jin, Yangyang Xu, Jun Zhu, Yibo Yang

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Tsinghua University(清华大学) Shanghai Jiao Tong University(上海交通大学)

专题命中 其他安全 :alignment(abstract);分类 cs.CL

AI总结 提出MemSlides层次化记忆框架,通过分离长期记忆(用户画像和工具记忆)与工作记忆,结合局部修订机制,实现个性化幻灯片生成中的用户偏好保持、多轮修订和可靠局部编辑。

Comments Code, website, project page, and video are linked in the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.11616 2026-06-17 cs.LG cs.IR 新提交 57%

DeMix: Debugging Training Data with Mixed Data Error Types by Investigating Influence Vectors

DeMix: 通过影响向量调试包含混合错误类型的训练数据

Jiale Deng, Yanyan Shen, Xiaogang Shi, Junjun Chai

机构 * Shanghai Jiao Tong University(上海交通大学) ByteDance Inc.(字节跳动) Tiktok

专题命中 其他安全 :alignment(abstract);分类 cs.LG

AI总结 提出DeMix框架,利用影响向量捕捉不同错误类型对模型行为的独特模式,将数据调试转化为多标签分类问题,并引入基于干预的学习策略,在11个任务上显著提升调试F1分数和修复后模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏