arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-12-10 至 2025-12-10 共收录 139 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 15 篇

2504.20290 2025-12-10 astro-ph.IM astro-ph.GA 78%

FALCO: a Foundation model of Astronomical Light Curves for time dOmain astronomy

FALCO:用于时间域天文学的天文光变曲线基础模型

Xiaoxiong Zuo, Yihan Tao, Yang Huang, Zhixuan Kang, Huaxi Chen, Chenzhou Cui, Jiashu Pan, Xiao Kong, Xiaoyu Tang, Henggeng Han, Haiyang Mu, Yunfei Xu, Dongwei Fan, Guirong Xue, Ali Luo, Jifeng Liu

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 FALCO是一款基于Transformer架构的时间域天文学基础模型,通过自监督学习在未标记数据上训练,实现了高精度的恒星变异性分类、表面重力估计和耀斑识别。

Journal ref AJ 171 10 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08185 2025-12-10 cs.CR cs.AI 70%

A Practical Framework for Evaluating Medical AI Security: Reproducible Assessment of Jailbreaking and Privacy Vulnerabilities Across Clinical Specialties

评估医疗AI安全的实用框架:在不同临床专科中可重复评估劫持和隐私漏洞

Jinghao Wang, Ping Zhang, Carter Yagemann

机构 * The Ohio State University(俄亥俄州立大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出了一种可重复评估医疗AI安全性的框架,针对不同临床专科的劫持和隐私漏洞进行评估,无需特殊资源或数据权限。

Comments 6 pages, 1 figure, framework proposal

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04263 2025-12-10 cs.SE cs.LG 70%

Polynomiogram: An Integrated Framework for Root Visualization and Generative Art

多项式图:用于根可视化和生成艺术的集成框架

Hoang Duc Nguyen, Anh Van Pham, Hien D. Nguyen

机构 * Georgia Institute of Technology(佐治亚理工学院) University of Information Technology(信息技术大学) Vietnam National University(越南国家大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 Polynomiogram通过灵活的采样方案和双引擎架构,实现多项式根的科学探索与生成艺术的结合。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07992 2025-12-10 cs.LG cs.SE 70%

Bridging the Clinical Expertise Gap: Development of a Web-Based Platform for Accessible Time Series Forecasting and Analysis

弥合临床专业知识差距:面向可及性的时间序列预测与分析的网页平台开发

Aaron D. Mullen, Daniel R. Harris, Svetla Slavova, V. K. Cody Bumgardner

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出了一款网页平台,旨在通过提供多种预测模型和定制化功能,使临床研究人员更便捷地进行时间序列预测与分析。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08567 2025-12-10 cs.LG cs.AI 62%

A Hybrid Model for Stock Market Forecasting: Integrating News Sentiment and Time Series Data with Graph Neural Networks

一种股票市场预测的混合模型:整合新闻情感与时间序列数据与图神经网络

Nader Sadek, Mirette Moawad, Christina Naguib, Mariam Elzahaby

专题命中 领域大模型 :language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出一种结合新闻情感与时间序列数据的混合模型,利用图神经网络提升股票市场预测性能,实验显示GNN在准确率和精度上均优于LSTM基线。

Comments 11 pages, 6 figures. Published in the Proceedings of the 5th International Conference on Artificial Intelligence Research (ICAIR 2025). Published version available at: https://papers.academic-conferences.org/index.php/icair/article/view/4294

Journal ref Proceedings of the 5th International Conference on AI Research (ICAIR 2025), Vol. 5, No. 1, pp. 452-462 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06039 2025-12-10 q-bio.QM cs.AI cs.CV cs.LG 62%

AI-powered virtual tissues from spatial proteomics for clinical diagnostics and biomedical discovery

基于空间蛋白质组学的AI虚拟组织用于临床诊断和生物医学发现

Johann Wenckstern, Eeshaan Jain, Yexiang Cheng, Benedikt von Querfurth, Kiril Vasilev, Matteo Pariset, Phil F. Cheng, Petros Liakopoulos, Olivier Michielin, Andreas Wicki, Gabriele Gut, Charlotte Bunne

专题命中 领域大模型 :foundation model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出VirTues模型,利用空间蛋白质组学数据实现生物标记物发现和患者分层,提升临床诊断和生物医学研究的准确性。

Comments 25 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08139 2025-12-10 cs.LG 57%

Robust Agents in Open-Ended Worlds

开放世界中的鲁棒智能体

Mikayel Samvelyan

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 本研究提出Maestro和MiniHack框架,通过对抗课程和程序化内容生成提升智能体在开放环境中的鲁棒性和泛化能力。

Comments PhD Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 4 篇

2502.12992 2025-12-10 cs.CL cs.AI 81%

B-cos LM: Efficiently Transforming Pre-trained Language Models for Improved Explainability

B-cos LM:高效地将预训练语言模型转换以提高可解释性

Yifan Wang, Sukrut Rao, Ji-Ung Lee, Mayank Jobanputra, Vera Demberg

机构 * Saarland University(萨尔兰大学) Max Planck Institute for Informatics(马克斯·普朗克信息研究所) Saarland Informatics Campus(萨尔兰计算机科学校区)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

AI总结 本文提出B-cos LM,通过结合B-cos转换和任务微调,提高预训练语言模型的可解释性与效率。

Comments TMLR 12/2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08439 2025-12-10 cs.CV 78%

LapFM: A Laparoscopic Segmentation Foundation Model via Hierarchical Concept Evolving Pre-training

LapFM:通过分层概念演化的预训练构建腹腔镜分割基础模型

Qing Xu, Kun Yuan, Yuxiang Luo, Yuhao Zhai, Wenting Duan, Nassir Navab, Zhen Chen

机构 * School of Computer Science, University of Lincoln, UK(英国林肯大学计算机科学学院) University of Nottingham, UK(英国诺丁汉大学) University of Nottingham Ningbo China, China(中国宁波诺丁汉大学) University of Strasbourg, France(法国斯特拉斯堡大学) Technical University of Munich, Germany(德国慕尼黑技术大学) Graduate School of Information, Production and Systems, Waseda University, Japan(日本早稻田大学信息、生产与系统研究生院) Department of Gastrointestinal Surgery, The Second Qilu Hospital, Shandong University, China(中国山东大学第二齐鲁医院胃肠外科) School of Engineering and Physical Science, University of Lincoln, Lincoln LN6 7TS, UK(英国林肯大学工程与物理科学学院) Yale University, New Haven, CT 06510, USA(美国耶鲁大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

AI总结 LapFM通过分层概念演化预训练方法,构建了基于腹腔镜手术图像的大型基准,实现了对复杂手术场景的高效分割和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04964 2025-12-10 cs.CL 77%

Uncertainty Quantification for LLMs through Minimum Bayes Risk: Bridging Confidence and Consistency

通过最小贝叶斯风险进行大语言模型的不确定性量化:连接置信度与一致性

Roman Vashurin, Maiya Goloburda, Albina Ilina, Aleksandr Rubashevskii, Preslav Nakov, Artem Shelmanov, Maxim Panov

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·泽亚德人工智能大学)

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出通过最小贝叶斯风险连接模型置信度与输出一致性,改进大语言模型的不确定性量化方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08700 2025-12-10 cs.CV 50%

Scale-invariant and View-relational Representation Learning for Full Surround Monocular Depth

尺度不变与视图关系的表示学习用于全环绕单目深度估计

Kyumin Hwang, Wonhyeok Choi, Kiljoon Han, Wonjoon Choi, Minwoo Choi, Yongcheon Na, Minwoo Park, Sunghoon Im

机构 * Department of Electrical Engineering & Computer Sciences, Daegu Gyeongbuk Institute of Science and Technology (DGIST)(电子工程与计算机科学系,大邱庆北科学技术大学) Department of Autonomous Driving Perception Technology Vanguard Team, Hyundai Motor Company(自动驾驶感知技术先锋团队,现代汽车公司)

专题命中 知识编辑与模型理解 :foundation model(abstract)

AI总结 本文提出了一种结合尺度不变与视图关系的知识蒸馏方法,用于提升全环绕单目深度估计的性能与效率。

Comments Accepted at IEEE Robotics and Automation Letters (RA-L) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 8 篇

2512.07869 2025-12-10 q-bio.NC cs.AI 79%

Manifolds and Modules: How Function Develops in a Neural Foundation Model

流形与模块:神经基础模型中函数如何发展

Johannes Bertram, Luciano Dyballa, T. Anderson Keller, Savik Kinger, Steven W. Zucker

机构 * University of Tübingen(图宾根大学) School of Science & Technology(科学与技术学院) IE University(IE大学) The Kempner Institute for Natural and Artificial Intelligence(自然与人工智能研究所) Harvard University(哈佛大学) Department of Computer Science(计算机科学系) Yale University(耶鲁大学) Depts. of Computer Science and Biomedical Engineering(计算机科学与生物医学工程系) Wu Tsai Institute(吴士怀研究所)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本文通过分析神经基础模型中不同处理阶段的流形结构,揭示其在生物相关性方面的贡献。

Comments 25 pages, 10 figures, accepted at Data on the Brain & Mind Findings, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13936 2025-12-10 cs.CL cs.AI 73%

EEG-to-Text Translation: A Model for Deciphering Human Brain Activity

EEG-to-Text翻译:一种解码人类脑活动的模型

Saydul Akbar Murad, Ashim Dahal, Nick Rahimi

机构 * School of Computing Sciences & Computer Engineering, University of Southern Mississippi(计算科学与计算机工程学院,密西西比大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 R1 Translator通过结合双向LSTM和Transformer解码器,提升EEG-to-text解码性能,ROUGE指标优于T5和Brain Translator。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08935 2025-12-10 cs.IR 71%

Personalize Before Retrieve: LLM-based Personalized Query Expansion for User-Centric Retrieval

在检索前个性化:基于LLM的个性化查询扩展用于以用户为中心的检索

Yingyi Zhang, Pengyue Jia, Derong Xu, Yi Wen, Xianneng Li, Yichao Wang, Wenlin Zhang, Xiaopeng Li, Weinan Gan, Huifeng Guo, Yong Liu, Xiangyu Zhao

专题命中 其他LLM :LLM(title)

AI总结 本文提出PBR框架,通过在检索前个性化查询扩展,提升用户个性化检索的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15771 2025-12-10 cs.CY cs.AI econ.GN q-fin.EC 70%

Left Leaning Models: How AI Evaluates Economic Policy?

左倾模型:人工智能如何评估经济政策?

Maxim Chupilkin

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文研究了AI在多因素约束下对经济政策的偏好,发现LLMs普遍偏好高增长、低失业和低不平等,而非传统宏观经济目标。

Comments 16 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03295 2025-12-10 cs.AI cs.RO cs.SE 70%

Capability-Driven Skill Generation with LLMs: A RAG-Based Approach for Reusing Existing Libraries and Interfaces

基于能力驱动的技能生成:一种基于RAG的方法,用于重用现有库和接口

Luis Miguel Vieira da Silva, Aljosha Köcher, Nicolas König, Felix Gehlhoff, Alexander Fay

机构 * Ruhr University, Bochum, Germany(鲁尔大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种基于RAG的方法,利用大语言模型生成可执行代码,通过整合现有库和接口实现能力驱动的技能生成。

Comments \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08374 2025-12-10 cs.CV 67%

The Unseen Bias: How Norm Discrepancy in Pre-Norm MLLMs Leads to Visual Information Loss

看不见的偏见:预规范MLLM中规范差异如何导致视觉信息丢失

Bozhou Li, Xinda Xue, Sihan Yang, Yang Shi, Xinlong Chen, Yushuo Guan, Yuanxing Zhang, Wentao Zhang

机构 * Peking University(北京大学) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Xi’an Jiaotong University(西安交通大学) Kling Team, Kuaishou Technology(快手科技 Kling 团队)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文揭示了预规范MLLM中规范差异导致的视觉信息丢失问题,并提出通过插入层规范层来解决这一问题,提升模型整体能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08922 2025-12-10 cs.CV 50%

Unified Diffusion Transformer for High-fidelity Text-Aware Image Restoration

统一扩散变压器用于高质量文本感知图像恢复

Jin Hyeon Kim, Paul Hyunbin Cho, Claire Kim, Jaewon Min, Jaeeun Lee, Jihye Park, Yeji Choi, Seungryong Kim

机构 * KAIST AI(韩国科学技术院人工智能研究所) Samsung Electronics(三星电子)

专题命中 其他LLM :language model(abstract)

AI总结 UniT通过整合扩散变压器、视觉-语言模型和文本识别模块,实现高质量文本感知图像恢复,有效抑制文本幻觉并取得最佳性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08737 2025-12-10 cs.CY cs.MA 50%

Insured Agents: A Decentralized Trust Insurance Mechanism for Agentic Economy

受保代理:一种用于代理经济的去中心化信任保险机制

Botao 'Amber' Hu, Bangdao Chen

专题命中 其他LLM :LLM(abstract)

AI总结 本文提出了一种去中心化的信任保险机制,通过抵押金和TEE技术解决代理经济中的信任问题,实现安全、高效的纠纷解决。

Comments Submitted to AAMAS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏