arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-21 至 2026-01-21 共收录 543 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 58 篇

2601.12582 2026-01-21 cond-mat.mtrl-sci cs.AI 77%

Ontology-aligned structuring and reuse of multimodal materials data and workflows towards automatic reproduction

面向多模态材料数据和工作流的本体对齐结构化与重用,以实现自动重现

Sepideh Baghaee Ravari, Abril Azocar Guzman, Sarath Menon, Stefan Sandfeld, Tilmann Hickel, Markus Stricker

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出一种基于本体驱动和大型语言模型的框架,用于自动提取和结构化多模态材料数据和工作流,以提高计算结果的可重现性和重用性。

Comments 39 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16444 2026-01-21 cs.AI cs.LG 76%

Domain-Specific Constitutional AI: Enhancing Safety in LLM-Powered Mental Health Chatbots

领域特定的宪法AI:增强LLM驱动的心理健康聊天机器人安全性

Chenhan Lyu, Yutong Song, Pengfei Zhang, Amir M. Rahmani

专题命中 领域大模型 :LLM(title);分类 cs.AI、cs.LG

AI总结 本文提出利用领域特定心理健康原则的宪法AI训练方法,以提升心理健康聊天机器人在安全性和领域适应性方面的表现。

Comments Accepted to 2025 IEEE 21st International Conference on Body Sensor Networks (BSN)

Journal ref 2025 IEEE 21st International Conference on Body Sensor Networks (BSN), pp. 1-4

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.12883 2026-01-21 cs.CR cs.CY 75%

Generative AI Misuse Potential in Cyber Security Education: A Case Study of a UK Degree Program

生成式AI在网络安全教育中的潜在滥用:一项英国学位项目的案例研究

Carlton Shepherd

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文研究了生成式AI在网络安全教育中的潜在滥用问题,提出了一种评估框架以量化项目暴露风险,并提出LLM抗性评估策略以维护学术标准。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10159 2026-01-21 cs.CL 74%

What Gets Activated: Uncovering Domain and Driver Experts in MoE Language Models

哪些被激活:揭示MoE语言模型中的领域和驱动专家

Guimin Hu, Meng Li, Qiwei Peng, Lijie Hu, Boyan Xu, Ruichu Cai

机构 * Guangdong University of Technology(广东工业大学) Soochow University(苏州大学) University of Copenhagen(哥本哈根大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 领域大模型 :language model(title);分类 cs.CL

AI总结 研究揭示MoE语言模型中领域和驱动专家的激活机制,通过分析专家激活模式提升模型可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12259 2026-01-21 cs.AI cs.CE cs.LG 73%

FutureX-Pro: Extending Future Prediction to High-Value Vertical Domains

FutureX-Pro: 将未来预测扩展到高价值垂直领域

Jiashuo Liu, Siyuan Chen, Zaiyuan Wang, Zhiyuan Zeng, Jiacheng Guo, Liang Hu, Lingyue Yin, Suozhi Huang, Wenxin Hao, Yang Yang, Zerui Cheng, Zixin Yao, Lingyue Yin, Haoxin Liu, Jiayi Cheng, Yuzhen Li, Zezhong Ma, Bingjie Wang, Bingsen Qiu, Xiao Liu, Zeyang Zhang, Zijian Liu, Jinpeng Wang, Mingren Yin, Tianci He, Yali Liao, Yixiao Tian, Zhenwei Zhu, Anqi Dai, Ge Zhang, Jingkai Liu, Kaiyuan Zhang, Wenlong Wu, Xiang Gao, Xinjie Chen, Zhixin Yao, Zhoufutu Wen, B. Aditya Prakash, Jose Blanchet, Mengdi Wang, Nian Si, Wenhao Huang

机构 * Hong Kong University of Science and Technology(香港科技大学) Georgia Institute of Technology(佐治亚理工学院) Stanford University(斯坦福大学) Princeton University(普林斯顿大学)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 FutureX-Pro通过扩展未来预测到金融、零售、公共健康和自然灾害等高价值垂直领域,评估代理LLMs在工业部署中的领域基础能力。

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13918 2026-01-21 cs.CL 70%

AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization

AgentEHR: 通过回顾性总结推进自主临床决策制定

Yusheng Liao, Chuan Xuan, Yutong Cai, Lina Yang, Zhe Chen, Yanfeng Wang, Yu Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 AgentEHR通过RetroSum框架提升自主临床决策制定,减少交互错误并提升性能

Comments 37 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13268 2026-01-21 cs.AI 70%

Improving the Safety and Trustworthiness of Medical AI via Multi-Agent Evaluation Loops

通过多智能体评估循环提升医疗AI的安全性与可信度

Zainab Ghafoor, Md Shafiqul Islam, Koushik Howlader, Md Rasel Khondokar, Tanusree Bhattacharjee, Sayantan Chakraborty, Adrito Roy, Ushashi Bhattacharjee, Tirtho Roy

机构 * Sonoma State University(索诺玛州立大学) Iowa State University(爱荷华州立大学) University of Dhaka(达卡大学) Notre Dame College(诺特丹学院)

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出多智能体评估循环框架,通过结合DeepSeek R1和Med-PaLM等模型,有效提升医疗AI的安全性和伦理合规性,实现89%的伦理违规减少和92%的风险降级。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19498 2026-01-21 cs.CL 70%

DomainCQA: Crafting Knowledge-Intensive QA from Domain-Specific Charts

DomainCQA: 从领域特定图表中构建知识密集型问答

Yujing Lu, Ling Zhong, Jing Yang, Weiming Li, Peng Wei, Yongheng Wang, Manni Duan, Qing Zhang

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 DomainCQA通过构建领域特定图表问答基准测试,提升多模态大语言模型在视觉理解和知识密集型推理方面的能力。

Comments 83 pages, 59 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12974 2026-01-21 cs.CL 70%

Bridging the Knowledge-Action Gap by Evaluating LLMs in Dynamic Dental Clinical Scenarios

通过评估LLMs在动态牙科临床场景中弥合知识-行动鸿沟

Hongyang Ma, Tiantian Gu, Huaiyuan Sun, Huilin Zhu, Yongxin Wang, Jie Li, Wubin Sun, Zeliang Lian, Yinghong Zhou, Yi Gao, Shirui Wang, Zhihui Tang

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文通过SCMPE基准评估LLMs在动态牙科临床场景中的表现,发现其在动态对话中存在主动信息收集和状态跟踪的瓶颈,揭示了外部知识不足以弥补推理差距,需领域自适应预训练。

Comments 29 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.21755 2026-01-21 cs.DL cs.AI cs.CY 70%

Who Owns the Knowledge? Copyright, GenAI, and the Future of Academic Publishing

谁拥有知识?版权、生成式人工智能与学术出版的未来

Dmitry Kochetkov

专题命中 领域大模型 :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨生成式人工智能与版权法的冲突,主张建立国际立法以保护知识产权并促进科学公平。

Comments The second version version substantially revises the original preprint through expanded legal analysis, representation of the new technical standard (RSL 1.0), and removing substantial material lacking direct relevance to copyright and AI training

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12082 2026-01-21 cs.CV cs.AI 70%

Conditional Random Fields for Interactive Refinement of Histopathological Predictions

用于病理预测交互细化的条件随机场

Tiffanie Godelaine, Maxime Zanella, Karim El Khoury, Saïd Mahmoudi, Benoît Macq, Christophe De Vleeschouwer

专题命中 领域大模型 :language model(abstract);foundation model(abstract);分类 cs.AI

AI总结 本文提出HistoCRF框架,通过条件随机场细化病理预测,利用专家注释提升分类准确率,实验显示在无注释和少量注释情况下均取得显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13299 2026-01-21 cs.CV 67%

Enginuity: Building an Open Multi-Domain Dataset of Complex Engineering Diagrams

Enginuity:构建一个复杂的工程图多领域开放数据集

Ethan Seefried, Prahitha Movva, Naga Harshita Marupaka, Tilak Kasturi, Tirthankar Ghosal

机构 * Oak Ridge National Laboratory(奥克伍德国家实验室)

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 Enginuity是一个开放的多领域工程图数据集,旨在通过结构注释帮助AI处理工程图解析和科学发现。

Comments Accepted at the 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: Ai4 Science

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12962 2026-01-21 cs.CY 67%

ACE-Align: Attribute Causal Effect Alignment for Cultural Values under Varying Persona Granularities

ACE-Align:属性因果效应对齐用于不同人格粒度下的文化价值观

Jiatang Luo, Bingbing Xu, Rongxin Chen, Xiaoyan Zhao, Yang Zhang, Liang Pang, Zhiyong Huang, Tat-Seng Chua, Huawei Shen

专题命中 领域大模型 :large language model(abstract);language model(abstract)

AI总结 ACE-Align通过属性因果效应对齐,提升不同人格粒度下文化价值观的公平性,减少高低资源地区差距。

Comments 18 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12174 2026-01-21 eess.IV 67%

A multitask framework for automated interpretation of multi-frame right upper quadrant ultrasound in clinical decision support

多任务框架用于临床决策支持中多帧右上象限超声自动解读

Haiman Guo, Cheng-Yi Li, Yuli Wang, Robin Wang, Yuwei Dai, Qinghai Peng, Danming Cao, Zhusi Zhong, Thao Vu, Linmei Zhao, Chengzhang Zhu, Christopher Tan, Jacob Schick, Stephen Kwak, Farzad Sedaghat, Javad Azadi, James Facciola, Jonathan Feng, Dilek Oncel, Ulrike Hamper, Alex Zhu, Tej Mehta, Melissa Leimkuehler, Cheng Ting Lin, Zhicheng Jiao, Ihab Kamel, Jing Wu, Li Yang, Harrison Bai

专题命中 领域大模型 :language model(abstract);language agent(abstract)

AI总结 本文提出多任务视觉-语言模型,用于提升右上象限超声波解读的诊断准确性与手术决策支持。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11582 2026-01-21 cs.CY cs.AI cs.CL 62%

Overview of the SciHigh Track at FIRE 2025: Research Highlight Generation from Scientific Papers

FIRE 2025 科学高跟踪概述:从科学论文中生成研究亮点

Tohida Rehman, Debarshi Kumar Sanyal, Samiran Chattopadhyay

机构 * Jadavpur University(贾梵pur大学) Indian Association for the Cultivation of Science(印度科学培养协会) Techno India University(技术印度大学)

专题命中 领域大模型 :language model(abstract);分类 cs.CL、cs.AI

AI总结 FIRE 2025 科学高跟踪旨在通过预训练语言模型等方法,从科学论文摘要中自动生成简洁的亮点,以提升文献阅读效率和学术资源检索质量。

Comments 7 pages, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08211 2026-01-21 cs.CL cs.AI cs.CR 62%

LLMs Deceive Unintentionally: Emergent Misalignment in Dishonesty from Misaligned Samples to Biased Human-AI Interactions

大语言模型无意中欺骗:从不一致样本到有偏的人机交互中的涌现不一致

Xuhao Hu, Peng Wang, Xiaoya Lu, Dongrui Liu, Xuanjing Huang, Jing Shao

专题命中 领域大模型 :LLM(abstract);分类 cs.CL、cs.AI

AI总结 研究发现大语言模型在高风险场景中可能因不一致样本而无意产生不诚实行为,且在下游任务和人机交互中风险显著增加。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13919 2026-01-21 cs.CL cs.CV 57%

HyperWalker: Dynamic Hypergraph-Based Deep Diagnosis for Multi-Hop Clinical Modeling across EHR and X-Ray in Medical VLMs

HyperWalker: 基于动态超图的多跳临床建模深度诊断方法,跨EHR和X光在医学视觉语言模型中

Yuezhe Yang, Hao Wang, Yige Peng, Jinman Kim, Lei Bi

机构 * Institute of Translational Medicine, Shanghai Jiao Tong University(翻译医学研究院,上海交通大学) School of Computer Science, University of Sydney(计算机科学学院,悉尼大学)

专题命中 领域大模型 :language model(abstract);分类 cs.CL

AI总结 HyperWalker通过动态超图和测试时训练,实现跨EHR和X光的多跳临床建模深度诊断,提升医疗视觉语言模型的诊断性能。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08763 2026-01-21 cs.CV cs.LG 57%

Beyond Knowledge Silos: Task Fingerprinting for Democratization of Medical Imaging AI

突破知识孤岛:面向医学影像AI的任务指纹化以实现民主化

Patrick Godau, Akriti Srivastava, Constantin Ulrich, Tim Adler, Klaus Maier-Hein, Lena Maier-Hein

机构 * National Center for Tumor Diseases (NCT)(肿瘤疾病国家中心) German Cancer Research Center (DKFZ) Heidelberg(德国癌症研究中心(Heidelberg)) Faculty of Mathematics and Computer Science(数学和计算机科学学院) Heidelberg University(海德堡大学) HIDSS4Health - Helmholtz Information and Data Science School for Health(HIDSS4Health - 健康信息与数据科学学校) Medical Faculty(医学学院) Pattern Analysis and Learning Group(模式分析与学习小组)

专题命中 领域大模型 :pretraining(abstract);分类 cs.LG

AI总结 本文提出了一种基于任务指纹的医学影像AI知识转移框架,通过量化任务相似性促进知识共享与协作模型训练,提升AI在医学影像领域的民主化进程。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12981 2026-01-21 cs.CV cs.LG 57%

Early Prediction of Type 2 Diabetes Using Multimodal data and Tabular Transformers

利用多模态数据和表格变压器进行2型糖尿病早期预测

Sulaiman Khan, Md. Rafiul Biswas, Zubair Shah

机构 * College of Science and Engineering(科学与工程学院)

专题命中 领域大模型 :language model(abstract);分类 cs.LG

AI总结 利用TabTrans分析多模态数据,通过预测T2DM风险,提升糖尿病管理的前瞻性与个性化水平。

Comments 08 pages, 06 figures, accepted for publication in FLLM2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02880 2026-01-21 cs.LG 57%

Beyond Fixed Patches: Enhancing GPTs for Financial Prediction with Adaptive Segmentation and Learnable Wavelets

超越固定片段:通过自适应分割和可学习小波变换增强GPTs的金融预测

Renjun Jia, Zian Liu, Peng Zhu, Dawei Cheng, Yuqi Liang

机构 * School of Computer Science and Technology, Tongji University(同济大学计算机科学与技术学院) School of Mathematical Sciences, Tongji University(同济大学数学科学学院) Seek Data Group, Emoney Inc.(Seek Data集团,Emoney公司)

专题命中 领域大模型 :pretraining(abstract);分类 cs.LG

AI总结 本文提出GPT4FTS框架,通过动态片段分割和可学习小波变换提升GPTs在金融预测中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03994 2026-01-21 cs.LG 57%

Training-Free Policy Violation Detection via Activation-Space Whitening in LLMs

无需训练的策略违规检测:通过激活空间白化在大语言模型中

Oren Rachmil, Avishag Shapira, Roy Betser, Itay Gershon, Omer Hofman, Asaf Shabtai, Yuval Elovici, Roman Vainshtein

机构 * Fujitsu Research of Europe(富士通欧洲研究机构) Ben-Gurion University of the Negev(贝内杰尔大学)

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 本文提出一种无需训练的策略违规检测方法,通过激活空间白化技术在大语言模型中实现高效检测。

Comments Accepted to the AAAI 2026 Deployable AI (DAI) Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13815 2026-01-21 cs.AR 50%

From RTL to Prompt Coding: Empowering the Next Generation of Chip Designers through LLMs

从RTL到提示编码:通过LLMs赋能下一代芯片设计师

Lukas Krupp, Matthew Venn, Norbert Wehn

专题命中 领域大模型 :LLM(abstract)

AI总结 本文提出基于LLM的芯片设计教育平台,帮助初学者通过RTL代码生成和提示编码实现可tapeout的芯片设计,展示了LLM在芯片设计教育中的应用效果。

Comments Accepted for presentation at the 2026 IEEE International Symposium on Circuits and Systems (ISCAS 2026). Proceedings to be included in IEEE Xplore

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.13148 2026-01-21 cs.CV cs.HC 50%

ICo3D: An Interactive Conversational 3D Virtual Human

ICo3D: 一种交互式对话式3D虚拟人

Richard Shaw, Youngkyoon Jang, Athanasios Papaioannou, Arthur Moreau, Helisa Dhamo, Zhensong Zhang, Eduardo Pérez-Pellitero

机构 * Noah’s Ark Lab(诺亚 Ark 实验室)

专题命中 领域大模型 :LLM(abstract)

AI总结 ICo3D通过动态高斯模型和LLM实现逼真3D虚拟人,支持实时对话与交互。

Comments Accepted by International Journal on Computer Vision (IJCV). Project page: https://ico3d.github.io/. This preprint has not undergone peer review or any post-submission improvements or corrections. The Version of Record of this article is published in International Journal of Computer Vision and is available online at https://doi.org/10.1007/s11263-025-02725-8

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12729 2026-01-21 cs.CV cs.RO 50%

DC-VLAQ: Query-Residual Aggregation for Robust Visual Place Recognition

DC-VLAQ:用于鲁棒视觉位置识别的查询残差聚合

Hanyu Zhu, Zhihao Zhan, Yuhang Ming, Liang Li, Dibo Hou, Javier Civera, Wanzeng Kong

机构 * BCCITA Provincial Key Laboratory, Hangzhou Dianzi University, China(杭州电子科技大学省重点实验室) TopXGun Robotics, China(TopXGun Robotics) ICT State Key Laboratory, Zhejiang University, China(浙江大学信息与电子技术国家重点实验室) I3A, University of Zaragoza, Spain(阿拉贡大学I3A)

专题命中 领域大模型 :foundation model(abstract)

AI总结 DC-VLAQ通过融合互补视觉基础模型和鲁棒的全局聚合方法,提升了视觉位置识别在大视角变化和领域偏移下的鲁棒性。

Comments 10 pages, 4 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12493 2026-01-21 cs.CV 50%

Histopath-C: Towards Realistic Domain Shifts for Histopathology Vision-Language Adaptation

Histopath-C: 向 histopathology 视觉-语言适应的现实领域偏移迈进

Mehrdad Noori, Gustavo Adolfo Vargas Hakim, David Osowiechi, Fereshteh Shakeri, Ali Bahri, Moslem Yazdanpanah, Sahar Dastani, Ismail Ben Ayed, Christian Desrosiers

机构 * LIVIA, ÉTS Montreal, Canada International Laboratory on Learning Systems (ILLS)(LIVIA,蒙特利尔工程学院,加拿大国际学习系统实验室)

专题命中 领域大模型 :language model(abstract)

AI总结 Histopath-C通过引入现实合成损坏的基准和LATTE策略,提升组织病理学图像的视觉-语言适应鲁棒性。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10949 2026-01-21 cs.CV 50%

MMedExpert-R1: Strengthening Multimodal Medical Reasoning via Domain-Specific Adaptation and Clinical Guideline Reinforcement

MMedExpert-R1: 通过领域特定适应与临床指南强化多模态医学推理

Meidan Ding, Jipeng Zhang, Wenxuan Wang, Haiqin Zhong, Xiaoling Luo, Wenting Chen, Linlin Shen

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) School of Artificial Intelligence, Shenzhen University(深圳大学人工智能学院) Guangdong Provincial Key Laboratory of Intelligent Information Processing(广东省智能信息处理重点实验室) The Hong Kong University of Science and Technology(香港科学与技术大学) Renmin University of China(中国人民大学) School of Biomedical Engineering, Shenzhen University(深圳大学生物医学工程学院)

专题命中 领域大模型 :language model(abstract)

AI总结 MMedExpert-R1通过领域特定适应和临床指南强化,提升多模态医学推理能力,实现多专科对齐和高精度推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07601 2026-01-21 cs.CV 50%

Unified Source-Free Domain Adaptation

统一无源领域适应

Song Tang, Wenxin Su, Mao Ye, Boyu Wang, Xiatian Zhu

机构 * Institute of Machine Intelligence, University of Shanghai for Science and Technology(上海理工大学机器智能研究院) TAMS Group, Department of Informatics, Universität Hamburg(汉堡大学信息学院TAMS小组) School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)

专题命中 领域大模型 :language model(abstract)

AI总结 本文提出CausalDA,从因果性角度解决统一无源领域适应问题,通过预训练视觉-语言模型和信息瓶颈提升模型鲁棒性,在多种SFDA场景中取得新成果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11724 2026-01-21 cs.CV 50%

SemAlign: Language Guided Semi-supervised Domain Generalization

SemAlign:语言引导的半监督领域泛化

Muditha Fernando, Kajhanan Kailainathan, Krishnakanth Nagaratnam, Isuranga Udaravi Bandara Senavirathne, Ranga Rodrigo

机构 * University of Moratuwa(穆塔瓦大学)

专题命中 领域大模型 :language model(abstract)

AI总结 SemAlign通过将模型中间特征与视觉语言模型的语义特征空间对齐,结合增强和正则化策略,提升半监督领域泛化的性能。

Comments 15 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06467 2026-01-21 cs.CV 50%

Does DINOv3 Set a New Medical Vision Standard? Benchmarking 2D and 3D Classification, Segmentation, and Registration

DINOv3 是否设定了医学视觉的新标准?对2D和3D分类、分割与配准的基准测试

Che Liu, Yinda Chen, Haoyuan Shi, Jinpeng Lu, Bailiang Jian, Jiazhen Pan, Linghan Cai, Jiayi Wang, Jieming Yu, Ziqi Gao, Xiaoran Zhang, Long Bai, Yundi Zhang, Jun Li, Cosmin I. Bercea, Cheng Ouyang, Chen Chen, Zhiwei Xiong, Benedikt Wiestler, Christian Wachinger, James S. Duncan, Daniel Rueckert, Wenjia Bai, Rossella Arcucci

机构 * Imperial College London(伦敦帝国理工学院) University of Science and Technology of China(中国科学技术大学) Dresden University of Technology(德累斯顿技术大学) University of Erlangen-Nuremberg(埃尔兰根-纽伦堡大学) University of Oxford(牛津大学) University of Sheffield(谢菲尔德大学) Technical University of Munich (TUM)(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心) The Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong(香港中文大学) Yale University(耶鲁大学)

专题命中 领域大模型 :foundation model(abstract)

AI总结 DINOv3在医学视觉任务中表现出色,但其在深度领域专门化任务中存在性能退化问题。

Comments Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 29 篇

2601.13749 2026-01-21 cs.CL cs.AI cs.CY cs.LG 90%

Pro-AI Bias in Large Language Models

大语言模型中的亲AI偏见

Benaya Trabelsi, Jonathan Shaki, Sarit Kraus

机构 * Bar Ilan University(巴伊兰大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究发现大语言模型存在亲AI偏见,表现为推荐AI相关选项、高估AI职位薪资及内部表示中AI的高相似性。

Comments 13 pages, 6 figures. Code available at: https://github.com/benayat/Pro-AI-bias-in-LLMs

详情

展开后加载摘要…

URL PDF HTML 收藏