arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12193 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12193 篇

2601.02831 2026-01-07 cs.CV 78%

DGA-Net: Enhancing SAM with Depth Prompting and Graph-Anchor Guidance for Camouflaged Object Detection

DGA-Net:通过深度提示和图锚引导增强SAM用于伪装物检测

Yuetong Li, Qing Zhang, Yilin Zhao, Gongyang Li, Zeming Liu

专题命中 其他LLM :prompting(title,abstract)

AI总结 DGA-Net通过深度提示和图锚引导增强SAM,提升伪装物检测的精度和一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00555 2026-01-05 cs.RO 78%

LLM-Based Agentic Exploration for Robot Navigation & Manipulation with Skill Orchestration

基于大语言模型的代理探索用于机器人导航与操作的技能编排

Abu Hanif Muhammad Syarubany, Farhan Zaki Rahmani, Trio Widianto

专题命中 其他LLM :LLM(title,abstract)

AI总结 本文提出基于大语言模型的代理探索系统,用于机器人导航与操作,通过技能编排实现端到端任务执行。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11239 2025-12-30 cs.CV 78%

Cross-modal Prompting for Balanced Incomplete Multi-modal Emotion Recognition

跨模态提示用于平衡不完整多模态情绪识别

Wen-Jue He, Xiaofeng Zhu, Zheng Zhang

专题命中 其他LLM :prompting(title,abstract)

AI总结 本文提出跨模态提示方法,通过增强模态特定特征和动态加权提升多模态情绪识别的准确性和鲁棒性。

Comments Accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12715 2025-12-23 cs.CV cs.RO 78%

AsyMoE: Leveraging Modal Asymmetry for Enhanced Expert Specialization in Large Vision-Language Models

AsyMoE:利用模态不对称性提升大视觉-语言模型专家专业化

Heng Zhang, Haichuan Hu, Yaomin Shen, Weihao Yu, Yilei Yuan, Haochen You, Guo Cheng, Zijian Zhang, Lubin Gan, Huihui Wei, Hao Zhang, Jin Huang

专题命中 其他LLM :language model(title,abstract)

AI总结 AsyMoE通过三个专门专家组解决视觉-语言模态不对称问题,提升大模型专家专业化性能。

Comments This submission has been withdrawn by the authors due to a fundamental error in the methodology that affects the validity of the main results

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.15804 2025-12-19 cs.SE 78%

XBIDetective: Leveraging Vision Language Models for Identifying Cross-Browser Visual Inconsistencies

XBIDetective: 利用视觉语言模型识别跨浏览器视觉不一致

Balreet Grewal, James Graham, Jeff Muizelaar, Jan Honza Odvarko, Suhaib Mujahid, Marco Castelluccio, Cor-Paul Bezemer

专题命中 其他LLM :language model(title,abstract)

AI总结 XBIDetective利用视觉语言模型识别跨浏览器视觉不一致,通过自动截图和分析检测渲染错误,准确率高达79%

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15871 2025-12-16 cs.CV 78%

Visual symbolic mechanisms: Emergent symbol processing in vision language models

视觉符号机制:视觉语言模型中的涌现符号处理

Rim Assouel, Declan Campbell, Yoshua Bengio, Taylor Webb

专题命中 其他LLM :language model(title,abstract)

AI总结 研究揭示了视觉语言模型中通过内容无关空间索引实现的新兴符号机制,用于解决视觉绑定问题并减少模型错误。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04129 2025-12-05 cs.CR 78%

Tipping the Dominos: Topology-Aware Multi-Hop Attacks on LLM-Based Multi-Agent Systems

推倒多米诺骨牌:面向基于大语言模型的多智能体系统的拓扑感知多跳攻击

Ruichao Liang, Le Yin, Jing Chen, Cong Wu, Xiaoyu Zhang, Huangpeng Gu, Zijian Zhang, Yang Liu

专题命中 其他LLM :LLM(title,abstract)

AI总结 针对基于大语言模型的多智能体系统,提出拓扑感知多跳攻击方法,揭示其内在安全漏洞并提出基于拓扑信任的防御框架。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03657 2025-12-02 cs.CV 78%

Dynamic Multimodal Prototype Learning in Vision-Language Models

视觉-语言模型中的动态多模态原型学习

Xingyu Zhu, Shuo Wang, Beier Zhu, Miaoge Li, Yunfan Li, Junfeng Fang, Zhicai Wang, Dongsheng Wang, Hanwang Zhang

机构 * University of Science and Technology of China(中国科学技术大学) Nanyang Technological University(南洋理工大学) The Hong Kong Polytechnic University(香港理工大学) Sichuan University(四川大学) National University of Singapore(新加坡国立大学) Shenzhen University(深圳大学)

专题命中 其他LLM :language model(title,abstract)

AI总结 本文提出ProtoMM框架,通过动态更新视觉粒子和多模态原型学习,提升视觉-语言模型在测试时的适应性能。

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12979 2025-12-01 q-bio.BM 78%

Guiding Generative Protein Language Models with Reinforcement Learning

通过强化学习引导生成蛋白质语言模型

Filippo Stocco, Maria Artigues-Lleixa, Andrea Hunklinger, Talal Widatalla, Marc Guell, Noelia Ferruz

专题命中 其他LLM :language model(title,abstract)

AI总结 通过强化学习引导生成蛋白质语言模型,实现EGFR结合物设计中结合亲和力的显著提升。

Comments 28 pages including main text and supporting information

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15118 2025-11-20 cs.CV 78%

Unbiased Semantic Decoding with Vision Foundation Models for Few-shot Segmentation

Jin Wang, Bingfeng Zhang, Jian Pang, Weifeng Liu, Baodi Liu, Honglong Chen

机构 * School of Control Science and Engineering, China University of Petroleum (East China)(控制科学与工程学院,中国石油大学(华东))

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09353 2025-11-18 cs.CR cs.CV 78%

DAVSP: Safety Alignment for Large Vision-Language Models via Deep Aligned Visual Safety Prompt

Yitong Zhang, Jia Li, Liyi Cai, Ge Li

专题命中 其他LLM :language model(title,abstract)

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08282 2025-11-12 cs.NI cs.CR cs.ET 78%

SRE-Llama -- Fine-Tuned Meta's Llama LLM, Federated Learning, Blockchain and NFT Enabled Site Reliability Engineering(SRE) Platform for Communication and Networking Software Services

Eranga Bandara, Safdar H. Bouk, Sachin Shetty, Ravi Mukkamala, Abdul Rahman, Peter Foytik, Ross Gore, Xueping Liang, Ng Wee Keong, Kasun De Zoysa

专题命中 其他LLM :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20303 2025-11-11 cs.CV 78%

DeepAndes: A Self-Supervised Vision Foundation Model for Multi-Spectral Remote Sensing Imagery of the Andes

Junlin Guo, James R. Zimmer-Dauphinee, Jordan M. Nieusma, Siqi Lu, Quan Liu, Ruining Deng, Can Cui, Jialin Yue, Yizhe Lin, Tianyuan Yao, Juming Xiong, Junchao Zhu, Chongyu Qu, Yuechen Yang, Mitchell Wilkes, Xiao Wang, Parker VanValkenburgh, Steven A. Wernke, Yuankai Huo

机构 * Department of Electrical and Computer Engineering, Vanderbilt University(电气与计算机工程系,范德比尔特大学) Department of Anthropology, Vanderbilt University(人类学系,范德比尔特大学) Data Science Institute, Vanderbilt University(数据科学研究院,范德比尔特大学) Department of Computer Science, Vanderbilt University(计算机科学系,范德比尔特大学) Department of Mathematics, Vanderbilt University(数学系,范德比尔特大学) Oak Ridge National Laboratory(橡树岭国家实验室) Department of Anthropology, Brown University(人类学系,布朗大学) Department of Radiology, Weill Cornell Medicine(放射学系,韦尔·科恩医学中心)

专题命中 其他LLM :foundation model(title,abstract)

Journal ref 10.1109/JSTARS.2025.3619423

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09013 2025-11-05 cs.CV 78%

Prompt to Restore, Restore to Prompt: Cyclic Prompting for Universal Adverse Weather Removal

Rongxin Liao, Feng Li, Yanyan Wei, Zenglin Shi, Le Zhang, Huihui Bai, Meng Wang

机构 * School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院) School of Information and Communication Engineering, University of Electronic Science and Technology of China(电子科技大学信息与通信工程学院) School of Computer Science and Technology, Beijing Jiaotong University(北京交通大学计算机科学与技术学院)

专题命中 其他LLM :prompting(title);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00821 2025-11-04 cs.CV 78%

OMEGA: Optimized Multimodal Position Encoding Index Derivation with Global Adaptive Scaling for Vision-Language Models

Ruoxiang Huang, Xindian Ma, Rundong Kong, Zhen Yuan, Peng Zhang

机构 * Tianjin University(天津大学)

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11553 2025-10-23 cs.CV 78%

How many samples to label for an application given a foundation model? Chest X-ray classification study

Nikolay Nechaev, Evgeniia Przhezdzetskaia, Viktor Gombolevskiy, Dmitry Umerenkov, Dmitry Dylov

机构 * AIRI Scoltech

专题命中 其他LLM :foundation model(title,abstract)

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16870 2025-10-21 cs.CV 78%

Uncovering Brain-Like Hierarchical Patterns in Vision-Language Models through fMRI-Based Neural Encoding

Yudan Ren, Xinlong Wang, Kexin Wang, Tian Xia, Zihan Ma, Zhaowei Li, Xiangrong Bi, Xiao Li, Xiaowei He

专题命中 其他LLM :language model(title,abstract)

Comments 14 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13821 2025-10-17 cs.NI 78%

LLM Agent Communication Protocol (LACP) Requires Urgent Standardization: A Telecom-Inspired Protocol is Necessary

Xin Li, Mengbing Liu, Chau Yuen

专题命中 其他LLM :LLM(title,abstract)

Comments Accepted at NeurIPS 2025 AI4NextG Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10533 2025-10-14 cs.CV 78%

Layout-Independent License Plate Recognition via Integrated Vision and Language Models

Elham Shabaninia, Fatemeh Asadi-zeydabadi, Hossein Nezamabadi-pour

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10084 2025-10-14 cs.CV 78%

Tracking the Spatiotemporal Evolution of Landslide Scars Using a Vision Foundation Model: A Novel and Universal Framework

Meijun Zhou, Gang Mei, Zhengjing Ma, Nengxiong Xu, Jianbing Peng

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08242 2025-10-10 cs.HC 78%

Simulating Teams with LLM Agents: Interactive 2D Environments for Studying Human-AI Dynamics

Mohammed Almutairi, Charles Chiang, Haoze Guo, Matthew Belcher, Nandini Banerjee, Maria Milkowski, Svitlana Volkova, Daniel Nguyen, Tim Weninger, Michael Yankoski, Trenton W. Ford, Diego Gomez-Zara

专题命中 其他LLM :LLM(title,abstract)

Comments 29 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24807 2025-09-30 cs.CR 78%

Active Authentication via Korean Keystrokes Under Varying LLM Assistance and Cognitive Contexts

Dong Hyun Roh, Rajesh Kumar

专题命中 其他LLM :LLM(title,abstract)

Comments Accepted for publication at IEEE-ICMLA 2025. Contains nine pages, six figures, and two tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24370 2025-09-30 cs.CV 78%

DINOReg: Strong Point Cloud Registration with Vision Foundation Model

Congjia Chen, Yufu Qu

机构 * School of Instrumentation and Optoelectronic Engineering(仪器与光电工程学院)

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23991 2025-09-30 cs.CV 78%

RPG360: Robust 360 Depth Estimation with Perspective Foundation Models and Graph Optimization

Dongki Jung, Jaehoon Choi, Yonghan Lee, Dinesh Manocha

机构 * University of Maryland, College Park(马里兰大学 College Park 分校)

专题命中 其他LLM :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04511 2025-09-30 cs.CV 78%

FA: Forced Prompt Learning of Vision-Language Models for Out-of-Distribution Detection

Xinhua Lu, Runhe Lai, Yanqi Wu, Kanghao Chen, Wei-Shi Zheng, Ruixuan Wang

机构 * Sun Yat-sen University(中山大学) Peng Cheng Laboratory(鹏城实验室) Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Key Laboratory of Machine Intelligence and Advanced Computing(人工智能与先进计算重点实验室)

专题命中 其他LLM :language model(title,abstract)

Comments 12 pages, 4 figures, Accepted by ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22638 2025-09-29 cs.CL cs.AI cs.LG 78%

Language Models Can Learn from Verbal Feedback Without Scalar Rewards

Renjie Luo, Zichen Liu, Xiangyan Liu, Chao Du, Min Lin, Wenhu Chen, Wei Lu, Tianyu Pang

机构 * Sea AI Lab(海智实验室) SUTD(新加坡科技设计大学) NUS(国立大学) NTU(南洋理工大学) University of Waterloo(滑铁卢大学)

专题命中 其他LLM :language model(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04931 2025-09-23 cs.HC 78%

Breaking the News: Taking the Roles of Influencer vs. Journalist in a LLM-Based Game for Raising Misinformation Awareness

Huiyun Tang, Songqi Sun, Kexin Nie, Ang Li, Anastasia Sergeeva, Ray LC

专题命中 其他LLM :LLM(title,abstract)

Comments Accepted to ACM CHI PLAY 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16193 2025-09-22 eess.AS 78%

Are Multimodal Foundation Models All That Is Needed for Emofake Detection?

Mohd Mujtaba Akhtar, Girish, Orchid Chetia Phukan, Swarup Ranjan Behera, Pailla Balakrishna Reddy, Ananda Chandra Nayak, Sanjib Kumar Nayak, Arun Balaji Buduru

专题命中 其他LLM :foundation model(title,abstract)

Comments Accepted to APSIPA-ASC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16724 2025-09-19 cs.SD eess.AS 78%

SALM: Spatial Audio Language Model with Structured Embeddings for Understanding and Editing

Jinbo Hu, Yin Cao, Ming Wu, Zhenbo Luo, Jun Yang

机构 * Institute of Acoustics, Chinese Academy of Sciences, China(中国科学院声学研究所) MiLM Plus, Xiaomi Inc., China(小米公司) Xi’an Jiaotong Liverpool University, China(西安交通大学利物浦大学) University of Chinese Academy of Sciences, China(中国科学院大学)

专题命中 其他LLM :language model(title,abstract)

Comments 5 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13982 2025-09-18 cs.PL 78%

CLMTracing: Black-box User-level Watermarking for Code Language Model Tracing

Boyu Zhang, Ping He, Tianyu Du, Xuhong Zhang, Lei Yun, Kingsum Chow, Jianwei Yin

专题命中 其他LLM :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏