arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 19037 信号源:cs.CL, cs.AI, cs.LG

1. 推理与问题求解 19037 篇

2503.14495 2025-12-01 cs.CL cs.AI cs.LG 78%

Temporal Consistency for LLM Reasoning Process Error Identification

LLM推理过程错误识别中的时序一致性

Jiacheng Guo, Yue Wu, Jiahao Qiu, Kaixuan Huang, Xinzhe Juan, Ling Yang, Mengdi Wang

机构 * Department of Electrical & Computer Engineering, Princeton University(普林斯顿大学电气与计算机工程系) AI Lab, Princeton University(普林斯顿大学人工智能实验室) Department of Computer Science & Engineering, University of Michigan(密歇根大学计算机科学与工程系)

专题命中 推理与问题求解 :LLM(title);分类 cs.CL、cs.AI、cs.LG

AI总结 本文提出了一种基于时序一致性的LLM推理过程错误识别方法,通过迭代自我反思提升验证准确性,在多个基准测试中实现了优于基线方法的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.16369 2025-11-21 eess.SP cs.NI 78%

Reasoning Meets Representation: Envisioning Neuro-Symbolic Wireless Foundation Models

推理与表征的结合:展望神经符号无线基础模型

Jaron Fontaine, Mohammad Cheraghinia, John Strassner, Adnan Shahid, Eli De Poorter

专题命中 推理与问题求解 :foundation model(title,abstract)

AI总结 本文提出神经符号无线基础模型,结合神经网络与符号推理,以解决无线通信中的可解释性、鲁棒性和合规性问题,推动6G网络的智能化发展。

Comments Accepted at the 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: AI and ML for Next-Generation Wireless Communications and Networking (AI4NextG)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15333 2025-11-20 cs.RO 78%

C2F-Space: Coarse-to-Fine Space Grounding for Spatial Instructions using Vision-Language Models

Nayoung Oh, Dohyun Kim, Junhyeong Bang, Rohan Paul, Daehyung Park

专题命中 推理与问题求解 :language model(title,abstract)

Comments 16 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21323 2025-11-18 cs.CR 78%

LLM-driven Provenance Forensics for Threat Investigation and Detection

Kunal Mukherjee, Murat Kantarcioglu

专题命中 推理与问题求解 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11576 2025-11-18 cs.CV 78%

Causality Matters: How Temporal Information Emerges in Video Language Models

Yumeng Shi, Quanyu Long, Yin Wu, Wenya Wang

专题命中 推理与问题求解 :language model(title,abstract)

Comments Accepted to AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11526 2025-11-17 cs.CV 78%

Bridging Hidden States in Vision-Language Models

Benjamin Fein-Ashley, Jacob Fein-Ashley

机构 * University of Southern California(南加州大学)

专题命中 推理与问题求解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09843 2025-11-14 cs.CV astro-ph.IM astro-ph.SR 78%

CORONA-Fields: Leveraging Foundation Models for Classification of Solar Wind Phenomena

Daniela Martin, Jinsu Hong, Connor O'Brien, Valmir P Moraes Filho, Jasmine R. Kobayashi, Evangelia Samara, Joseph Gallego

机构 * University of Delaware(德克萨斯大学) Georgia State University(佐治亚州立大学) Boston University(波士顿大学) Catholic University of America(美国天主教大学) Southwest Research Institute(西南研究院) NASA Goddard Space Flight Center(NASA戈达德空间飞行中心) Drexel University(德雷塞尔大学)

专题命中 推理与问题求解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06904 2025-11-13 cs.CV 78%

An Instance-Aware Prompting Framework for Training-free Camouflaged Object Segmentation

Chao Yin, Jide Li, Hang Yao, Xiaoqiang Li

机构 * Shanghai University(上海大学)

专题命中 推理与问题求解 :prompting(title,abstract)

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06238 2025-11-11 cs.CV 78%

Temporal-Guided Visual Foundation Models for Event-Based Vision

Ruihao Xia, Junhong Cai, Luziwei Leng, Liuyi Wang, Chengju Liu, Ran Cheng, Yang Tang, Pan Zhou

机构 * Key Laboratory of Smart Manufacturing in Energy Chemical Process, Ministry of Education, East China University of Science and Technology(能源化工过程智能制造重点实验室,东华大学) ACSLab, Huawei Technologies Company Ltd.(华为技术有限公司ACS实验室) Department of Computer Science and Engineering, Southern University of Science and Technology(南方科技大学计算机科学与工程系) Shanghai Institute of Intelligent Science and Technology, Tongji University(同济大学智能科学与技术研究院) Department of Data Science and Artificial Intelligence and the Department of Computing, The Hong Kong Polytechnic University(香港理工大学数据科学与人工智能系、计算系) Singapore Management University(新加坡国立大学)

专题命中 推理与问题求解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20263 2025-11-11 cs.CV 78%

Vector-Quantized Vision Foundation Models for Object-Centric Learning

Rongzhen Zhao, Vivienne Wang, Juho Kannala, Joni Pajarinen

机构 * Aalto University(阿alto大学) University of Oulu(奥卢大学)

专题命中 推理与问题求解 :foundation model(title,abstract)

Comments Accepted to ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04646 2025-11-07 cs.AI cs.CL cs.LG cs.MA 78%

DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration

Narjes Nourzad, Hanqing Yang, Shiyu Chen, Carlee Joe-Wong

机构 * University of Southern California(南加州大学) Carnegie Mellon University(卡内基梅隆大学)

专题命中 推理与问题求解 :LLM(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04023 2025-11-07 cs.SE cs.CR 78%

LLM-Driven Adaptive Source-Sink Identification and False Positive Mitigation for Static Analysis

Shiyin Lin

专题命中 推理与问题求解 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01423 2025-11-04 cs.SE 78%

LLM-Assisted Tool for Joint Generation of Formulas and Functions in Rule-Based Verification of Map Transformations

Ruidi He, Yu Zhang, Meng Zhang, Andreas Rausch

专题命中 推理与问题求解 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22978 2025-10-28 cs.HC 78%

Reasoning About Reasoning: Towards Informed and Reflective Use of LLM Reasoning in HCI

Ramaravind Kommiya Mothilal, Sally Zhang, Syed Ishtiaque Ahmed, Shion Guha

专题命中 推理与问题求解 :LLM(title,abstract)

Comments 14 pages, 3 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22838 2025-10-28 cs.CV 78%

Semantic-Preserving Cross-Style Visual Reasoning for Robust Multi-Modal Understanding in Large Vision-Language Models

Aya Nakayama, Brian Wong, Yuji Nishimura, Kaito Tanaka

专题命中 推理与问题求解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21757 2025-10-28 cs.CV 78%

Agro-Consensus: Semantic Self-Consistency in Vision-Language Models for Crop Disease Management in Developing Countries

Mihir Gupta, Pratik Desai, Ross Greer

机构 * The Harker School(哈克尔学校) University of California, Merced(加州大学默塞德分校)

专题命中 推理与问题求解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23566 2025-10-28 cs.CV 78%

Uni-MuMER: Unified Multi-Task Fine-Tuning of Vision-Language Model for Handwritten Mathematical Expression Recognition

Yu Li, Jin Jiang, Jianhua Zhu, Shuai Peng, Baole Wei, Yuxuan Zhou, Liangcai Gao

机构 * Wangxuan Institute of Computer Technology, Peking University(计算机技术研究院,北京大学)

专题命中 推理与问题求解 :language model(title,abstract)

Comments Accepted by NeurIPS 2025 as a spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14674 2025-10-27 cs.CV 78%

Recognition through Reasoning: Reinforcing Image Geo-localization with Large Vision-Language Models

Ling Li, Yao Zhou, Yuxuan Liang, Fugee Tsung, Jiaheng Wei

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 推理与问题求解 :language model(title,abstract)

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12448 2025-10-27 cs.CV 78%

SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning

Yang Liu, Ming Ma, Xiaomin Yu, Pengxiang Ding, Han Zhao, Mingyang Sun, Siteng Huang, Donglin Wang

机构 * Westlake University(西湖大学) Zhejiang University(浙江大学) Harbin Institute of Technology(哈尔滨工业大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Shanghai Innovation Institute(上海创新研究院)

专题命中 推理与问题求解 :language model(title,abstract)

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21401 2025-10-24 cs.CV 78%

JaiLIP: Jailbreaking Vision-Language Models via Loss Guided Image Perturbation

Md Jueal Mia, M. Hadi Amini

机构 * Knight Foundation School of Computing and Information Sciences (KFSCIS)(骑士基金会计算与信息科学学院) Florida International University(佛罗里达国际大学) Sustainability, Optimization, and Learning for InterDependent networks laboratory (solid lab)(可持续性、优化与互依赖网络学习实验室)

专题命中 推理与问题求解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14008 2025-10-22 cs.MA 78%

Stop Reducing Responsibility in LLM-Powered Multi-Agent Systems to Local Alignment

Jinwei Hu, Yi Dong, Shuang Ao, Zhuoyun Li, Boxuan Wang, Lokesh Singh, Guangliang Cheng, Sarvapali D. Ramchurn, Xiaowei Huang

专题命中 推理与问题求解 :LLM(title,abstract)

Comments Updated manuscript of our previous version (arXiv:2502.01714). Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17376 2025-10-21 cs.SE 78%

AdapTrack: Constrained Decoding without Distorting LLM's Output Intent

Yongmin Li, Jia Li, Ge Li, Zhi Jin

专题命中 推理与问题求解 :LLM(title);language model(abstract)

Comments to be published in ICSE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12190 2025-10-15 cs.CV 78%

Hierarchical Reasoning with Vision-Language Models for Incident Reports from Dashcam Videos

Shingo Yokoi, Kento Sasaki, Yu Yamaguchi

机构 * Turing Inc.(图灵公司)

专题命中 推理与问题求解 :language model(title,abstract)

Comments 2nd Place Winner, ICCV 2025 2COOOL Competition

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15192 2025-10-08 cs.CV 78%

Leveraging Foundation Models for Multimodal Graph-Based Action Recognition

Fatemeh Ziaeetabar, Florentin Wörgötter

机构 * School of Mathematics, Statistics and Computer Science, College of Science, University of Tehran(数学、统计与计算机科学学院,科学学院,塔里斯坦大学)

专题命中 推理与问题求解 :foundation model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03903 2025-10-07 cs.CV 78%

Zero-Shot Fine-Grained Image Classification Using Large Vision-Language Models

Md. Atabuzzaman, Andrew Zhang, Chris Thomas

机构 * Department of Computer Science(计算机科学系) Virginia Tech(弗吉尼亚理工大学)

专题命中 推理与问题求解 :language model(title,abstract)

Comments Accepted to EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22010 2025-10-02 cs.CV 78%

CoFFT: Chain of Foresight-Focus Thought for Visual Language Models

Xinyu Zhang, Yuxuan Dong, Lingling Zhang, Chengyou Jia, Zhuohang Dang, Basura Fernando, Jun Liu, Mike Zheng Shou

机构 * School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院) Ministry of Education Key Laboratory of Intelligent Networks and Network Security, China(教育部智能网络与网络安全重点实验室) Shaanxi Province Key Laboratory of Big Data Knowledge Engineering, China(陕西省大数据知识工程重点实验室) IHPC, Agency for Science, Technology and Research, Singapore(新加坡科技研究局IHPC) Show Lab, National University of Singapore(新加坡国立大学Show实验室) College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算与数据科学学院)

专题命中 推理与问题求解 :language model(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15865 2025-10-01 cs.CV 78%

How Do Large Vision-Language Models See Text in Image? Unveiling the Distinctive Role of OCR Heads

Ingeol Baek, Hwan Chang, Sunghyun Ryu, Hwanhee Lee

专题命中 推理与问题求解 :language model(title,abstract)

Comments EMNLP 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14142 2025-10-01 cs.SD eess.AS 78%

AudSemThinker: Enhancing Audio-Language Models through Reasoning over Semantics of Sound

Gijs Wijngaard, Elia Formisano, Michele Esposito, Michel Dumontier

机构 * Maastricht University(马斯特里赫特大学)

专题命中 推理与问题求解 :language model(title,abstract)

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22674 2025-09-30 cs.CV 78%

Pathological Truth Bias in Vision-Language Models

Yash Thube

机构 * Savitribai Phule Pune University (SPPU)(萨维特里·布尔大学(SPPU))

专题命中 推理与问题求解 :language model(title,abstract)

Comments 10 pages, 12 figures. Code for MATS released at https://github.com/thubZ09/mats-spatial-reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15293 2025-09-23 cs.CV cs.RO 78%

How Good are Foundation Models in Step-by-Step Embodied Reasoning?

Dinura Dissanayake, Ahmed Heakl, Omkar Thawakar, Noor Ahsan, Ritesh Thawkar, Ketan More, Jean Lahoud, Rao Anwer, Hisham Cholakkal, Ivan Laptev, Fahad Shahbaz Khan, Salman Khan

机构 * Mohamed bin Zayed University of AI(穆罕默德·本·扎耶德人工智能大学) Linköping University(林雪平大学) Australian National University(澳大利亚国立大学)

专题命中 推理与问题求解 :foundation model(title,abstract)

Comments Project page: https://mbzuai-oryx.github.io/FoMER-Bench/

详情

展开后加载摘要…

URL PDF HTML 收藏