arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 5001 信号源:cs.CL, cs.AI, cs.LG

1. 复杂问题求解 5001 篇

2502.08180 2025-03-28 cs.CL cs.AI 62%

Enhancing LLM Character-Level Manipulation via Divide and Conquer

Zhen Xiong, Yujun Cai, Bryan Hooi, Nanyun Peng, Zhecheng Li, Yiwei Wang

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18002 2025-03-26 cs.NE cs.AI cs.AR cs.LG 62%

Neuromorphic Principles for Efficient Large Language Models on Intel Loihi 2

Steven Abreu, Sumit Bam Shrestha, Rui-Jie Zhu, Jason Eshraghian

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted to International Conference on Learning Representations (ICLR) Workshop on Scalable Optimization for Efficient and Adaptive Foundation Models (SCOPE)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14630 2025-03-20 cs.SE cs.AI cs.LG 62%

Assessing Large Language Models for Automated Feedback Generation in Learning Programming Problem Solving

Priscylla Silva, Evandro Costa

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.07763 2025-03-18 cs.CL cs.AI cs.DB 62%

Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Fangyu Lei, Jixuan Chen, Yuxiao Ye, Ruisheng Cao, Dongchan Shin, Hongjin Su, Zhaoqing Suo, Hongcheng Gao, Wenjing Hu, Pengcheng Yin, Victor Zhong, Caiming Xiong, Ruoxi Sun, Qian Liu, Sida Wang, Tao Yu

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments ICLR 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01262 2025-03-18 cs.CL cs.AI cs.HC 62%

Exploring ReAct Prompting for Task-Oriented Dialogue: Insights and Shortcomings

Michelle Elizabeth, Morgan Veyret, Miguel Couceiro, Ondrej Dusek, Lina M. Rojas-Barahona

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07155 2025-03-18 cs.AI cs.CL cs.MA cs.NI cs.SI 62%

Scaling Large Language Model-based Multi-Agent Collaboration

Chen Qian, Zihao Xie, YiFei Wang, Wei Liu, Kunlun Zhu, Hanchen Xia, Yufan Dang, Zhuoyun Du, Weize Chen, Cheng Yang, Zhiyuan Liu, Maosong Sun

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted to ICLR-2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11444 2025-03-17 cs.MA cs.AI cs.CL cs.OS 62%

Cerebrum (AIOS SDK): A Platform for Agent Development, Deployment, Distribution, and Discovery

Balaji Rama, Kai Mei, Yongfeng Zhang

专题命中 复杂问题求解 :CoT(abstract);分类 cs.CL、cs.AI

Comments Accepted to the 2025 Annual Conference of the North American Chapter of the Association for Computational Linguistics (NAACL) - System Demonstration Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10857 2025-03-17 cs.GR cs.AI cs.CL cs.CV 62%

Towards Understanding Graphical Perception in Large Multimodal Models

Kai Zhang, Jianwei Yang, Jeevana Priya Inala, Chandan Singh, Jianfeng Gao, Yu Su, Chenglong Wang

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Work in Progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02056 2025-03-13 eess.AS cs.AI cs.CL 62%

Synthio: Augmenting Small-Scale Audio Classification Datasets with Synthetic Data

Sreyan Ghosh, Sonal Kumar, Zhifeng Kong, Rafael Valle, Bryan Catanzaro, Dinesh Manocha

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted at ICLR 2025. Code and Checkpoints available here: https://github.com/Sreyan88/Synthio

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.15131 2025-03-13 cs.CL cs.AI 62%

Interactive-KBQA: Multi-Turn Interactions for Knowledge Base Question Answering with Large Language Models

Guanming Xiong, Junwei Bao, Wen Zhao

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments This work has been accepted by the ACL 2024 main conference. Code and data are available at: https://github.com/JimXiongGM/Interactive-KBQA

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06430 2025-03-11 cs.CL cs.AI cs.IR 62%

Graph Retrieval-Augmented LLM for Conversational Recommendation Systems

Zhangchi Qiu, Linhao Luo, Zicheng Zhao, Shirui Pan, Alan Wee-Chung Liew

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted by PAKDD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11721 2025-03-11 cs.CL cs.LG 62%

Explain-Query-Test: Self-Evaluating LLMs Via Explanation and Comprehension Discrepancy

Saeid Asgari Taghanaki, Joao Monteiro

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.LG

Comments Accepted to ICLR 2025, SSI-FM

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.15183 2025-03-11 cs.LG cs.AI 62%

GraphEdit: Large Language Models for Graph Structure Learning

Zirui Guo, Lianghao Xia, Yanhua Yu, Yuling Wang, Kangkang Lu, Zhiyong Huang, Chao Huang

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05919 2025-03-11 cs.CL cs.LG 62%

From Style to Facts: Mapping the Boundaries of Knowledge Injection with Finetuning

Eric Zhao, Pranjal Awasthi, Nika Haghtalab

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00242 2025-03-10 cs.CL cs.AI 62%

DeFT: Decoding with Flash Tree-attention for Efficient Tree-structured LLM Inference

Jinwei Yao, Kaiqi Chen, Kexun Zhang, Jiaxuan You, Binhang Yuan, Zeke Wang, Tao Lin

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Update DeFT-v4, accepted by ICLR'25 (https://openreview.net/forum?id=2c7pfOqu9k). Our code is available at https://github.com/LINs-lab/DeFT

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18168 2025-03-05 cs.CL cs.AI 62%

SECURA: Sigmoid-Enhanced CUR Decomposition with Uninterrupted Retention and Low-Rank Adaptation in Large Language Models

Yuxuan Zhang

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments New work on PEFT for LLMs, introducing S-MagNorm and CABR-LoRA to enhance fine-tuning performance and knowledge retention. In v4, we renamed Sigmoid-based Magnitude Normalization to S-MagNorm for clarity and added a gradient comparison between SECURA and CABR-LoRA to highlight their contributions

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.01241 2025-02-27 cs.AI cs.HC cs.LG 62%

Diagrammatization and Abduction to Improve AI Interpretability With Domain-Aligned Explanations for Medical Diagnosis

Brian Y. Lim, Joseph P. Cahaly, Chester Y. F. Sng, Adam Chew

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

Comments CHI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18873 2025-02-27 cs.AI cs.CL 62%

Multi-LLM Collaborative Search for Complex Problem Solving

Sen Yang, Yafu Li, Wai Lam, Yu Cheng

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.09046 2025-02-26 cs.IR cs.AI cs.LG 62%

HyPA-RAG: A Hybrid Parameter Adaptive Retrieval-Augmented Generation System for AI Legal and Policy Applications

Rishi Kalra, Zekun Wu, Ayesha Gulley, Airlie Hilliard, Xin Guan, Adriano Koshiyama, Philip Treleaven

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

Comments NAACL 2025 Industry Track & EMNLP 2024 CustomNLP4U Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16797 2025-02-25 cs.CL cs.AI 62%

Enhancing Answer Reliability Through Inter-Model Consensus of Large Language Models

Alireza Amiri-Margavi, Iman Jebellat, Ehsan Jebellat, Seyed Pouyan Mousavi Davoudi

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments 14 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08820 2025-02-20 cs.AI cs.CL 62%

Can a Single Model Master Both Multi-turn Conversations and Tool Use? CoALM: A Unified Conversational Agentic Language Model

Emre Can Acikgoz, Jeremiah Greer, Akul Datta, Ze Yang, William Zeng, Oussama Elachqar, Emmanouil Koukoumidis, Dilek Hakkani-Tür, Gokhan Tur

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12126 2025-02-19 cs.AI cs.LG cs.SI 62%

What Do LLMs Need to Understand Graphs: A Survey of Parametric Representation of Graphs

Dongqi Fu, Liri Fang, Zihao Li, Hanghang Tong, Vetle I. Torvik, Jingrui He

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Preprint, 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05085 2025-02-10 cs.LG cs.AI 62%

Causality can systematically address the monsters under the bench(marks)

Felix Leeb, Zhijing Jin, Bernhard Schölkopf

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04421 2025-02-10 cs.CR cs.AI cs.LG 62%

Assessing and Prioritizing Ransomware Risk Based on Historical Victim Data

Spencer Massengale, Philip Huff

专题命中 复杂问题求解 :chain-of-thought(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00988 2025-02-04 cs.CL cs.AI 62%

PlotGen: Multi-Agent LLM-based Scientific Data Visualization via Multimodal Feedback

Kanika Goswami, Puneet Mathur, Ryan Rossi, Franck Dernoncourt

专题命中 复杂问题求解 :planning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17167 2025-01-29 cs.HC cs.AI cs.CL 62%

StressPrompt: Does Stress Impact Large Language Models and Human Performance Similarly?

Guobin Shen, Dongcheng Zhao, Aorigele Bao, Xiang He, Yiting Dong, Yi Zeng

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments 11 pages, 9 figures, Accepted by AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.13514 2025-01-28 cs.CL cs.AI 62%

Self-DC: When to Reason and When to Act? Self Divide-and-Conquer for Compositional Unknown Questions

Hongru Wang, Boyang Xue, Baohang Zhou, Tianhua Zhang, Cunxiang Wang, Huimin Wang, Guanhua Chen, Kam-fai Wong

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14731 2025-01-28 cs.SE cs.AI cs.CL 62%

From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models

Zexing Xu, Zhuang Luo, Yichuan Li, Kyumin Lee, S. Rasoul Etesami

专题命中 复杂问题求解 :self-correction(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14210 2025-01-27 cs.CV cs.AI cs.LG 62%

PuzzleGPT: Emulating Human Puzzle-Solving Ability for Time and Location Prediction

Hammad Ayyubi, Xuande Feng, Junzhang Liu, Xudong Lin, Zhecan Wang, Shih-Fu Chang

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.AI、cs.LG

Comments NAACL 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13731 2025-01-24 cs.CL cs.AI 62%

Pseudocode-Injection Magic: Enabling LLMs to Tackle Graph Computational Tasks

Chang Gong, Wanrui Bian, Zhijie Zhang, Weiguo Zheng

专题命中 复杂问题求解 :reasoning(abstract);分类 cs.CL、cs.AI

Comments 24 pages

详情

展开后加载摘要…

URL PDF HTML 收藏