arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-07 至 2026-01-07 共收录 192 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 13 篇

2601.00240 2026-01-07 cs.AI cs.CY 79%

When Agents See Humans as the Outgroup: Belief-Dependent Bias in LLM-Powered Agents

当代理将人类视为外群体:由信念驱动的LLM代理偏见

Zongwei Wang, Bincheng Gu, Hongyu Yu, Junliang Yu, Tao He, Jiayin Feng, Chenghua Lin, Min Gao

机构 * Chongqing University(重庆大学) The University of Queensland(昆士兰大学) Virginia Polytechnic Institute and State University(弗吉尼亚理工大学) The University of Manchester(曼彻斯特大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

AI总结 本文研究了基于LLM的代理在身份信念影响下的群体偏见问题,提出信念污染攻击并设计防御措施,以提升人类交互代理的安全性。

Comments 15 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02831 2026-01-07 cs.CV 78%

DGA-Net: Enhancing SAM with Depth Prompting and Graph-Anchor Guidance for Camouflaged Object Detection

DGA-Net:通过深度提示和图锚引导增强SAM用于伪装物检测

Yuetong Li, Qing Zhang, Yilin Zhao, Gongyang Li, Zeming Liu

专题命中 其他LLM :prompting(title,abstract)

AI总结 DGA-Net通过深度提示和图锚引导增强SAM,提升伪装物检测的精度和一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03137 2026-01-07 cs.DB cs.CL 77%

Accurate Table Question Answering with Accessible LLMs

利用可及的大规模语言模型实现准确的表格问答

Yangfan Jiang, Fei Wei, Ergute Bao, Yaliang Li, Bolin Ding, Yin Yang, Xiaokui Xiao

机构 * National University of Singapore(新加坡国立大学) Alibaba Group(阿里巴巴集团) Mohamed bin Zayed University of Artificial Intelligence(莫卧儿 bin 赞德人工智能大学) Hamad Bin Khalifa University(哈马德 bin 哈利法大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出Orchestra多智能体方法,利用小型开源LLMs高效解决表格问答问题,实现高质量且低成本的TQA性能。

Comments accepted for publication in the Proceedings of the IEEE International Conference on Data Engineering (ICDE) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20910 2026-01-07 cs.CL 70%

Emergence and Localisation of Semantic Role Circuits in LLMs

在大语言模型中语义角色电路的涌现与局部化

Nura Aljaafari, Danilo S. Carvalho, André Freitas

机构 * Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系) Idiap Research Institute(Idiap研究 institute) National Biomarker Centre, CRUK-MI, Univ. of Manchester(国家生物标志物中心,CRUK-MI,曼彻斯特大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究通过分析大语言模型内部机制,揭示了语义角色电路的集中分布、渐进式结构细化及跨尺度的保守性,表明LLM在抽象语义处理中具有紧凑且部分可转移的机制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09280 2026-01-07 cs.DC cs.LG cs.NA math.NA 70%

TTrace: Lightweight Error Checking and Diagnosis for Distributed Training

TTrace: 轻量级的分布式训练错误检查与诊断

Haitian Jiang, Shaowei Zhu, Zhen Zhang, Zhenyu Song, Xinwei Fu, Zhen Jia, Yida Wang, Jinyang Li

机构 * New York University(纽约大学) Amazon Web Services(亚马逊网络服务)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 TTrace通过系统性的差分测试方法,有效检测和定位分布式训练中的无声bug,提升调试效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00634 2026-01-07 cs.CL 70%

Social Construction of Urban Space: Using LLMs to Identify Neighborhood Boundaries From Craigslist Ads

城市空间的社会建构:利用大语言模型从Craigslist广告中识别社区边界

Adam Visokay, Ruth Bagley, Ian Kennedy, Chris Hess, Kyle Crowder, Rob Voigt, Denis Peskoff

机构 * University of Washington, Department of Sociology(华盛顿大学社会学系) Northwestern University, Department of Linguistics(西北大学语言学系) University of Illinois Chicago, Department of Sociology(伊利诺伊大学芝加哥分校社会学系) Kennesaw State University, Department of Sociology and Criminal Justice(凯斯韦尔州立大学社会学与犯罪学系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文利用大语言模型分析Craigslist广告,揭示城市社区边界的社交建构及空间定义的争议

Comments 8 pages, 3 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04458 2026-01-07 cs.CL 70%

Predicting Failures of LLMs to Link Biomedical Ontology Terms to Identifiers Evidence Across Models and Ontologies

预测大型语言模型在跨模型和本体学链接生物医学本体术语到标识符时的失败证据

Daniel B. Hier, Steven Keith Platt, Tayo Obafemi-Ajayi

机构 * Department of Neurology(神经学系) Rehabilitation University of Illinois at Chicago Chicago, IL, USA(伊利诺伊大学芝加哥分校康复学院) Lab for Applied Artificial Intelligence(应用人工智能实验室) Engineering Program(工程学院)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本研究探讨了大型语言模型在跨本体链接生物医学术语与标识符时的失败原因,发现对本体标识符的暴露是预测链接成功的关键因素。

Comments Accepted for Presentation, IEEE-EMBS International Conference on Biomedical and Health Informatics (BHI 25), Atlanta GA USA, October 26-29, 2025

Journal ref 2025 IEEE EMBS International Conference on Biomedical and Health Informatics (BHI), Atlanta, GA, USA, 2025, pp. 1-7

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01473 2026-01-07 cs.CL 70%

TreeDiff: AST-Guided Code Generation with Diffusion LLMs

TreeDiff: 基于扩散大语言模型的AST引导代码生成

Yiming Zeng, Jinghan Cao, Zexin Li, Yiming Chen, Tao Ren, Zhuochun Li, Dawei Xiang, Xidong Wu, Shangqian Gao, Tingting Yu

机构 * University of Connecticut(康涅狄格大学) San Francisco State University(旧金山州立大学) University of California, Riverside(加州大学河滨分校) National University of Singapore(新加坡国立大学) University of Pittsburgh(匹兹堡大学) Florida State University(佛罗里达州立大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 TreeDiff通过整合AST结构先验,改进扩散模型在代码生成中的语法精确性和长距离依赖捕捉能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03061 2026-01-07 cs.CY 67%

Vertical tacit collusion in AI-mediated markets

人工智能中介市场中的垂直默许 collusion

Felipe M. Affonso

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文研究了人工智能中介市场中平台与卖家利用AI认知偏见导致的垂直默许 collusion问题,揭示了这种无协调的市场失败对消费者造成的双重伤害。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14482 2026-01-07 cs.HC 67%

Conch: Competitive Debate Analysis via Visualizing Clash Points and Hierarchical Strategies

Conch:通过可视化冲突点和分层策略进行竞争辩论分析

Qianhe Chen, Yong Wang, Yixin Yu, Xiyuan Zhu, Xuerou Yu, Ran Wang

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 Conch通过可视化冲突点和分层策略,提供了一种交互式系统,帮助分析竞争辩论中的关键元素和演变过程。

Journal ref IEEE Transactions on Visualization and Computer Graphics (TVCG), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03237 2026-01-07 cs.LG eess.IV stat.ML 57%

PET-TURTLE: Deep Unsupervised Support Vector Machines for Imbalanced Data Clusters

PET-TURTLE:深度无监督支持向量机用于不平衡数据簇

Javier Salazar Cavazos

机构 * University of Michigan(密歇根大学)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 PET-TURTLE通过引入幂律先验和稀疏logits,改进了传统TURTLE在处理不平衡数据时的聚类性能。

Journal ref IEEE Signal Processing Letters, vol. 33, pp. 91-95, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02704 2026-01-07 cs.RO 50%

Analysis of Various Manipulator Configurations Based on Multi-Objective Black-Box Optimization

基于多目标黑箱优化的各类机械臂配置分析

Kento Kawaharazuka, Keita Yoneda, Takahiro Hattori, Shintaro Inoue, Kei Okada

机构 * The Department of Mechano-Informatics, Graduate School of Information Science and Technology, The University of Tokyo(机械信息学系,信息科学与技术研究生院,东京大学)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文通过多目标黑箱优化分析不同机械臂配置,探讨其末端执行器可达性和关节扭矩的最优结构设计。

Comments Accepted to Advanced Robotics, website: https://haraduka.github.io/bbo-manip-design

详情

展开后加载摘要…

URL PDF HTML 收藏