arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12266 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12266 篇

2310.01459 2023-10-04 cs.CL cs.AI cs.HC 73%

NarrativePlay: Interactive Narrative Understanding

Runcong Zhao, Wenjia Zhang, Jiazheng Li, Lixing Zhu, Yanran Li, Yulan He, Lin Gui

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.01088 2023-10-03 cs.CL cs.LG cs.SD eess.AS 73%

Towards human-like spoken dialogue generation between AI agents from written dialogue

Kentaro Mitsui, Yukiya Hono, Kei Sawada

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 18 pages, 8 figures, 9 tables, audio samples: https://rinnakk.github.io/research/publications/CHATS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.07420 2023-09-26 cs.CL cs.AI cs.NE 73%

Named entity recognition using GPT for identifying comparable companies

Eurico Covas

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 10 pages, 1 figure, to be submited to a journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05918 2023-09-15 cs.CL cs.AI 73%

Stochastic LLMs do not Understand Language: Towards Symbolic, Explainable and Ontologically Based LLMs

Walid S. Saba

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 17 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.03242 2023-09-08 q-bio.GN cs.AI cs.LG cs.MA 73%

Automated Bioinformatics Analysis via AutoBA

Juexiao Zhou, Bin Zhang, Xiuying Chen, Haoyang Li, Xiaopeng Xu, Siyuan Chen, Xin Gao

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.06954 2023-09-06 cs.CL cs.AI 73%

ACTI at EVALITA 2023: Overview of the Conspiracy Theory Identification Task

Giuseppe Russo, Niklas Stoehr, Manoel Horta Ribeiro

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at the Evalita Workshop 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00915 2023-08-29 q-bio.NC cs.AI cs.LG cs.RO 73%

The feasibility of artificial consciousness through the lens of neuroscience

Jaan Aru, Matthew Larkum, James M. Shine

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01284 2023-08-21 cs.CL cs.AI 73%

Fighting Fire with Fire: Can ChatGPT Detect AI-generated Text?

Amrita Bhattacharjee, Huan Liu

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments to appear in SIGKDD Explorations (December 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.08637 2023-08-17 cs.CL cs.AI 73%

Analyzing the Limits of Self-Supervision in Handling Bias in Language

Lisa Bauer, Karthik Gopalakrishnan, Spandana Gella, Yang Liu, Mohit Bansal, Dilek Hakkani-Tur

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

Comments Accepted at Findings of the Conference on Empirical Methods in Natural Language Processing (EMNLP) 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00017 2023-07-28 cs.CL cs.AI 73%

Towards Explainable and Language-Agnostic LLMs: Symbolic Reverse Engineering of Language at Scale

Walid S. Saba

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Draft, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10633 2023-07-21 cs.CL cs.LG 73%

Multi-Method Self-Training: Improving Code Generation With Text, And Vice Versa

Shriyash K. Upadhyay, Etan J. Ginsberg

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 23 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.01784 2023-07-06 cs.CL cs.AI 73%

The Inner Sentiments of a Thought

Chris Gagne, Peter Dayan

专题命中 其他LLM :LLM(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.01201 2023-07-06 cs.CL cs.AI 73%

Schema-learning and rebinding as mechanisms of in-context learning and emergence

Sivaramakrishnan Swaminathan, Antoine Dedieu, Rajkumar Vasudeva Raju, Murray Shanahan, Miguel Lazaro-Gredilla, Dileep George

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.14060 2023-06-27 cs.CV cs.CL cs.LG 73%

DesCo: Learning Object Recognition with Rich Language Descriptions

Liunian Harold Li, Zi-Yi Dou, Nanyun Peng, Kai-Wei Chang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.13000 2023-06-23 cs.CY cs.AI cs.CL 73%

Apolitical Intelligence? Auditing Delphi's responses on controversial political issues in the US

Jonathan H. Rystrøm

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.01904 2023-06-12 cs.CL cs.AI 73%

Robust Multi-bit Natural Language Watermarking through Invariant Features

KiYoon Yoo, Wonhyuk Ahn, Jiho Jang, Nojun Kwak

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments ACL 2023 long

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.04990 2023-06-07 cs.CL cs.LG 73%

Explanation-based Finetuning Makes Models More Robust to Spurious Cues

Josh Magnus Ludan, Yixuan Meng, Tai Nguyen, Saurabh Shah, Qing Lyu, Marianna Apidianaki, Chris Callison-Burch

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.11042 2023-06-06 cs.CL cs.LG 73%

In-context Example Selection with Influences

Tai Nguyen, Eric Wong

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01150 2023-06-05 cs.CL cs.AI 73%

Did You Read the Instructions? Rethinking the Effectiveness of Task Definitions in Instruction Learning

Fan Yin, Jesse Vig, Philippe Laban, Shafiq Joty, Caiming Xiong, Chien-Sheng Jason Wu

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments ACL 2023, camera-ready; 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00198 2023-06-02 cs.CL cs.LG 73%

An Invariant Learning Characterization of Controlled Text Generation

Carolina Zheng, Claudia Shi, Keyon Vafa, Amir Feder, David M. Blei

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments To appear in the 2023 Conference of the Association for Computational Linguistics (ACL 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16837 2023-05-29 cs.SE cs.AI cs.LG 73%

ChatGPT: A Study on its Utility for Ubiquitous Software Engineering Tasks

Giriprasad Sridhara, Ranjani H. G., Sourav Mazumdar

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.14556 2023-05-25 cs.CL cs.AI 73%

Unraveling ChatGPT: A Critical Analysis of AI-Generated Goal-Oriented Dialogues and Annotations

Tiziano Labruna, Sofia Brenna, Andrea Zaninello, Bernardo Magnini

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.12463 2023-05-23 cs.CL cs.AI 73%

Teaching the Pre-trained Model to Generate Simple Texts for Text Simplification

Renliang Sun, Wei Xu, Xiaojun Wan

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by ACL Findings: 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.08005 2023-05-16 cs.CR cs.AI cs.CL cs.CY cs.HC 73%

Beyond the Safeguards: Exploring the Security Risks of ChatGPT

Erik Derner, Kristina Batistič

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02469 2023-05-05 cs.HC cs.AI cs.LG 73%

The System Model and the User Model: Exploring AI Dashboard Design

Fernanda Viégas, Martin Wattenberg

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 10 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.01750 2023-05-05 cs.CL cs.AI 73%

Few-shot In-context Learning for Knowledge Base Question Answering

Tianle Li, Xueguang Ma, Alex Zhuang, Yu Gu, Yu Su, Wenhu Chen

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09871 2023-04-26 cs.LG cs.AI math.OC 73%

A Theory on Adam Instability in Large-Scale Machine Learning

Igor Molybog, Peter Albert, Moya Chen, Zachary DeVito, David Esiobu, Naman Goyal, Punit Singh Koura, Sharan Narang, Andrew Poulton, Ruan Silva, Binh Tang, Diana Liskovich, Puxin Xu, Yuchen Zhang, Melanie Kambadur, Stephen Roller, Susan Zhang

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.10453 2023-04-21 cs.CL cs.AI 73%

Phoenix: Democratizing ChatGPT across Languages

Zhihong Chen, Feng Jiang, Junying Chen, Tiannan Wang, Fei Yu, Guiming Chen, Hongbo Zhang, Juhao Liang, Chen Zhang, Zhiyi Zhang, Jianquan Li, Xiang Wan, Benyou Wang, Haizhou Li

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.07333 2023-04-18 cs.CY cs.AI cs.CL cs.HC 73%

The Self-Perception and Political Biases of ChatGPT

Jérôme Rutinowski, Sven Franke, Jan Endendyk, Ina Dormuth, Markus Pauly

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.13312 2023-04-04 cs.CL cs.AI 73%

Neural Theory-of-Mind? On the Limits of Social Intelligence in Large LMs

Maarten Sap, Ronan LeBras, Daniel Fried, Yejin Choi

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Originally published at EMNLP 2022, extended to include ChatGPT and GPT-4 models on March 30th 2023 (extension not peer reviewed)

详情

展开后加载摘要…

URL PDF HTML 收藏