arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 8064 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 其他安全 8064 篇

2404.08566 2024-04-15 eess.SP cs.LG 57%

Mitigating Receiver Impact on Radio Frequency Fingerprint Identification via Domain Adaptation

Liu Yang, Qiang Li, Xiaoyang Ren, Yi Fang, Shafei Wang

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments Accepted by IEEE Internet of Things Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.07430 2024-04-15 cs.CL 57%

Adapted Large Language Models Can Outperform Medical Experts in Clinical Text Summarization

Dave Van Veen, Cara Van Uden, Louis Blankemeier, Jean-Benoit Delbrouck, Asad Aali, Christian Bluethgen, Anuj Pareek, Malgorzata Polacin, Eduardo Pontes Reis, Anna Seehofnerova, Nidhi Rohatgi, Poonam Hosamani, William Collins, Neera Ahuja, Curtis P. Langlotz, Jason Hom, Sergios Gatidis, John Pauly, Akshay S. Chaudhari

专题命中 其他安全 :safety(abstract);分类 cs.CL

Comments 27 pages, 19 figures

Journal ref Nature Medicine, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06911 2024-04-11 cs.CL 57%

GraSAME: Injecting Token-Level Structural Information to Pretrained Language Models via Graph-guided Self-Attention Mechanism

Shuzhou Yuan, Michael Färber

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments NAACL 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13980 2024-04-11 cs.CV cs.LG 57%

Carve3D: Improving Multi-view Reconstruction Consistency for Diffusion Models with RL Finetuning

Desai Xie, Jiahao Li, Hao Tan, Xin Sun, Zhixin Shu, Yi Zhou, Sai Bi, Sören Pirk, Arie E. Kaufman

专题命中 其他安全 :alignment(abstract);分类 cs.LG

Comments 22 pages, 16 figures. Our code, training and testing data, and video results are available at: https://desaixie.github.io/carve-3d. This paper has been accepted to CVPR 2024. v2: incorporated changes from the CVPR 2024 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.05195 2024-04-11 cs.CV cs.CL 57%

$λ$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space

Maitreya Patel, Sangmin Jung, Chitta Baral, Yezhou Yang

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Project page: https://eclipse-t2i.github.io/Lambda-ECLIPSE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06217 2024-04-10 cs.CL 57%

VI-OOD: A Unified Representation Learning Framework for Textual Out-of-distribution Detection

Li-Ming Zhan, Bo Liu, Xiao-Ming Wu

专题命中 其他安全 :safety(abstract);分类 cs.CL

Comments COLING 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04925 2024-04-09 cs.CL 57%

Multilingual Large Language Model: A Survey of Resources, Taxonomy and Frontiers

Libo Qin, Qiguang Chen, Yuhang Zhou, Zhi Chen, Yinghui Li, Lizi Liao, Min Li, Wanxiang Che, Philip S. Yu

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04492 2024-04-09 cs.RO cs.AI cs.CV 57%

Automated Lane Change Behavior Prediction and Environmental Perception Based on SLAM Technology

Han Lei, Baoming Wang, Zuwei Shui, Peiyuan Yang, Penghao Liang

专题命中 其他安全 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.00176 2024-04-08 cs.CL 57%

ChipNeMo: Domain-Adapted LLMs for Chip Design

Mingjie Liu, Teodor-Dumitru Ene, Robert Kirby, Chris Cheng, Nathaniel Pinckney, Rongjian Liang, Jonah Alben, Himyanshu Anand, Sanmitra Banerjee, Ismet Bayraktaroglu, Bonita Bhaskaran, Bryan Catanzaro, Arjun Chaudhuri, Sharon Clay, Bill Dally, Laura Dang, Parikshit Deshpande, Siddhanth Dhodhi, Sameer Halepete, Eric Hill, Jiashang Hu, Sumit Jain, Ankit Jindal, Brucek Khailany, George Kokai, Kishor Kunal, Xiaowei Li, Charley Lind, Hao Liu, Stuart Oberman, Sujeet Omar, Ghasem Pasandi, Sreedhar Pratty, Jonathan Raiman, Ambar Sarkar, Zhengjiang Shao, Hanfei Sun, Pratik P Suthar, Varun Tej, Walker Turner, Kaizhe Xu, Haoxing Ren

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Updated results for ChipNeMo-70B model

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03590 2024-04-05 cs.CV cs.AI 57%

SemGrasp: Semantic Grasp Generation via Language Aligned Discretization

Kailin Li, Jingbo Wang, Lixin Yang, Cewu Lu, Bo Dai

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02872 2024-04-04 cs.AI 57%

Integrating Explanations in Learning LTL Specifications from Demonstrations

Ashutosh Gupta, John Komp, Abhay Singh Rajput, Krishna Shankaranarayanan, Ashutosh Trivedi, Namrita Varshney

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments 21 Pages, 13 Page Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01156 2024-04-02 cs.CV cs.AI 57%

SyncMask: Synchronized Attentional Masking for Fashion-centric Vision-Language Pretraining

Chull Hwan Song, Taebaek Hwang, Jooyoung Yoon, Shunghyun Choi, Yeong Hyeon Gu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments CVPR2024 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00925 2024-04-02 cs.CV cs.CL 57%

LLMs are Good Sign Language Translators

Jia Gong, Lin Geng Foo, Yixuan He, Hossein Rahmani, Jun Liu

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.20026 2024-04-01 cs.CV cs.CL 57%

FSMR: A Feature Swapping Multi-modal Reasoning Approach with Joint Textual and Visual Clues

Shuang Li, Jiahua Wang, Lijie Wen

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.02192 2024-03-29 cs.DB cs.AI 57%

The FormAI Dataset: Generative AI in Software Security Through the Lens of Formal Verification

Norbert Tihanyi, Tamas Bisztray, Ridhi Jain, Mohamed Amine Ferrag, Lucas C. Cordeiro, Vasileios Mavroeidis

专题命中 其他安全 :safety(abstract);分类 cs.AI

Comments https://github.com/FormAI-Dataset PLEASE USE PUBLISHED VERSION FOR CITATION: https://doi.org/10.1145/3617555.3617874

Journal ref PROMISE 2023: Proceedings of the 19th International Conference on Predictive Models and Data Analytics in Software Engineering December 2023 Pages 33 to 43

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18183 2024-03-28 cs.AI cs.IR 57%

Can AI Models Appreciate Document Aesthetics? An Exploration of Legibility and Layout Quality in Relation to Prediction Confidence

Hsiu-Wei Yang, Abhinav Agrawal, Pavlos Fragkogiannis, Shubham Nitin Mulay

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.16081 2024-03-27 cs.CV cs.AI 57%

ViT-Lens: Towards Omni-modal Representations

Weixian Lei, Yixiao Ge, Kun Yi, Jianfeng Zhang, Difei Gao, Dylan Sun, Yuying Ge, Ying Shan, Mike Zheng Shou

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments This work is a follow-up of arXiv:2308.10185. Accepted to CVPR2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16702 2024-03-26 cs.CL cs.IR cs.SE 57%

ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search

Zehan Li, Jianfei Zhang, Chuantao Yin, Yuanxin Ouyang, Wenge Rong

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments Accepted to LREC-COLING 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.11401 2024-03-26 cs.CV cs.AI 57%

Scene-LLM: Extending Language Model for 3D Visual Understanding and Reasoning

Rao Fu, Jingyu Liu, Xilun Chen, Yixin Nie, Wenhan Xiong

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08849 2024-03-26 cs.CL 57%

OFA: A Framework of Initializing Unseen Subword Embeddings for Efficient Large-scale Multilingual Continued Pretraining

Yihong Liu, Peiqin Lin, Mingyang Wang, Hinrich Schütze

专题命中 其他安全 :alignment(abstract);分类 cs.CL

Comments NAACL 2024 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.12862 2024-03-20 cs.CL 57%

Epistemology of Language Models: Do Language Models Have Holistic Knowledge?

Minsu Kim, James Thorne

专题命中 其他安全 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.09585 2024-03-20 cs.CL 57%

LifeTox: Unveiling Implicit Toxicity in Life Advice

Minbeom Kim, Jahyun Koo, Hwanhee Lee, Joonsuk Park, Hwaran Lee, Kyomin Jung

专题命中 其他安全 :safety(abstract);分类 cs.CL

Comments 11 pages, 5 figures, NAACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.19778 2024-03-20 cs.HC cs.AI 57%

Human-AI collaboration is not very collaborative yet: A taxonomy of interaction patterns in AI-assisted decision making from a systematic review

Catalina Gomez, Sue Min Cho, Shichang Ke, Chien-Ming Huang, Mathias Unberath

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments 25 pages; 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.13923 2024-03-19 cs.LG cs.IR q-bio.BM 57%

Towards 3D Molecule-Text Interpretation in Language Models

Sihang Li, Zhiyuan Liu, Yanchen Luo, Xiang Wang, Xiangnan He, Kenji Kawaguchi, Tat-Seng Chua, Qi Tian

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07718 2024-03-19 cs.LG math.OC 57%

CaVE: A Cone-Aligned Approach for Fast Predict-then-optimize with Binary Linear Programs

Bo Tang, Elias B. Khalil

专题命中 其他安全 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11656 2024-03-19 cs.LG 57%

SIFU: Sequential Informed Federated Unlearning for Efficient and Provable Client Unlearning in Federated Optimization

Yann Fraboni, Martin Van Waerebeke, Kevin Scaman, Richard Vidal, Laetitia Kameni, Marco Lorenzi

专题命中 其他安全 :safety(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10368 2024-03-18 stat.ML cs.CR cs.LG 57%

Conformal Predictions for Probabilistically Robust Scalable Machine Learning Classification

Alberto Carlevaro, Teodoro Alamo Cantarero, Fabrizio Dabbene, Maurizio Mongelli

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments 19 pages, 6 figures, journal paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.00038 2024-03-18 eess.IV cs.CV cs.LG q-bio.QM 57%

Detecting Brain Tumors through Multimodal Neural Networks

Antonio Curci, Andrea Esposito

专题命中 其他安全 :safety(abstract);分类 cs.LG

Comments Presented at NeroPRAI 2024 (co-located with ICPRAM 2024). This version did not undergo peer review: refer to the open access version of record (see DOI)

Journal ref Proceedings of the 13th International Conference on Pattern Recognition Applications and Methods (ICPRAM 2024) - NeroPRAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.06072 2024-03-15 cs.CV cs.AI 57%

Actional Atomic-Concept Learning for Demystifying Vision-Language Navigation

Bingqian Lin, Yi Zhu, Xiaodan Liang, Liang Lin, Jianzhuang Liu

专题命中 其他安全 :alignment(abstract);分类 cs.AI

Comments Accepted by AAAI 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08426 2024-03-14 cs.CV cs.AI 57%

Language-Driven Visual Consensus for Zero-Shot Semantic Segmentation

Zicheng Zhang, Tong Zhang, Yi Zhu, Jianzhuang Liu, Xiaodan Liang, QiXiang Ye, Wei Ke

专题命中 其他安全 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏