arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7608 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7608 篇

2311.11047 2023-11-21 cs.RO 50%

CLIPSwarm: Converting text into formations of robots

Pablo Pueyo, Eduardo Montijano, Ana C. Murillo, Mac Schwager

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments Please cite this article as "P. Pueyo, E. Montijano, A. C. Murillo, and M. Schwager, CLIPSwarm: Converting text into formations of robots. ICRA 2023 Workshop on Multi-Robot Learning"

Journal ref ICRA 2023, Workshop on Multi-Robot Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.12334 2023-11-15 cs.CV 50%

Improving Representation Learning for Histopathologic Images with Cluster Constraints

Weiyi Wu, Chongyang Gao, Joseph DiPalma, Soroush Vosoughi, Saeed Hassanpour

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted by ICCV2023

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2023, pp. 21404-21414

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.06242 2023-11-13 cs.CV 50%

Florence-2: Advancing a Unified Representation for a Variety of Vision Tasks

Bin Xiao, Haiping Wu, Weijian Xu, Xiyang Dai, Houdong Hu, Yumao Lu, Michael Zeng, Ce Liu, Lu Yuan

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09270 2023-11-07 cs.CV 50%

SpectralCLIP: Preventing Artifacts in Text-Guided Style Transfer from a Spectral Perspective

Zipeng Xu, Songlong Xing, Enver Sangineto, Nicu Sebe

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments WACV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.00859 2023-10-31 cs.CV 50%

Vision-Language Adaptive Mutual Decoder for OOV-STR

Jinshui Hu, Chenyu Liu, Qiandong Yan, Xuyang Zhu, Jiajia Wu, Jun Du, Lirong Dai

专题命中 知识编辑与模型理解 :language model(abstract)

Comments 1st Place Solution to ECCV 2022 OOV-ST Challenge; ICIG 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18812 2023-10-31 cs.CV 50%

UniCat: Crafting a Stronger Fusion Baseline for Multimodal Re-Identification

Jennifer Crawford, Haoli Yin, Luke McDermott, Daniel Cummings

专题命中 知识编辑与模型理解 :prompting(abstract)

Comments Accepted NeurIPS 2023 UniReps, 9 pages, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.16667 2023-10-26 cs.CV 50%

CoDet: Co-Occurrence Guided Region-Word Alignment for Open-Vocabulary Object Detection

Chuofan Ma, Yi Jiang, Xin Wen, Zehuan Yuan, Xiaojuan Qi

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Accepted by NeurIPS 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03923 2023-10-09 cs.CV cs.RO 50%

Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation

Kashu Yamazaki, Taisei Hanyu, Khoa Vo, Thang Pham, Minh Tran, Gianfranco Doretto, Anh Nguyen, Ngan Le

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.00011 2023-10-04 cs.CV eess.IV 50%

Joint Self-supervised Depth and Optical Flow Estimation towards Dynamic Objects

Zhengyang Lu, Ying Chen

专题命中 知识编辑与模型理解 :pretraining(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.14532 2023-09-25 cs.CV 50%

Scale-MAE: A Scale-Aware Masked Autoencoder for Multiscale Geospatial Representation Learning

Colorado J. Reed, Ritwik Gupta, Shufan Li, Sarah Brockman, Christopher Funk, Brian Clipp, Kurt Keutzer, Salvatore Candido, Matt Uyttendaele, Trevor Darrell

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments International Conference on Computer Vision 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10922 2023-09-21 eess.AS cs.SD 50%

Discrete Audio Representation as an Alternative to Mel-Spectrograms for Speaker and Speech Recognition

Krishna C. Puvvada, Nithin Rao Koluguri, Kunal Dhawan, Jagadeesh Balam, Boris Ginsburg

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Preprint. Submitted to ICASSP 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09485 2023-09-19 cs.CV 50%

Distributional Estimation of Data Uncertainty for Surveillance Face Anti-spoofing

Mouxiao Huang

专题命中 知识编辑与模型理解 :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12757 2023-08-25 cs.CV 50%

PartSeg: Few-shot Part Segmentation via Part-aware Prompt Learning

Mengya Han, Heliang Zheng, Chaoyue Wang, Yong Luo, Han Hu, Jing Zhang, Yonggang Wen

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.11722 2023-08-23 cs.CV 50%

Implicit Neural Representation for Cooperative Low-light Image Enhancement

Shuzhou Yang, Moxuan Ding, Yanmin Wu, Zihan Li, Jian Zhang

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.03026 2023-08-11 cs.CV 50%

Context Autoencoder for Self-Supervised Representation Learning

Xiaokang Chen, Mingyu Ding, Xiaodi Wang, Ying Xin, Shentong Mo, Yunhao Wang, Shumin Han, Ping Luo, Gang Zeng, Jingdong Wang

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted by International Journal of Computer Vision (IJCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03975 2023-08-09 cs.CV 50%

Prompted Contrast with Masked Motion Modeling: Towards Versatile 3D Action Representation Learning

Jiahang Zhang, Lilang Lin, Jiaying Liu

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted by ACM Multimedia 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13452 2023-08-02 astro-ph.SR physics.space-ph 50%

Probing the variations in the timing of the Sun's polar magnetic field reversals through observations and surface flux transport simulations

Elena M. Golubeva, Akash Biswas, Anna I. Khlystova, Pawan Kumar, Bidya Binay Karak

专题命中 知识编辑与模型理解 :SFT(abstract)

Comments Accepted for publication in MNRAS

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.13244 2023-07-26 cs.CV 50%

Multi-Granularity Prediction with Learnable Fusion for Scene Text Recognition

Cheng Da, Peng Wang, Cong Yao

专题命中 知识编辑与模型理解 :language model(abstract)

Comments submitted to TPAMI; an extension to our previous ECCV 2022 paper arXiv:2209.03592

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10046 2023-07-20 cs.CV 50%

Divert More Attention to Vision-Language Object Tracking

Mingzhe Guo, Zhipeng Zhang, Liping Jing, Haibin Ling, Heng Fan

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments 16 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.02677 2023-07-07 cs.CV 50%

Caption Anything: Interactive Image Description with Diverse Multimodal Controls

Teng Wang, Jinrui Zhang, Junjie Fei, Hao Zheng, Yunlong Tang, Zhe Li, Mingqi Gao, Shanshan Zhao

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments Tech-report

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.16991 2023-06-30 cs.CV 50%

Integrating Large Pre-trained Models into Multimodal Named Entity Recognition with Evidential Fusion

Weide Liu, Xiaoyang Zhong, Jingwen Hou, Shaohua Li, Haozhe Huang, Yuming Fang

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.05171 2023-06-14 cs.CV 50%

ULIP: Learning a Unified Representation of Language, Images, and Point Clouds for 3D Understanding

Le Xue, Mingfei Gao, Chen Xing, Roberto Martín-Martín, Jiajun Wu, Caiming Xiong, Ran Xu, Juan Carlos Niebles, Silvio Savarese

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18721 2023-06-12 cs.CV 50%

LayoutMask: Enhance Text-Layout Interaction in Multi-modal Pre-training for Document Understanding

Yi Tu, Ya Guo, Huan Chen, Jinyang Tang

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Accepted by ACL 2023 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.16193 2023-06-02 eess.AS cs.SD 50%

Probing phoneme, language and speaker information in unsupervised speech representations

Maureen de Seyssel, Marvin Lavechin, Yossi Adi, Emmanuel Dupoux, Guillaume Wisniewski

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Submitted to INTERSPEECH 2022, 5 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18203 2023-06-01 cs.CV 50%

Concept Decomposition for Visual Exploration and Inspiration

Yael Vinker, Andrey Voynov, Daniel Cohen-Or, Ariel Shamir

专题命中 知识编辑与模型理解 :language model(abstract)

Comments https://inspirationtree.github.io/inspirationtree/

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15727 2023-05-26 cs.CV 50%

POPE: 6-DoF Promptable Pose Estimation of Any Object, in Any Scene, with One Reference

Zhiwen Fan, Panwang Pan, Peihao Wang, Yifan Jiang, Dejia Xu, Hanwen Jiang, Zhangyang Wang

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.00788 2023-05-18 cs.CV 50%

Open-Vocabulary Point-Cloud Object Detection without 3D Annotation

Yuheng Lu, Chenfeng Xu, Xiaobao Wei, Xiaodong Xie, Masayoshi Tomizuka, Kurt Keutzer, Shanghang Zhang

专题命中 知识编辑与模型理解 :prompting(abstract)

Comments I want to update this manuscript

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.02578 2023-05-15 cs.IR 50%

SimLM: Pre-training with Representation Bottleneck for Dense Passage Retrieval

Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, Furu Wei

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Accepted to ACL 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.12083 2023-04-25 cs.IR 50%

Joint Semantic and Structural Representation Learning for Enhancing User Preference Modelling

Xuhui Ren, Wei Yuan, Tong Chen, Chaoqun Yang, Quoc Viet Hung Nguyen, Hongzhi Yin

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11330 2023-04-25 cs.CV 50%

Self-supervised Learning by View Synthesis

Shaoteng Liu, Xiangyu Zhang, Tao Hu, Jiaya Jia

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments 13 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏