arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7608 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7608 篇

2503.23094 2025-04-01 cs.CV 50%

FRAME: Floor-aligned Representation for Avatar Motion from Egocentric Video

Andrea Boscolo Camiletto, Jian Wang, Eduardo Alvarado, Rishabh Dabral, Thabo Beeler, Marc Habermann, Christian Theobalt

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17721 2025-04-01 cs.HC cs.SE 50%

Prototyping with Prompts: Emerging Approaches and Challenges in Generative AI Design for Collaborative Software Teams

Hari Subramonyam, Divy Thakkar, Andrew Ku, Jürgen Dieber, Anoop Sinha

专题命中 知识编辑与模型理解 :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15482 2025-03-28 cs.CV 50%

SplatFlow: Self-Supervised Dynamic Gaussian Splatting in Neural Motion Flow Field for Autonomous Driving

Su Sun, Cheng Zhao, Zhuoyang Sun, Yingjie Victor Chen, Mei Chen

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13652 2025-03-26 cs.CV 50%

RelationField: Relate Anything in Radiance Fields

Sebastian Koch, Johanna Wald, Mirco Colosi, Narunas Vaskevicius, Pedro Hermosilla, Federico Tombari, Timo Ropinski

专题命中 知识编辑与模型理解 :language model(abstract)

Comments CVPR 2025. Project page: https://relationfield.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13769 2025-03-25 cs.CV 50%

Continual Unlearning for Foundational Text-to-Image Models without Generalization Erosion

Kartik Thakral, Tamar Glaser, Tal Hassner, Mayank Vatsa, Richa Singh

专题命中 知识编辑与模型理解 :foundation model(abstract)

Comments Under submission to T-PAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08216 2025-03-17 cs.CV 50%

Attention Hijackers: Detect and Disentangle Attention Hijacking in LVLMs for Hallucination Mitigation

Beitao Chen, Xinyu Lyu, Lianli Gao, Jingkuan Song, Heng Tao Shen

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06744 2025-03-11 cs.CV 50%

CoDa-4DGS: Dynamic Gaussian Splatting with Context and Deformation Awareness for Autonomous Driving

Rui Song, Chenwei Liang, Yan Xia, Walter Zimmer, Hu Cao, Holger Caesar, Andreas Festag, Alois Knoll

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16981 2025-03-07 cs.CV 50%

Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection

Xiangyu Gao, Yu Dai, Benliu Qiu, Lanxiao Wang, Heqian Qiu, Hongliang Li

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02748 2025-03-05 cs.RO 50%

Bridging VLM and KMP: Enabling Fine-grained robotic manipulation via Semantic Keypoints Representation

Junjie Zhu, Huayu Liu, Jin Wang, Bangrong Wen, Kaixiang Huang, Xiaofei Li, Haiyun Zhan, Guodong Lu

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20026 2025-03-04 cs.CV 50%

Towards Robust Algorithms for Surgical Phase Recognition via Digital Twin Representation

Hao Ding, Yuqian Zhang, Wenzheng Cheng, Xinyu Wang, Xu Lian, Chenhao Yu, Hongchao Shu, Ji Woong Kim, Axel Krieger, Mathias Unberath

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20850 2025-03-03 cs.CV 50%

VLEER: Vision and Language Embeddings for Explainable Whole Slide Image Representation

Anh Tien Nguyen, Keunho Byeon, Kyungeun Kim, Jin Tae Kwak

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19293 2025-02-28 cs.CV 50%

Pathology Report Generation and Multimodal Representation Learning for Cutaneous Melanocytic Lesions

Ruben T. Lucassen, Sander P. J. Moonemans, Tijn van de Luijtgaarden, Gerben E. Breimer, Willeke A. M. Blokx, Mitko Veta

专题命中 知识编辑与模型理解 :language model(abstract)

Comments 11 pages, 2 figures. arXiv admin note: text overlap with arXiv:2502.19285

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16528 2025-02-25 cs.RO 50%

OpenVox: Real-time Instance-level Open-vocabulary Probabilistic Voxel Representation

Yinan Deng, Bicheng Yao, Yihang Tang, Yi Yang, Yufeng Yue

专题命中 知识编辑与模型理解 :language model(abstract)

Comments Project website: https://open-vox.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11729 2025-02-18 eess.IV 50%

On Quantizing Neural Representation for Variable-Rate Video Coding

Junqi Shi, Zhujia Chen, Hanfei Li, Qi Zhao, Ming Lu, Tong Chen, Zhan Ma

专题命中 知识编辑与模型理解 :post-training(abstract)

Comments to be pulished in ICLR'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.13785 2025-02-18 cs.CV 50%

A Spatiotemporal Approach to Tri-Perspective Representation for 3D Semantic Occupancy Prediction

Sathira Silva, Savindu Bhashitha Wannigama, Gihan Jayatilaka, Muhammad Haris Khan, Roshan Ragel

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted to the 2025 Workshop on Machine Learning for Autonomous Driving at AAAI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.00250 2025-02-04 cs.CV 50%

Transformer-Based Vector Font Classification Using Different Font Formats: TrueType versus PostScript

Takumu Fujioka, Gouhei Tanaka

专题命中 知识编辑与模型理解 :language model(abstract)

Comments 8 pages, 8 figures, 4 tables, Submitted to IJCNN 2025. Code available at https://github.com/fjktkm/truetype-vs-postscript-transformer

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13449 2025-01-24 cs.CV 50%

MultiDreamer3D: Multi-concept 3D Customization with Concept-Aware Diffusion Guidance

Wooseok Song, Seunggyu Chang, Jaejun Yoo

专题命中 知识编辑与模型理解 :LLM(abstract)

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.17857 2025-01-22 cs.CV 50%

SAGD: Boundary-Enhanced Segment Anything in 3D Gaussian via Gaussian Decomposition

Xu Hu, Yuxi Wang, Lue Fan, Chuanchen Luo, Junsong Fan, Zhen Lei, Qing Li, Junran Peng, Zhaoxiang Zhang

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.10144 2025-01-20 cs.CV 50%

A Vision-Language Framework for Multispectral Scene Representation Using Language-Grounded Features

Enes Karanfil, Nevrez Imamoglu, Erkut Erdem, Aykut Erdem

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04416 2025-01-09 eess.AS cs.SD 50%

ZSVC: Zero-shot Style Voice Conversion with Disentangled Latent Diffusion Models and Adversarial Training

Xinfa Zhu, Lei He, Yujia Xiao, Xi Wang, Xu Tan, Sheng Zhao, Lei Xie

专题命中 知识编辑与模型理解 :prompting(abstract)

Comments 5 pages, 3 figures, accepted by ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03469 2025-01-08 cs.CV 50%

Information-Maximized Soft Variable Discretization for Self-Supervised Image Representation Learning

Chuang Niu, Wenjun Xia, Hongming Shan, Ge Wang

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.12032 2025-01-07 math.ST stat.TH 50%

Clusterization in D-optimal designs: the case against linearization

Yair Daon

专题命中 知识编辑与模型理解 :prompting(abstract)

Comments 29 pages, 5 figures, to be published in Bayesian Analysis. Code in https://github.com/yairdaon/OED

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23754 2024-12-31 cs.HC q-bio.NC 50%

RealMind: Advancing Visual Decoding and Language Interaction via EEG Signals

Dongyang Li, Haoyang Qin, Mingyang Wu, Jiahua Tang, Yuang Cao, Chen Wei, Quanying Liu

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18604 2024-12-25 cs.CV 50%

Explaining in Diffusion: Explaining a Classifier Through Hierarchical Semantics with Text-to-Image Diffusion Models

Tahira Kazimi, Ritika Allada, Pinar Yanardag

专题命中 知识编辑与模型理解 :language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.07458 2024-12-18 physics.acc-ph 50%

Uncertainty Aware ML-based surrogate models for particle accelerators: A Study at the Fermilab Booster Accelerator Complex

Malachi Schram, Kishansingh Rajput, Karthik Somayaji Peng Li, Jason St. John, Himanshu Sharma

专题命中 知识编辑与模型理解 :post-training(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17474 2024-12-17 cs.CV 50%

Probing the Mid-level Vision Capabilities of Self-Supervised Learning

Xuweiyi Chen, Markus Marks, Zezhou Cheng

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Project Page: https://midvision-probe.cs.virginia.edu/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10431 2024-12-17 cs.CV cs.RO 50%

CUPS: Improving Human Pose-Shape Estimators with Conformalized Deep Uncertainty

Harry Zhang, Luca Carlone

专题命中 知识编辑与模型理解 :post-training(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09953 2024-12-10 physics.ins-det 50%

Photocathode characterisation for robust PICOSEC Micromegas precise-timing detectors

M. Lisowska, R. Aleksan, Y. Angelis, S. Aune, J. Bortfeldt, F. Brunbauer, M. Brunoldi, E. Chatzianagnostou, J. Datta, K. Dehmelt, G. Fanourakis, S. Ferry, D. Fiorina, K. J. Floethner, M. Gallinaro, F. Garcia, I. Giomataris, K. Gnanvo, F. J. Iguaz, D. Janssens, A. Kallitsopoulou, M. Kovacic, B. Kross, C. C. Lai, P. Legou, J. Liu, M. Lupberger, I. Maniatis, J. McKisson, Y. Meng, H. Muller, R. De Oliveira, E. Oliveri, G. Orlandini, A. Pandey, T. Papaevangelou, M. Pomorski, M. Robert, L. Ropelewski, D. Sampsonidis, L. Scharenberg, T. Schneider, E. Scorsone, L. Sohl, M. van Stenis, Y. Tsipolitis, S. Tzamarias, A. Utrobicic, I. Vai, R. Veenhof, L. Viezzi, P. Vitulo, C. Volpato, X. Wang, S. White, W. Xi, Z. Zhang, Y. Zhou

专题命中 知识编辑与模型理解 :prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09401 2024-12-06 cs.CV 50%

Unsupervised Modality-Transferable Video Highlight Detection with Representation Activation Sequence Learning

Tingtian Li, Zixun Sun, Xinyu Xiao

专题命中 知识编辑与模型理解 :pretraining(abstract)

Comments Accepted by IEEE Transactions on Image Processing, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03314 2024-12-05 cs.CV 50%

Equivariant Representation Learning for Augmentation-based Self-Supervised Learning via Image Reconstruction

Qin Wang, Kai Krajsek, Hanno Scharr

专题命中 知识编辑与模型理解 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏