International Conference on Computer Vision · 会议 · Computer Vision
至 收录 4770 篇
2507.016542026-03-09cs.CVcs.LG
SPoT: Subpixel Placement of Tokens in Vision Transformers
SPoT: 视觉变换器中令牌的子像素位置
Martine Hjelkrem-Tan, Marius Aasan, Gabriel Y. Arteaga, Adín Ramírez Rivera
AI总结
SPoT通过子像素位置策略提升视觉变换器的稀疏性利用,减少令牌数量以提高效率与准确性。
CommentsAppeared in Workshop on Efficient Computing under Limited Resources: Visual Computing (ICCV 2025). Code available at https://github.com/dsb-ifi/SPoT
UNet-Based Keypoint Regression for 3D Cone Localization in Autonomous Racing
基于UNet的关键点回归用于自动驾驶赛车中3D圆锥定位
Mariia Baidachna, James Carty, Aidan Ferguson, Joseph Agrane, Varad Kulkarni, Aubrey Agub, Michael Baxendale, Aaron David, Rachel Horton, Elliott Atkinson
机构
*
School of Computer Science, University of Glasgow(计算机科学学院,格拉斯哥大学)
;
Amazon(亚马逊)
Comments8 pages, 9 figures. Accepted to ICCV End-to-End 3D Learning Workshop 2025 and presented as a poster; not included in the final proceedings due to a conference administrative error
EHWGesture -- A dataset for multimodal understanding of clinical gestures
EHWGesture -- 一个用于多模态理解临床手势的数据集
Gianluca Amprimo, Alberto Ancilotto, Alessandro Savino, Fabio Quazzolo, Claudia Ferraris, Gabriella Olmo, Elisabetta Farella, Stefano Di Carlo
机构
*
Department of Control and Computer Engineering, Politecnico di Torino(控制与计算机工程系,都灵理工大学)
;
Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会)
;
CNR-IEIIT(意大利国家研究委员会-IEIIT)
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电子与计算机工程系)
;
Amazon(亚马逊)
;
Samsung Research(三星研究院)
;
School of Computer Science and Engineering, Soongsil University(顺天大学计算机科学与工程学院)
;
IPAI, AIIS, ASRI, INMC, and ISRC, Seoul National University(首尔国立大学IPAI、AIIS、ASRI、INMC和ISRC)
机构
*
Drexel University(德雷塞尔大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
LLNL(劳伦斯利弗莫尔国家实验室)
;
Lehigh University(莱斯大学)
Image-to-Image Translation with Diffusion Transformers and CLIP-Based Image Conditioning
基于扩散变换器和CLIP的图像到图像翻译
Qiang Zhu, Kuan Lu, Menghao Huo, Yuxiao Li
机构
*
Department of Mechanical and Aerospace Engineering(机械与航空航天工程系)
;
University of Houston(休斯顿大学)
;
School of Engineering(工程学院)
;
Santa Clara University(圣克拉拉大学)
;
School of Electrical and Computer Engineering(电气与计算机工程学院)
;
Cornell University(康奈尔大学)
;
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
Northeastern University(东北大学)
Tune-Your-Style: Intensity-tunable 3D Style Transfer with Gaussian Splatting
Tune-Your-Style: 可调强度的3D风格迁移与高斯点散布
Yian Zhao, Rushi Ye, Ruochong Zheng, Zesen Cheng, Chaoran Feng, Jiashu Yang, Pengchong Qiao, Chang Liu, Jie Chen
机构
*
School of Electronic and Computer Engineering, Peking University, Shenzhen, China(电子与计算机工程学院,北京大学深圳校区)
;
Pengcheng Laboratory, Shenzhen, China(鹏城实验室)
;
AI for Science (AI4S)-Preferred Program, Peking University Shenzhen Graduate School, China(人工智能科学(AI4S)优选计划,北京大学深圳研究生院)
;
Department of Automation and BNRist, Tsinghua University, Beijing, China(自动化系和BNRist,清华大学,北京,中国)
;
Dalian University of Technology, China(大连理工大学)
DeepShield: Fortifying Deepfake Video Detection with Local and Global Forgery Analysis
DeepShield: 通过局部和全局伪造分析强化深度伪造视频检测
Yinqi Cai, Jichang Li, Zhaolun Li, Weikai Chen, Rushi Lan, Xi Xie, Xiaonan Luo, Guanbin Li
机构
*
Sun Yat-sen University(中山大学)
;
Pengcheng Laboratory(鹏城实验室)
;
Guilin University of Electronic Technology(桂林电子科技大学)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)