arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2401.01482 2024-04-02 cs.CV cs.AI cs.LG

Incorporating Geo-Diverse Knowledge into Prompting for Increased Geographical Robustness in Object Recognition

Kyle Buettner, Sina Malakouti, Xiang Lorraine Li, Adriana Kovashka

Comments To appear in IEEE/CVF Computer Vision and Pattern Recognition Conference (CVPR), 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.00374 2024-04-02 cs.CV

EMAGE: Towards Unified Holistic Co-Speech Gesture Generation via Expressive Masked Audio Gesture Modeling

Haiyang Liu, Zihao Zhu, Giorgio Becherini, Yichen Peng, Mingyang Su, You Zhou, Xuefei Zhe, Naoya Iwamoto, Bo Zheng, Michael J. Black

Comments Fix typos; Conflict of Interest Disclosure; CVPR Camera Ready; Project Page: https://pantomatrix.github.io/EMAGE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.16084 2024-04-02 cs.CV

LangSplat: 3D Language Gaussian Splatting

Minghan Qin, Wanhua Li, Jiawei Zhou, Haoqian Wang, Hanspeter Pfister

Comments CVPR 2024. Project Page: https://langsplat.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.14985 2024-04-02 cs.CV

UniHuman: A Unified Model for Editing Human Images in the Wild

Nannan Li, Qing Liu, Krishna Kumar Singh, Yilin Wang, Jianming Zhang, Bryan A. Plummer, Zhe Lin

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09788 2024-04-02 cs.CV cs.AI cs.LG

Collaborating Foundation Models for Domain Generalized Semantic Segmentation

Yasser Benigmim, Subhankar Roy, Slim Essid, Vicky Kalogeiton, Stéphane Lathuilière

Comments https://github.com/yasserben/CLOUDS ; Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08869 2024-04-02 cs.CV

I'M HOI: Inertia-aware Monocular Capture of 3D Human-Object Interactions

Chengfeng Zhao, Juze Zhang, Jiashen Du, Ziwei Shan, Junye Wang, Jingyi Yu, Jingya Wang, Lan Xu

Comments Accepted to CVPR 2024. Project page: https://afterjourney00.github.io/IM-HOI.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.06742 2024-04-02 cs.CV cs.AI cs.CL cs.LG

Honeybee: Locality-enhanced Projector for Multimodal LLM

Junbum Cha, Wooyoung Kang, Jonghwan Mun, Byungseok Roh

Comments CVPR 2024 camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05716 2024-04-02 cs.CV

Initialization Matters for Adversarial Transfer Learning

Andong Hua, Jindong Gu, Zhiyu Xue, Nicholas Carlini, Eric Wong, Yao Qin

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05264 2024-04-02 cs.CR cs.LG

All Rivers Run to the Sea: Private Learning with Asymmetric Flows

Yue Niu, Ramy E. Ali, Saurav Prakash, Salman Avestimehr

Comments Camera-ready for CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04553 2024-04-02 cs.CV eess.IV

SPIDeRS: Structured Polarization for Invisible Depth and Reflectance Sensing

Tomoki Ichikawa, Shohei Nobuhara, Ko Nishino

Comments to be published in CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02142 2024-04-02 cs.CV

Object Recognition as Next Token Prediction

Kaiyu Yue, Bor-Chun Chen, Jonas Geiping, Hengduo Li, Tom Goldstein, Ser-Nam Lim

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02010 2024-04-02 cs.CV cs.AI

Towards Learning a Generalist Model for Embodied Navigation

Duo Zheng, Shijia Huang, Lin Zhao, Yiwu Zhong, Liwei Wang

Comments Accepted by CVPR 2024 (14 pages, 3 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01998 2024-04-02 cs.CV cs.IR

Language-only Efficient Training of Zero-shot Composed Image Retrieval

Geonmo Gu, Sanghyuk Chun, Wonjae Kim, Yoohoon Kang, Sangdoo Yun

Comments CVPR 2024 camera-ready; First two authors contributed equally; 17 pages, 3.1MB

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01196 2024-04-02 cs.CV

Neural Parametric Gaussians for Monocular Non-Rigid Object Reconstruction

Devikalyan Das, Christopher Wewer, Raza Yunus, Eddy Ilg, Jan Eric Lenssen

Comments Accepted at CVPR 2024 | Project Website: https://geometric-rl.mpi-inf.mpg.de/npg

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00600 2024-04-02 cs.LG

Improving Plasticity in Online Continual Learning via Collaborative Learning

Maorong Wang, Nicolas Michel, Ling Xiao, Toshihiko Yamasaki

Comments Update Camera-ready revision for CVPR'24

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00081 2024-04-02 cs.CV

Synthesize, Diagnose, and Optimize: Towards Fine-Grained Vision-Language Understanding

Wujian Peng, Sicheng Xie, Zuyao You, Shiyi Lan, Zuxuan Wu

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17516 2024-04-02 cs.CR cs.CV

MMA-Diffusion: MultiModal Attack on Diffusion Models

Yijun Yang, Ruiyuan Gao, Xiaosen Wang, Tsung-Yi Ho, Nan Xu, Qiang Xu

Comments CVPR 2024. Our codes and benchmarks are available at https://github.com/cure-lab/MMA-Diffusion

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15855 2024-04-02 cs.CV

SiTH: Single-view Textured Human Reconstruction with Image-Conditioned Diffusion

Hsuan-I Ho, Jie Song, Otmar Hilliges

Comments 23 pages, 23 figures, CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.12194 2024-04-02 cs.CV

DiffAvatar: Simulation-Ready Garment Optimization with Differentiable Simulation

Yifei Li, Hsiao-yu Chen, Egor Larionov, Nikolaos Sarafianos, Wojciech Matusik, Tuur Stuyck

Comments CVPR 2024; Project page: https://people.csail.mit.edu/liyifei/publication/diffavatar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.03149 2024-04-02 cs.CV

Asymmetric Masked Distillation for Pre-Training Small Foundation Models

Zhiyu Zhao, Bingkun Huang, Sen Xing, Gangshan Wu, Yu Qiao, Limin Wang

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11281 2024-04-02 cs.CV

Language-driven Object Fusion into Neural Radiance Fields with Pose-Conditioned Dataset Updates

Ka Chun Shum, Jaeyeon Kim, Binh-Son Hua, Duc Thanh Nguyen, Sai-Kit Yeung

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01327 2024-04-02 cs.CV cs.AI cs.MM

Can I Trust Your Answer? Visually Grounded Video Question Answering

Junbin Xiao, Angela Yao, Yicong Li, Tat Seng Chua

Comments Accepted to CVPR'24. (Compared with preprint version, we mainly improve the presentation, discuss more related works, and extend experiments in Appendix.)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08919 2024-04-02 cs.CV cs.LG

Systematic comparison of semi-supervised and self-supervised learning for medical image classification

Zhe Huang, Ruijie Jiang, Shuchin Aeron, Michael C. Hughes

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.07735 2024-04-02 cs.CR

Permutation Equivariance of Transformers and Its Applications

Hengyuan Xu, Liyao Xiang, Hangyu Ye, Dixi Yao, Pengzhi Chu, Baochun Li

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13612 2024-04-02 cs.CV

NOPE: Novel Object Pose Estimation from a Single Image

Van Nguyen Nguyen, Thibault Groueix, Yinlin Hu, Mathieu Salzmann, Vincent Lepetit

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.11797 2024-04-02 cs.CV

CAT-Seg: Cost Aggregation for Open-Vocabulary Semantic Segmentation

Seokju Cho, Heeseong Shin, Sunghwan Hong, Anurag Arnab, Paul Hongsuck Seo, Seungryong Kim

Comments Accepted to CVPR 2024. Project page: https://ku-cvlab.github.io/CAT-Seg/

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09383 2024-04-02 cs.CV cs.AI

Unifying Top-down and Bottom-up Scanpath Prediction Using Transformers

Zhibo Yang, Sounak Mondal, Seoyoung Ahn, Ruoyu Xue, Gregory Zelinsky, Minh Hoai, Dimitris Samaras

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09373 2024-04-02 cs.CV cs.AI cs.LG

MAPSeg: Unified Unsupervised Domain Adaptation for Heterogeneous Medical Image Segmentation Based on 3D Masked Autoencoding and Pseudo-Labeling

Xuzhe Zhang, Yuhao Wu, Elsa Angelini, Ang Li, Jia Guo, Jerod M. Rasmussen, Thomas G. O'Connor, Pathik D. Wadhwa, Andrea Parolin Jackowski, Hai Li, Jonathan Posner, Andrew F. Laine, Yun Wang

Comments CVPR 2024 camera-ready (8 pages, 3 figures) with the supplemental materials (5 pages, 4 figures). Xuzhe Zhang and Yuhao Wu are co-first authors. Andrew F. Laine and Yun Wang are co-senior supervising authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08314 2024-04-02 cs.CV

Guided Slot Attention for Unsupervised Video Object Segmentation

Minhyeok Lee, Suhwan Cho, Dogyoon Lee, Chaewon Park, Jungho Lee, Sangyoun Lee

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.02835 2024-04-02 cs.CV

Traffic Scene Parsing through the TSP6K Dataset

Peng-Tao Jiang, Yuqi Yang, Yang Cao, Qibin Hou, Ming-Ming Cheng, Chunhua Shen

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏