arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

Trevor Darrell

Computer Vision

至 收录 357
2306.11180 2024-06-05 cs.CV cs.AI

Hyperbolic Active Learning for Semantic Segmentation under Domain Shift

Luca Franco, Paolo Mandica, Konstantinos Kallidromitis, Devin Guillory, Yu-Teng Li, Trevor Darrell, Fabio Galasso

Comments ICML 2024. Project repository: https://github.com/paolomandica/HALO

URL PDF HTML 收藏
2405.08597 2024-05-30 cs.LG

Risks and Opportunities of Open-Source Generative AI

Francisco Eiras, Aleksandar Petrov, Bertie Vidgen, Christian Schroeder, Fabio Pizzati, Katherine Elkins, Supratik Mukhopadhyay, Adel Bibi, Aaron Purewal, Csaba Botos, Fabro Steibel, Fazel Keshtkar, Fazl Barez, Genevieve Smith, Gianluca Guadagni, Jon Chun, Jordi Cabot, Joseph Imperial, Juan Arturo Nolazco, Lori Landay, Matthew Jackson, Phillip H. S. Torr, Trevor Darrell, Yong Lee, Jakob Foerster

Comments Extension of arXiv:2404.17047

URL PDF HTML 收藏
2404.17047 2024-05-27 cs.LG

Near to Mid-term Risks and Opportunities of Open-Source Generative AI

Francisco Eiras, Aleksandar Petrov, Bertie Vidgen, Christian Schroeder de Witt, Fabio Pizzati, Katherine Elkins, Supratik Mukhopadhyay, Adel Bibi, Botos Csaba, Fabro Steibel, Fazl Barez, Genevieve Smith, Gianluca Guadagni, Jon Chun, Jordi Cabot, Joseph Marvin Imperial, Juan A. Nolazco-Flores, Lori Landay, Matthew Jackson, Paul Röttger, Philip H. S. Torr, Trevor Darrell, Yong Suk Lee, Jakob Foerster

Comments Accepted to ICML'24 as a position paper

URL PDF HTML 收藏
2310.17688 2024-05-24 cs.CY cs.AI cs.CL cs.LG

Managing extreme AI risks amid rapid progress

Yoshua Bengio, Geoffrey Hinton, Andrew Yao, Dawn Song, Pieter Abbeel, Trevor Darrell, Yuval Noah Harari, Ya-Qin Zhang, Lan Xue, Shai Shalev-Shwartz, Gillian Hadfield, Jeff Clune, Tegan Maharaj, Frank Hutter, Atılım Güneş Baydin, Sheila McIlraith, Qiqi Gao, Ashwin Acharya, David Krueger, Anca Dragan, Philip Torr, Stuart Russell, Daniel Kahneman, Jan Brauner, Sören Mindermann

Comments Published in Science: https://www.science.org/doi/10.1126/science.adn0117

URL PDF HTML 收藏
2309.17444 2024-05-07 cs.CV cs.AI cs.CL

LLM-grounded Video Diffusion Models

Long Lian, Baifeng Shi, Adam Yala, Trevor Darrell, Boyi Li

Comments ICLR 2024. Project Page: https://llm-grounded-video-diffusion.github.io/

URL PDF HTML 收藏
2312.02974 2024-04-30 cs.CV cs.CL cs.CY cs.LG

Describing Differences in Image Sets with Natural Language

Lisa Dunlap, Yuhui Zhang, Xiaohan Wang, Ruiqi Zhong, Trevor Darrell, Jacob Steinhardt, Joseph E. Gonzalez, Serena Yeung-Levy

Comments CVPR 2024 Oral

URL PDF HTML 收藏
2404.09991 2024-04-16 cs.RO cs.CV

EgoPet: Egomotion and Interaction Data from an Animal's Perspective

Amir Bar, Arya Bakhtiar, Danny Tran, Antonio Loquercio, Jathushan Rajasegaran, Yann LeCun, Amir Globerson, Trevor Darrell

Comments https://www.amirbar.net/egopet

URL PDF HTML 收藏
2212.10564 2024-04-15 cs.CL cs.AI cs.LG

Re-evaluating the Need for Multimodal Signals in Unsupervised Grammar Induction

Boyi Li, Rodolfo Corona, Karttikeya Mangalam, Catherine Chen, Daniel Flaherty, Serge Belongie, Kilian Q. Weinberger, Jitendra Malik, Trevor Darrell, Dan Klein

Comments NAACL Findings 2024

URL PDF HTML 收藏
2311.06694 2024-04-11 cs.CL cs.AI cs.CV cs.RO

Which One? Leveraging Context Between Objects and Multiple Views for Language Grounding

Chancharik Mitra, Abrar Anwar, Rodolfo Corona, Dan Klein, Trevor Darrell, Jesse Thomason

Journal ref North American Chapter of the Association for Computational Linguistics (NAACL), 2024

URL PDF HTML 收藏
2303.17546 2024-04-10 cs.CV cs.AI cs.LG

PAIR-Diffusion: A Comprehensive Multimodal Object-Level Image Editor

Vidit Goel, Elia Peruzzo, Yifan Jiang, Dejia Xu, Xingqian Xu, Nicu Sebe, Trevor Darrell, Zhangyang Wang, Humphrey Shi

Comments Accepted in CVPR 2024, Project page https://vidit98.github.io/publication/conference-paper/pair_diff.html

URL PDF HTML 收藏
2404.02904 2024-04-04 cs.CV cs.AI cs.CL cs.LG

ALOHa: A New Measure for Hallucination in Captioning Models

Suzanne Petryk, David M. Chan, Anish Kachinthaya, Haodi Zou, John Canny, Joseph E. Gonzalez, Trevor Darrell

Comments To appear at NAACL 2024

URL PDF HTML 收藏
2312.02150 2024-04-04 cs.CV

Readout Guidance: Learning Control from Diffusion Features

Grace Luo, Trevor Darrell, Oliver Wang, Dan B Goldman, Aleksander Holynski

Comments CVPR 2024

URL PDF HTML 收藏
2305.14334 2024-04-03 cs.CV

Diffusion Hyperfeatures: Searching Through Time and Space for Semantic Correspondence

Grace Luo, Lisa Dunlap, Dong Huk Park, Aleksander Holynski, Trevor Darrell

Comments NeurIPS 2023

URL PDF HTML 收藏
2212.00210 2024-04-02 cs.CV cs.AI cs.LG

Shape-Guided Diffusion with Inside-Outside Attention

Dong Huk Park, Grace Luo, Clayton Toste, Samaneh Azadi, Xihui Liu, Maka Karalashvili, Anna Rohrbach, Trevor Darrell

Comments WACV 2024

URL PDF HTML 收藏
2311.17076 2024-04-02 cs.CV cs.AI cs.CL cs.LG

Compositional Chain-of-Thought Prompting for Large Multimodal Models

Chancharik Mitra, Brandon Huang, Trevor Darrell, Roei Herzig

URL PDF HTML 收藏
2305.13655 2024-03-05 cs.CV

LLM-grounded Diffusion: Enhancing Prompt Understanding of Text-to-Image Diffusion Models with Large Language Models

Long Lian, Boyi Li, Adam Yala, Trevor Darrell

Comments Transactions on Machine Learning Research (TMLR) 2024, with Featured Certification

URL PDF HTML 收藏
2402.19469 2024-03-01 cs.RO cs.CV cs.LG

Humanoid Locomotion as Next Token Prediction

Ilija Radosavovic, Bike Zhang, Baifeng Shi, Jathushan Rajasegaran, Sarthak Kamat, Trevor Darrell, Koushil Sreenath, Jitendra Malik

URL PDF HTML 收藏
2308.00566 2024-02-28 cs.CV cs.AI cs.LG

Stochastic positional embeddings improve masked image modeling

Amir Bar, Florian Bordes, Assaf Shocher, Mahmoud Assran, Pascal Vincent, Nicolas Ballas, Trevor Darrell, Amir Globerson, Yann LeCun

Comments Code and models available in https://github.com/amirbar/StoP

URL PDF HTML 收藏
2402.03290 2024-02-06 cs.CV cs.AI cs.LG

InstanceDiffusion: Instance-level Control for Image Generation

Xudong Wang, Trevor Darrell, Sai Saketh Rambhatla, Rohit Girdhar, Ishan Misra

Comments Preprint; Project page: https://people.eecs.berkeley.edu/~xdwang/projects/InstDiff/

URL PDF HTML 收藏
2401.01885 2024-01-04 cs.CV

From Audio to Photoreal Embodiment: Synthesizing Humans in Conversations

Evonne Ng, Javier Romero, Timur Bagautdinov, Shaojie Bai, Trevor Darrell, Angjoo Kanazawa, Alexander Richard

URL PDF HTML 收藏
2312.17243 2023-12-29 cs.CV

Unsupervised Universal Image Segmentation

Dantong Niu, Xudong Wang, Xinyang Han, Long Lian, Roei Herzig, Trevor Darrell

URL PDF HTML 收藏
2307.00764 2023-12-22 cs.CV cs.AI cs.LG

Hierarchical Open-vocabulary Universal Image Segmentation

Xudong Wang, Shufan Li, Konstantinos Kallidromitis, Yusuke Kato, Kazuki Kozuka, Trevor Darrell

Comments Project web-page: http://people.eecs.berkeley.edu/~xdwang/projects/HIPIE/; NeurIPS 2023 Camera-ready

URL PDF HTML 收藏
2306.10007 2023-12-15 cs.RO cs.CV cs.LG

Robot Learning with Sensorimotor Pre-training

Ilija Radosavovic, Baifeng Shi, Letian Fu, Ken Goldberg, Trevor Darrell, Jitendra Malik

Comments CoRL 2023; Project page: https://robotic-pretrained-transformer.github.io

URL PDF HTML 收藏
2303.03381 2023-12-15 cs.RO cs.LG

Real-World Humanoid Locomotion with Reinforcement Learning

Ilija Radosavovic, Tete Xiao, Bike Zhang, Trevor Darrell, Jitendra Malik, Koushil Sreenath

Comments Project page: https://learning-humanoid-locomotion.github.io

URL PDF HTML 收藏
2312.08366 2023-12-14 cs.CV

See, Say, and Segment: Teaching LMMs to Overcome False Premises

Tsung-Han Wu, Giscard Biamby, David Chan, Lisa Dunlap, Ritwik Gupta, Xudong Wang, Joseph E. Gonzalez, Trevor Darrell

Comments Project Page: https://see-say-segment.github.io

URL PDF HTML 收藏
2212.04821 2023-12-08 cs.CV

PromptonomyViT: Multi-Task Prompt Learning Improves Video Transformers using Synthetic Scene Data

Roei Herzig, Ofir Abramovich, Elad Ben-Avraham, Assaf Arbelle, Leonid Karlinsky, Ariel Shamir, Trevor Darrell, Amir Globerson

Comments WACV 2024

URL PDF HTML 收藏
2312.01771 2023-12-05 cs.CV

IMProv: Inpainting-based Multimodal Prompting for Computer Vision Tasks

Jiarui Xu, Yossi Gandelsman, Amir Bar, Jianwei Yang, Jianfeng Gao, Trevor Darrell, Xiaolong Wang

Comments Project page: https://jerryxu.net/IMProv

URL PDF HTML 收藏
2312.00785 2023-12-04 cs.CV

Sequential Modeling Enables Scalable Learning for Large Vision Models

Yutong Bai, Xinyang Geng, Karttikeya Mangalam, Amir Bar, Alan Yuille, Trevor Darrell, Jitendra Malik, Alexei A Efros

Comments Website: https://yutongbai.com/lvm.html

URL PDF HTML 收藏
2311.18823 2023-12-01 cs.LG cs.CV

Initializing Models with Larger Ones

Zhiqiu Xu, Yanjie Chen, Kirill Vishniakov, Yida Yin, Zhiqiang Shen, Trevor Darrell, Lingjie Liu, Zhuang Liu

URL PDF HTML 收藏
2311.17942 2023-12-01 cs.CV

Object-based (yet Class-agnostic) Video Domain Adaptation

Dantong Niu, Amir Bar, Roei Herzig, Trevor Darrell, Anna Rohrbach

URL PDF HTML 收藏