arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

作者

Trevor Darrell

Computer Vision

至 收录 357
2311.16090 2023-11-28 cs.CV

Self-correcting LLM-controlled Diffusion Models

Tsung-Han Wu, Long Lian, Joseph E. Gonzalez, Boyi Li, Trevor Darrell

Comments 16 pages, 10 figures

URL PDF HTML 收藏
2311.12391 2023-11-22 cs.CV

From Wrong To Right: A Recursive Approach Towards Vision-Language Explanation

Jiaxin Ge, Sanjay Subramanian, Trevor Darrell, Boyi Li

Comments EMNLP 2023 Main

URL PDF HTML 收藏
2311.01011 2023-11-03 cs.LG cs.CR

Tensor Trust: Interpretable Prompt Injection Attacks from an Online Game

Sam Toyer, Olivia Watkins, Ethan Adrian Mendes, Justin Svegliato, Luke Bailey, Tiffany Wang, Isaac Ong, Karim Elmaaroufi, Pieter Abbeel, Trevor Darrell, Alan Ritter, Stuart Russell

URL PDF HTML 收藏
2305.16289 2023-10-31 cs.CV cs.AI

Diversify Your Vision Datasets with Automatic Diffusion-Based Augmentation

Lisa Dunlap, Alyssa Umino, Han Zhang, Jiezhi Yang, Joseph E. Gonzalez, Trevor Darrell

Comments Update: replaced Planes dataset with Waterbirds & updated results after bug fix

URL PDF HTML 收藏
2310.12971 2023-10-26 cs.CV cs.AI cs.CL

CLAIR: Evaluating Image Captions with Large Language Models

David Chan, Suzanne Petryk, Joseph E. Gonzalez, Trevor Darrell, John Canny

Comments To Appear at EMNLP 2023

URL PDF HTML 收藏
2305.06343 2023-10-26 cs.CV

Incorporating Structured Representations into Pretrained Vision & Language Models Using Scene Graphs

Roei Herzig, Alon Mendelson, Leonid Karlinsky, Assaf Arbelle, Rogerio Feris, Trevor Darrell, Amir Globerson

Comments EMNLP 2023

URL PDF HTML 收藏
2310.15166 2023-10-24 cs.CV cs.CL

Large Language Models are Visual Reasoning Coordinators

Liangyu Chen, Bo Li, Sheng Shen, Jingkang Yang, Chunyuan Li, Kurt Keutzer, Trevor Darrell, Ziwei Liu

Comments Accepted at NeurIPS 2023

URL PDF HTML 收藏
2210.06984 2023-09-29 cs.CV

QDTrack: Quasi-Dense Similarity Learning for Appearance-Only Multiple Object Tracking

Tobias Fischer, Thomas E. Huang, Jiangmiao Pang, Linlu Qiu, Haofeng Chen, Trevor Darrell, Fisher Yu

URL PDF HTML 收藏
2309.14525 2023-09-27 cs.CV cs.CL

Aligning Large Multimodal Models with Factually Augmented RLHF

Zhiqing Sun, Sheng Shen, Shengcao Cao, Haotian Liu, Chunyuan Li, Yikang Shen, Chuang Gan, Liang-Yan Gui, Yu-Xiong Wang, Yiming Yang, Kurt Keutzer, Trevor Darrell

Comments Preprint

URL PDF HTML 收藏
2212.14532 2023-09-25 cs.CV

Scale-MAE: A Scale-Aware Masked Autoencoder for Multiscale Geospatial Representation Learning

Colorado J. Reed, Ritwik Gupta, Shufan Li, Sarah Brockman, Christopher Funk, Brian Clipp, Kurt Keutzer, Salvatore Candido, Matt Uyttendaele, Trevor Darrell

Comments International Conference on Computer Vision 2023

URL PDF HTML 收藏
2302.06692 2023-09-18 cs.LG cs.AI cs.CL

Guiding Pretraining in Reinforcement Learning with Large Language Models

Yuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas, Trevor Darrell, Pieter Abbeel, Abhishek Gupta, Jacob Andreas

Comments ICML 2023

URL PDF HTML 收藏
2308.14710 2023-08-29 cs.CV cs.AI cs.LG

VideoCutLER: Surprisingly Simple Unsupervised Video Instance Segmentation

Xudong Wang, Ishan Misra, Ziyun Zeng, Rohit Girdhar, Trevor Darrell

Comments Preprint. Code: https://github.com/facebookresearch/CutLER

URL PDF HTML 收藏
2106.04550 2023-07-21 cs.CV

DETReg: Unsupervised Pretraining with Region Priors for Object Detection

Amir Bar, Xin Wang, Vadim Kantorov, Colorado J Reed, Roei Herzig, Gal Chechik, Anna Rohrbach, Trevor Darrell, Amir Globerson

Comments Project page: https://www.amirbar.net/detreg/

URL PDF HTML 收藏
2305.15542 2023-07-12 cs.CV cs.CL cs.LG

TOAST: Transfer Learning via Attention Steering

Baifeng Shi, Siyu Gai, Trevor Darrell, Xin Wang

Comments Code is available at https://github.com/bfshi/TOAST

URL PDF HTML 收藏
2305.14705 2023-07-06 cs.CL

Mixture-of-Experts Meets Instruction Tuning:A Winning Combination for Large Language Models

Sheng Shen, Le Hou, Yanqi Zhou, Nan Du, Shayne Longpre, Jason Wei, Hyung Won Chung, Barret Zoph, William Fedus, Xinyun Chen, Tu Vu, Yuexin Wu, Wuyang Chen, Albert Webson, Yunxuan Li, Vincent Zhao, Hongkun Yu, Kurt Keutzer, Trevor Darrell, Denny Zhou

Comments Preprint

URL PDF HTML 收藏
2207.03442 2023-06-22 cs.LG cs.CV

Back to the Source: Diffusion-Driven Test-Time Adaptation

Jin Gao, Jialing Zhang, Xihui Liu, Trevor Darrell, Evan Shelhamer, Dequan Wang

Comments published at CVPR 2023

URL PDF HTML 收藏
2306.05392 2023-06-09 cs.CL

Modular Visual Question Answering via Code Generation

Sanjay Subramanian, Medhini Narasimhan, Kushal Khangaonkar, Kevin Yang, Arsha Nagrani, Cordelia Schmid, Andy Zeng, Trevor Darrell, Dan Klein

Comments ACL 2023

URL PDF HTML 收藏
2303.01500 2023-06-01 cs.LG cs.AI cs.CV

Dropout Reduces Underfitting

Zhuang Liu, Zhiqiu Xu, Joseph Jin, Zhiqiang Shen, Trevor Darrell

Comments ICML 2023

URL PDF HTML 收藏
2305.07021 2023-05-12 cs.CV

Simple Token-Level Confidence Improves Caption Correctness

Suzanne Petryk, Spencer Whitehead, Joseph E. Gonzalez, Trevor Darrell, Anna Rohrbach, Marcus Rohrbach

URL PDF HTML 收藏
2210.09520 2023-05-02 cs.CV

Using Language to Extend to Unseen Domains

Lisa Dunlap, Clara Mohri, Devin Guillory, Han Zhang, Trevor Darrell, Joseph E. Gonzalez, Aditi Raghunathan, Anja Rohrbach

URL PDF HTML 收藏
2303.13043 2023-03-27 cs.CV

Top-Down Visual Attention from Analysis by Synthesis

Baifeng Shi, Trevor Darrell, Xin Wang

Comments CVPR2023 highlight; Project page: https://sites.google.com/view/absvit

URL PDF HTML 收藏
2303.13519 2023-03-24 cs.CV cs.AI cs.CL cs.LG

Learning and Verification of Task Structure in Instructional Videos

Medhini Narasimhan, Licheng Yu, Sean Bell, Ning Zhang, Trevor Darrell

Comments Wesbite at https://medhini.github.io/task_structure

URL PDF HTML 收藏
2303.07226 2023-03-14 cs.CV cs.CL

Scaling Vision-Language Models with Sparse Mixture of Experts

Sheng Shen, Zhewei Yao, Chunyuan Li, Trevor Darrell, Kurt Keutzer, Yuxiong He

Comments Preprint

URL PDF HTML 收藏
2203.11197 2023-02-21 cs.LG cs.AI

Teachable Reinforcement Learning via Advice Distillation

Olivia Watkins, Trevor Darrell, Pieter Abbeel, Jacob Andreas, Abhishek Gupta

Comments Published at NeurIPS 2021

URL PDF HTML 收藏
2208.11821 2022-12-22 cs.CV

Refine and Represent: Region-to-Object Representation Learning

Akash Gokul, Konstantinos Kallidromitis, Shufan Li, Yusuke Kato, Kazuki Kozuka, Trevor Darrell, Colorado J Reed

URL PDF HTML 收藏
2211.11720 2022-12-06 cs.CV cs.CL

Multitask Vision-Language Prompt Tuning

Sheng Shen, Shijia Yang, Tianjun Zhang, Bohan Zhai, Joseph E. Gonzalez, Kurt Keutzer, Trevor Darrell

Comments Preprint

URL PDF HTML 收藏
2112.05744 2022-12-06 cs.CV cs.GR

More Control for Free! Image Synthesis with Semantic Diffusion Guidance

Xihui Liu, Dong Huk Park, Samaneh Azadi, Gong Zhang, Arman Chopikyan, Yuxiao Hu, Humphrey Shi, Anna Rohrbach, Trevor Darrell

Comments WACV 2023. Project page https://xh-liu.github.io/sdg/

URL PDF HTML 收藏
2112.10936 2022-12-05 cs.CV cs.AI cs.CL cs.CR cs.MM

Watch Those Words: Video Falsification Detection Using Word-Conditioned Facial Motion

Shruti Agarwal, Liwen Hu, Evonne Ng, Trevor Darrell, Hao Li, Anna Rohrbach

Comments Accepted in WACV 2023

URL PDF HTML 收藏
2206.06346 2022-11-30 cs.CV

Bringing Image Scene Structure to Video via Frame-Clip Consistency of Object Tokens

Elad Ben-Avraham, Roei Herzig, Karttikeya Mangalam, Amir Bar, Anna Rohrbach, Leonid Karlinsky, Trevor Darrell, Amir Globerson

Comments Tech report

URL PDF HTML 收藏
2211.15521 2022-11-29 cs.CV cs.CL

G^3: Geolocation via Guidebook Grounding

Grace Luo, Giscard Biamby, Trevor Darrell, Daniel Fried, Anna Rohrbach

Comments Findings of EMNLP 2022

URL PDF HTML 收藏