作者
Trevor Darrell
Computer Vision
Dynamic Scale Inference by Entropy Minimization
Disentangling Propagation and Generation for Video Prediction
Comments ICCV 2019
Are You Looking? Grounding to Multiple Modalities in Vision-and-Language Navigation
Comments ACL 2019
Reinforcement Learning from Imperfect Demonstrations
Monocular Plan View Networks for Autonomous Driving
Comments 8 pages, 9 figures
Accurate Visual Localization for Automotive Applications
Blurring the Line Between Structure and Learning to Optimize and Adapt Receptive Fields
Robust Change Captioning
Hierarchical Discrete Distribution Decomposition for Match Density Estimation
Comments To appear at CVPR 2019
Adversarial Inference for Multi-Sentence Video Description
Comments Accepted to Computer Vision and Pattern Recognition (CVPR) 2019
TAFE-Net: Task-Aware Feature Embeddings for Low Shot Learning
Comments Accepted at CVPR 2019
Deep Mixture of Experts via Shallow Embedding
Generalized Zero- and Few-Shot Learning via Aligned Variational Autoencoders
Comments Accepted at CVPR 2019
Object Hallucination in Image Captioning
Comments Rohrbach and Hendricks contributed equally; accepted to EMNLP 2018
Compositional GAN: Learning Image-Conditional Binary Composition
Women also Snowboard: Overcoming Bias in Captioning Models
Comments 22 pages, 6 figures, Burns and Hendricks contributed equally
Explainable Neural Computation via Stack Neural Module Networks
Comments ECCV 2018
Rethinking the Value of Network Pruning
Comments ICLR 2019. Significant revisions from the previous version
Deep Object-Centric Policies for Autonomous Driving
Comments Accepted at ICRA 2019
Discriminator Rejection Sampling
Comments Published as a conference paper at ICLR 2019
Deep Layer Aggregation
Comments Published at the Conference on Computer Vision and Pattern Recognition (CVPR) 2018
Similarity R-C3D for Few-shot Temporal Activity Detection
SPLAT: Semantic Pixel-Level Adaptation Transforms for Detection
Modular Architecture for StarCraft II with Deep Reinforcement Learning
Comments Accepted to The 14th AAAI Conference on Artificial Intelligence and Interactive Digital Entertainment (AIIDE'18)
Speaker-Follower Models for Vision-and-Language Navigation
Comments NIPS 2018
Toward Multimodal Image-to-Image Translation
Comments NIPS 2017 Final paper. v4 updated acknowledgment. Website: https://junyanz.github.io/BicycleGAN/
Localizing Moments in Video with Temporal Language
Comments EMNLP 2018
Large-Scale Study of Curiosity-Driven Learning
Comments First three authors contributed equally and ordered alphabetically. Website at https://pathak22.github.io/large-scale-curiosity/
Grounding Visual Explanations
Comments Accepted to ECCV 2018
Journal ref European Conference on Computer Vision (ECCV), 2018