作者
Fei-Fei Li
Computer Vision
SURREAL-System: Fully-Integrated Stack for Distributed Deep Reinforcement Learning
Comments Technical report of the SURREAL system. See more details at https://surreal.stanford.edu
Causal Induction from Visual Observations for Goal Directed Tasks
Comments 13 pages, 6 figures
Regression Planning Networks
Comments Accepted at NeurIPS 2019
Making Sense of Vision and Touch: Learning Multimodal Representations for Contact-Rich Tasks
Comments arXiv admin note: substantial text overlap with arXiv:1810.10191
Peeking into the Future: Predicting Future Person Activities and Locations in Videos
Comments In CVPR 2019. Code, models and more results are available at: https://next.cs.cmu.edu/
D3TW: Discriminative Differentiable Dynamic Time Warping for Weakly Supervised Action Alignment and Segmentation
Comments To appear in CVPR 2019
Auto-DeepLab: Hierarchical Neural Architecture Search for Semantic Image Segmentation
Comments To appear in CVPR 2019 as oral. Code for Auto-DeepLab released at https://github.com/tensorflow/models/tree/master/research/deeplab
Information Maximizing Visual Question Generation
Comments CVPR 2019
Journal ref IEEE Conference on Computer Vision and Pattern Recognition, 2019
Scene Memory Transformer for Embodied Agents in Long-Horizon Tasks
Comments CVPR 2019 paper with supplementary material
Making Sense of Vision and Touch: Self-Supervised Learning of Multimodal Representations for Contact-Rich Tasks
Comments ICRA 2019
Neural Task Graphs: Generalizing to Unseen Tasks from a Single Video Demonstration
Comments CVPR 2019
Audio-Linguistic Embeddings for Spoken Sentences
Comments International Conference on Acoustics, Speech, and Signal Processing (ICASSP) 2019
DenseFusion: 6D Object Pose Estimation by Iterative Dense Fusion
Composing Text and Image for Image Retrieval - An Empirical Odyssey
Vision-Based Gait Analysis for Senior Care
Comments Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216
Privacy-Preserving Action Recognition for Smart Hospitals using Low-Resolution Depth Images
Comments Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216
Measuring Depression Symptom Severity from Spoken Language and 3D Facial Expressions
Comments Machine Learning for Health (ML4H) Workshop at NeurIPS 2018 arXiv:1811.07216
Faster CryptoNets: Leveraging Sparsity for Real-World Encrypted Inference
A Fully Private Pipeline for Deep Learning on Electronic Health Records
RoboTurk: A Crowdsourcing Platform for Robotic Skill Learning through Imitation
Comments Published at the Conference on Robot Learning (CoRL) 2018
Learning to Play with Intrinsically-Motivated Self-Aware Agents
Comments In NIPS 2018. 10 pages, 5 figures
Flexible Neural Representation for Physics Prediction
Comments 23 pages, 20 figures
Learning to Decompose and Disentangle Representations for Video Prediction
MentorNet: Learning Data-Driven Curriculum for Very Deep Neural Networks on Corrupted Labels
Journal ref published at ICML 2018
Graph Distillation for Action Detection with Privileged Modalities
Comments ECCV 2018
Progressive Neural Architecture Search
Comments To appear in ECCV 2018 as oral. The code and checkpoint for PNASNet-5 trained on ImageNet (both Mobile and Large) can now be downloaded from https://github.com/tensorflow/models/tree/master/research/slim#Pretrained. Also see https://github.com/chenxi116/PNASNet.TF for refactored and simplified TensorFlow code; see https://github.com/chenxi116/PNASNet.pytorch for exact conversion to PyTorch
HiDDeN: Hiding Data With Deep Networks
Tool Detection and Operative Skill Assessment in Surgical Videos Using Region-Based Convolutional Neural Networks
Comments arXiv admin note: text overlap with arXiv:1806.02031 by other authors
Topological light-trapping on a dislocation
Journal ref Nature Communications 9, 2462 (2018)