作者
Fei-Fei Li
Computer Vision
OccFusion: Rendering Occluded Humans with Generative Diffusion Priors
An Interactive Agent Foundation Model
BEHAVIOR Vision Suite: Customizable Dataset Generation via Simulation
Comments CVPR 2024 (Highlight). Project website: https://behavior-vision-suite.github.io/
ZeroNVS: Zero-Shot 360-Degree View Synthesis from a Single Image
Comments Accepted to CVPR 2024. 12 pages
BEHAVIOR-1K: A Human-Centered, Embodied AI Benchmark with 1,000 Everyday Activities and Realistic Simulation
Comments A preliminary version was published at 6th Conference on Robot Learning (CoRL 2022)
Position Paper: Agent AI Towards a Holistic Intelligence
Comments 22 pages, 4 figures. arXiv admin note: substantial text overlap with arXiv:2401.03568
Agent AI: Surveying the Horizons of Multimodal Interaction
Mini-BEHAVIOR: A Procedurally Generated Benchmark for Long-horizon Decision-Making in Embodied AI
Model-Based Control with Sparse Neural Dynamics
Comments Accepted at NeurIPS 2023. For tutorial code and additional visualizations, see https://robopil.github.io/Sparse-Dynamics/
Photorealistic Video Generation with Diffusion Models
Comments Project website https://walt-video-diffusion.github.io/
Holistic Evaluation of Text-To-Image Models
Comments NeurIPS 2023. First three authors contributed equally
NOIR: Neural Signal Operated Intelligent Robots for Everyday Activities
VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models
Sequential Dexterity: Chaining Dexterous Policies for Long-Horizon Manipulation
Comments 7th Conference on Robot Learning (CoRL 2023)
MimicPlay: Long-Horizon Imitation Learning by Watching Human Play
Comments 7th Conference on Robot Learning (CoRL 2023 oral presentation)
Co-GAIL: Learning Diverse Strategies for Human-Robot Collaboration
Comments CoRL 2021
MindAgent: Emergent Gaming Interaction
Comments The first three authors contributed equally. 28 pages
Sonicverse: A Multisensory Simulation Platform for Embodied Household Agents that See and Hear
Comments In ICRA 2023. Project page: https://ai.stanford.edu/~rhgao/sonicverse/. Code: https://github.com/StanfordVL/sonicverse. Gao and Li contributed equally to this work and are in alphabetical order
Rendering Humans from Object-Occluded Monocular Videos
Comments ICCV 2023, project page: https://cs.stanford.edu/~xtiange/projects/occnerf/
Primitive Skill-based Robot Learning from Human Evaluative Feedback
Dynamic-Resolution Model Learning for Object Pile Manipulation
Comments Accepted to Robotics: Science and Systems (RSS) 2023. The first two authors contributed equally. Project Page: https://robopil.github.io/dyn-res-pile-manip
Differentially Private Video Activity Recognition
Task-Driven Graph Attention for Hierarchical Relational Object Navigation
Modeling Dynamic Environments with Scene Graph Memory
HomE: Homography-Equivariant Video Representation Learning
Comments 10 pages, 4 figures, 4 tables
The ObjectFolder Benchmark: Multisensory Learning with Neural and Real Objects
Comments In CVPR 2023. Project page: https://objectfolder.stanford.edu/. ObjectFolder Real demo: https://www.objectfolder.org/swan_vis/. Gao, Dou, and Li contributed equally to this work
VIMA: General Robot Manipulation with Multimodal Prompts
Comments ICML 2023 Camera-ready version. Project website: https://vimalabs.github.io/
Siamese Masked Autoencoders
Comments Project page https://siam-mae-video.github.io/