Predicting Object Dynamics in Scenes

David F. Fouhey; C. L. Zitnick

2014 CVPR CVPR 2014

Predicting Object Dynamics in Scenes

Abstract

Given a static scene, a human can trivially enumerate the myriad of things that can happen next and characterize the relative likelihood of each. In the process, we make use of enormous amounts of commonsense knowledge about how the world works. In this paper, we investigate learning this commonsense knowledge from data. To overcome a lack of densely annotated spatiotemporal data, we learn from sequences of abstract images gathered using crowdsourcing. The abstract scenes provide both object location and attribute information. We demonstrate qualitatively and quantitatively that our models produce plausible scene predictions on both the abstract images, as well as natural images taken from the Internet.

🌉 Interdisciplinary Bridge — Computer Vision and Interdisciplinary

🧭 Keyword Pioneer — commonsense knowledge

🐣 Hot Topic Early Bird — commonsense knowledge

🐝 Cross-Pollinator — Artificial Intelligence, Computer Science, Computer Vision, Data Science & Analytics, Deep Learning, Interdisciplinary, Knowledge & Reasoning, Machine Learning, Mathematics & Optimization, Natural Language Processing, Reinforcement Learning, Robotics

Authors

David F. Fouhey , C. L. Zitnick

Topics

Computer Vision > Analysis > Scene Understanding Interdisciplinary > Cognitive Science > Cognitive Modeling

Keywords

commonsense knowledge spatiotemporal datum scene prediction object dynamics abstract image

Download PDF

Related papers

Efficient Nonlinear Markov Models for Human Motion 2014

Occlusion Geodesics for Online Multi-Object Tracking 2014

A Principled Approach for Coarse-to-Fine MAP Inference 2014

Locally Optimized Product Quantization for Approximate Nearest Neighbor Search 2014

Fast and Accurate Image Matching with Cascade Hashing for 3D Reconstruction 2014