Skip to main navigation Skip to search Skip to main content

Learning robot activities from first-person human videos using convolutional future regression

  • Indiana University Bloomington

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

7 Scopus citations

Abstract

We design a new approach that allows robot learning of new activities from unlabeled human example videos. Given videos of humans executing the same activity from a human's viewpoint (i.e., first-person videos), our objective is to make the robot learn the temporal structure of the activity as its future regression network, and learn to transfer such model for its own motor execution. We present a new deep learning model: We extend the state-of-the-art convolutional object detection network for the representation/estimation of human hands in training videos, and newly introduce the concept of using a fully convolutional network to regress (i.e., predict) the intermediate scene representation corresponding to the future frame (e.g., 1-2 seconds later). Combining these allows direct prediction of future locations of human hands and objects, which enables the robot to infer the motor control plan using our manipulation network. We experimentally confirm that our approach makes learning of robot activities from unlabeled human interaction videos possible, and demonstrate that our robot is able to execute the learned collaborative activities in real-time directly based on its camera input.

Original languageEnglish
Title of host publicationIROS 2017 - IEEE/RSJ International Conference on Intelligent Robots and Systems
PublisherInstitute of Electrical and Electronics Engineers Inc.
Pages1497-1504
Number of pages8
ISBN (Electronic)9781538626825
DOIs
StatePublished - Dec 13 2017
Event2017 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2017 - Vancouver, Canada
Duration: Sep 24 2017Sep 28 2017

Publication series

NameIEEE International Conference on Intelligent Robots and Systems
Volume2017-September
ISSN (Print)2153-0858
ISSN (Electronic)2153-0866

Conference

Conference2017 IEEE/RSJ International Conference on Intelligent Robots and Systems, IROS 2017
Country/TerritoryCanada
CityVancouver
Period09/24/1709/28/17

Fingerprint

Dive into the research topics of 'Learning robot activities from first-person human videos using convolutional future regression'. Together they form a unique fingerprint.

Cite this