Robot Learning Manipulation Action Plans by "Watching" Unconstrained Videos from the World Wide Web-Reference-Cited by-同舟云学术

Robot Learning Manipulation Action Plans by "Watching" Unconstrained Videos from the World Wide Web

Published:2015-03-04 Issue:1 Volume:29 Page:
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Yang Yezhou,Li Yi,Fermuller Cornelia,Aloimonos Yiannis

Abstract

In order to advance action generation and creation in robots beyond simple learned schemas we need computational tools that allow us to automatically interpret and represent human actions. This paper presents a system that learns manipulation action plans by processing unconstrained videos from the World Wide Web. Its goal is to robustly generate the sequence of atomic actions of seen longer actions in video in order to acquire knowledge for robots. The lower level of the system consists of two convolutional neural network (CNN) based recognition modules, one for classifying the hand grasp type and the other for object recognition. The higher level is a probabilistic manipulation action grammar based parsing module that aims at generating visual sentences for robot manipulation. Experiments conducted on a publicly available unconstrained video dataset show that the system is able to learn manipulation actions by ``watching'' unconstrained videos with high accuracy.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 27 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. A Review of Natural-Language-Instructed Robot Execution Systems;AI;2024-06-26

2. Cook2LTL: Translating Cooking Recipes to LTL Formulae using Large Language Models;2024 IEEE International Conference on Robotics and Automation (ICRA);2024-05-13

3. TiV-ODE: A Neural ODE-based Approach for Controllable Video Generation From Text-Image Pairs;2024 IEEE International Conference on Robotics and Automation (ICRA);2024-05-13

4. Self-Supervised Bayesian Visual Imitation Learning Applied to Robotic Pouring;2024 IEEE International Conference on Industrial Technology (ICIT);2024-03-25

5. Imitation Learning of Long-Horizon Manipulation Tasks Through Temporal Sub-action Sequencing;Communications in Computer and Information Science;2024