A Dataset for Developing and Benchmarking Active Vision

Phil Ammirato,Patrick Poirson,Eunbyung Park,Jana Kosecka,Alexander C. Berg

2017 IEEE International Conference on Robotics and Automation (ICRA)（2017）

引用 192|浏览123

暂无评分

摘要

We present a new public dataset with a focus on simulating robotic vision tasks in everyday indoor environments using real imagery. The dataset includes 20,000+ RGB-D images and 50,000+ 2D bounding boxes of object instances densely captured in 9 unique scenes. We train a fast object category detector for instance detection on our data. Using the dataset we show that, although increasingly accurate and fast, the state of the art for object detection is still severely impacted by object scale, occlusion, and viewing direction all of which matter for robotics applications. We next validate the dataset for simulating active vision, and use the dataset to develop and evaluate a deep-network-based system for next best move prediction for object classification using reinforcement learning. Our dataset is available for download at cs.unc.edu/~ammirato/active_vision_dataset_website/.

查看译文

关键词

active vision,robotic vision,indoor environments,RGB-D images,2D bounding boxes,object category detector,instance detection,object detection,deep-network-based system,object classification,reinforcement learning

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要