Are we ready for autonomous driving? The KITTI vision benchmark suite

Abstract

Today, visual recognition systems are still rarely employed in robotics applications. Perhaps one of the main reasons for this is the lack of demanding benchmarks that mimic such scenarios. In this paper, we take advantage of our autonomous driving platform to develop novel challenging benchmarks for the tasks of stereo, optical flow, visual odometry/SLAM and 3D object detection. Our recording platform is equipped with four high resolution video cameras, a Velodyne laser scanner and a state-of-the-art localization system. Our benchmarks comprise 389 stereo and optical flow image pairs, stereo visual odometry sequences of 39.2 km length, and more than 200k 3D object annotations captured in cluttered scenarios (up to 15 cars and 30 pedestrians are visible per image). Results from state-of-the-art algorithms reveal that methods ranking high on established datasets such as Middlebury perform below average when being moved outside the laboratory to the real world. Our goal is to reduce this bias by providing challenging benchmarks with novel difficulties to the computer vision community. Our benchmarks are available online at: www.cvlibs.net/datasets/kitti.

Keywords

Artificial intelligenceVisual odometryComputer scienceComputer visionSuiteBenchmark (surveying)OdometryRoboticsOptical flowObject detectionVisualizationRobotImage (mathematics)Mobile robotPattern recognition (psychology)

Affiliated Institutions

Related Publications

VINS-Mono: A Robust and Versatile Monocular Visual-Inertial State Estimator

Tong Qin , Peiliang Li , Shaojie Shen

One camera and one low-cost inertial measurement unit (IMU) form a monocular visual-inertial system (VINS), which is the minimum sensor suite (in size, weight, and power) for th...

2018 IEEE Transactions on Robotics 4020 citations

VoxelNet: End-to-End Learning for Point Cloud Based 3D Object Detection

Yin Zhou , Oncel Tuzel

Accurate detection of objects in 3D point clouds is a central problem in many applications, such as autonomous navigation, housekeeping robots, and augmented/virtual reality. To...

2018 4245 citations

Mobile Robot Localization and Mapping with Uncertainty using Scale-Invariant Visual Landmarks

Stephen Se , David Lowe , Jim Little

A key component of a mobile robot system is the ability to localize itself accurately and, simultaneously, to build a map of the environment. Most of the existing algorithms are...

2002 The International Journal of Robotics... 312 citations

HOGgles: Visualizing Object Detection Features

Carl Vondrick , Aditya Khosla , Tomasz Malisiewicz +1 more

We introduce algorithms to visualize feature spaces used by object detectors. The tools in this paper allow a human to put on 'HOG goggles' and perceive the visual world as a HO...

2013 284 citations

Caltech-256 Object Category Dataset

G. S. Griffin , Alex Holub , Pietro Perona

We introduce a challenging set of 256 object categories containing a total of 30607 images. The original Caltech-101 [1] was collected by choosing a set of object categories, do...

2007 The Caltech Institute Archives (Calif... 2388 citations

Publication Info

Year: 2012
Type: article
Pages: 3354-3361
Citations: 13591
Access: Closed

External Links

View on DOI.org

Social Impact

Altmetric

Are we ready for autonomous driving? The KITTI vision benchmark suite

PlumX Metrics

Social media, news, blog, policy document mentions

Citation Metrics

13591

OpenAlex

Cite This

APA Style

                            
                                    Andreas Geiger, 
                                
                                    P Lenz, 
                                
                                    R. Urtasun
                                
                            (2012). 
                            Are we ready for autonomous driving? The KITTI vision benchmark suite. 
                            
                            , 3354-3361.
                            https://doi.org/10.1109/cvpr.2012.6248074

Identifiers

DOI: 10.1109/cvpr.2012.6248074