Abstract
This paper presents a novel approach to detect and track people and cars based on the combined information retrieved from a camera and a laser range scanner. Laser data points are classified by using boosted Conditional Random Fields, while the image based detector uses an extension of the Implicit Shape Model (ISM), which learns a codebook of local descriptors from a set of hand-labeled images and uses them to vote for centers of detected objects. Our extensions to ISM include the learning of object parts and template masks to obtain more distinctive votes for the particular object classes. The detections from both sensors are then fused and the objects are tracked using a Kalman Filter with multiple motion models. Experiments conducted in real-world urban scenarios demonstrate the effectiveness of our approach.
Originalsprache | Englisch |
---|---|
Seiten (von - bis) | 1498-1515 |
Seitenumfang | 18 |
Fachzeitschrift | International Journal of Robotics Research |
Jahrgang | 29 |
Ausgabenummer | 12 |
DOIs | |
Publikationsstatus | Veröffentlicht - Okt. 2010 |
Extern publiziert | Ja |