Abstract
We address the task of recognizing objects from video input. This important problem is relatively unexplored, compared with image-based object recognition. To this end, we make the following contributions. First, we introduce two comprehensive data sets for video-based object recognition. Second, we propose latent bi-constraint SVM (LBSVM), a maximum-margin framework for video-based object recognition. LBSVM is based on structured-output SVM, but extends it to handle noisy video data and ensure consistency of the output decision throughout time. We apply LBSVM to recognize office objects and museum sculptures, and we demonstrate its benefits over image-based, set-based, and other video-based object recognition.
| Original language | English |
|---|---|
| Article number | 7944564 |
| Pages (from-to) | 3044-3052 |
| Number of pages | 9 |
| Journal | IEEE Transactions on Circuits and Systems for Video Technology |
| Volume | 28 |
| Issue number | 10 |
| DOIs | |
| State | Published - Oct 2018 |
Keywords
- Object recognition
- structured-output SVM
- video analysis
Fingerprint
Dive into the research topics of 'Latent Bi-Constraint SVM for Video-Based Object Recognition'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver