Retrieving objects from videos based on affine regions

V. Ferrari, T. Tuytelaars, L. Van Gool

Research output: Chapter in Book/Report/Conference proceedingConference contribution


We present a method to (semi-)automatically annotate video material. More precisely, we focus on recognizing specific objects and scenes in keyframes. Objects are learnt simply by having the user delineate them in one (or a few) images. The basic building block to achieve this goal consists of affine invariant regions. These are local image patches that adapt their shape based on the image content so as to be invariant to viewpoint changes. Instead of simply matching the regions and counting the number of matches, we propose to gather more evidence about the presence of the object by exploring the image around the initial matches. This boosts the performance, especially under difficult, real-world imaging conditions. Experimental results on news broadcast data demonstrate the viability of the approach.
Original languageEnglish
Title of host publicationSignal Processing Conference, 2004 12th European
PublisherInstitute of Electrical and Electronics Engineers (IEEE)
Number of pages4
ISBN (Print)978-320-0001-65-7
Publication statusPublished - Sep 2004


Dive into the research topics of 'Retrieving objects from videos based on affine regions'. Together they form a unique fingerprint.

Cite this