TY - GEN
T1 - Video Retrieval by Mimicking Poses
AU - Jammalamadaka, Nataraj
AU - Zisserman, Andrew
AU - Eichner, Marcin
AU - Ferrari, Vittorio
AU - Jawahar, C. V.
PY - 2012
Y1 - 2012
N2 - We describe a method for real time video retrieval where the task is to match the 2D human pose of a query. A user can form a query by (i) interactively controlling a stickman on a web based GUI, (ii) uploading an image of the desired pose, or (iii) using the Kinect and acting out the query himself. The method is scalable and is applied to a dataset of 18 films totaling more than three million frames. The real time performance is achieved by searching for approximate nearest neighbors to the query using a random forest of K-D trees. Apart from the query modalities, we introduce two other areas of novelty. First, we show that pose retrieval can proceed using a low dimensional representation. Second, we show that the precision of the results can be improved substantially by combining the outputs of independent human pose estimation algorithms. The performance of the system is assessed quantitatively over a range of pose queries
AB - We describe a method for real time video retrieval where the task is to match the 2D human pose of a query. A user can form a query by (i) interactively controlling a stickman on a web based GUI, (ii) uploading an image of the desired pose, or (iii) using the Kinect and acting out the query himself. The method is scalable and is applied to a dataset of 18 films totaling more than three million frames. The real time performance is achieved by searching for approximate nearest neighbors to the query using a random forest of K-D trees. Apart from the query modalities, we introduce two other areas of novelty. First, we show that pose retrieval can proceed using a low dimensional representation. Second, we show that the precision of the results can be improved substantially by combining the outputs of independent human pose estimation algorithms. The performance of the system is assessed quantitatively over a range of pose queries
KW - human pose search, video processing
U2 - 10.1145/2324796.2324838
DO - 10.1145/2324796.2324838
M3 - Conference contribution
T3 - ICMR '12
SP - 34:1-34:8
BT - Proceedings of the 2Nd ACM International Conference on Multimedia Retrieval
PB - ACM
CY - New York, NY, USA
ER -