Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Bag of Attributes for Video Event Retrieval

Jul 18, 2016

Leonardo A. Duarte, Otávio A. B. Penatti, Jurandy Almeida

Figure 1 for Bag of Attributes for Video Event Retrieval

Figure 2 for Bag of Attributes for Video Event Retrieval

Figure 3 for Bag of Attributes for Video Event Retrieval

Figure 4 for Bag of Attributes for Video Event Retrieval

Share this with someone who'll enjoy it:

Abstract:In this paper, we present the Bag-of-Attributes (BoA) model for video representation aiming at video event retrieval. The BoA model is based on a semantic feature space for representing videos, resulting in high-level video feature vectors. For creating a semantic space, i.e., the attribute space, we can train a classifier using a labeled image dataset, obtaining a classification model that can be understood as a high-level codebook. This model is used to map low-level frame vectors into high-level vectors (e.g., classifier probability scores). Then, we apply pooling operations on the frame vectors to create the final bag of attributes for the video. In the BoA representation, each dimension corresponds to one category (or attribute) of the semantic space. Other interesting properties are: compactness, flexibility regarding the classifier, and ability to encode multiple semantic concepts in a single video representation. Our experiments considered the semantic space created by a deep convolutional neural network (OverFeat) pre-trained on 1000 object categories of ImageNet. OverFeat was then used to classify each video frame and max pooling combined the frame vectors in the BoA representation for the video. Results using BoA outperformed the baselines with statistical significance in the task of video event retrieval using the EVVE dataset.

View paper on

Share this with someone who'll enjoy it:

Title:Bag of Attributes for Video Event Retrieval

Paper and Code