中文
相关论文

相关论文: Bag of Attributes for Video Event Retrieval

200 篇论文

Often, videos are composed of multiple concepts or even genres. For instance, news videos may contain sports, action, nature, etc. Therefore, encoding the distribution of such concepts/genres in a compact and effective representation is a…

计算机视觉与模式识别 · 计算机科学 2020-12-29 Leonardo A. Duarte , Otávio A. B. Penatti , Jurandy Almeida

This article gives a survey for bag-of-words (BoW) or bag-of-features model in image retrieval system. In recent years, large-scale image retrieval shows significant potential in both industry applications and research problems. As local…

信息检索 · 计算机科学 2013-04-19 Jialu Liu

This work proposes a simple instance retrieval pipeline based on encoding the convolutional features of CNN using the bag of words aggregation scheme (BoW). Assigning each local array of activations in a convolutional layer to a visual word…

计算机视觉与模式识别 · 计算机科学 2016-06-21 Eva Mohedano , Amaia Salvador , Kevin McGuinness , Ferran Marques , Noel E. O'Connor , Xavier Giro-i-Nieto

Visual learning problems such as object classification and action recognition are typically approached using extensions of the popular bag-of-words (BoW) model. Despite its great success, it is unclear what visual features the BoW model is…

计算机视觉与模式识别 · 计算机科学 2016-01-20 Ji Zhao , Liantao Wang , Ricardo Cabral , Fernando De la Torre

The past decade has seen the growing popularity of Bag of Features (BoF) approaches to many computer vision tasks, including image classification, video search, robot localization, and texture recognition. Part of the appeal is simplicity.…

计算机视觉与模式识别 · 计算机科学 2011-01-19 Stephen O'Hara , Bruce A. Draper

The use of bag of visual words (BOW) model for modelling images based on local invariant features computed at interest point locations has become a standard choice for many computer vision tasks. Visual vocabularies generated from image…

计算机视觉与模式识别 · 计算机科学 2022-10-18 Yousef Alqasrawi

A new class of applications based on visual search engines are emerging, especially on smart-phones that have evolved into powerful tools for processing images and videos. The state-of-the-art algorithms for large visual content recognition…

计算机视觉与模式识别 · 计算机科学 2016-04-15 Giuseppe Amato , Fabrizio Falchi , Claudio Gennaro

Most popular deep models for action recognition in videos generate independent predictions for short clips, which are then pooled heuristically to assign an action label to the full video segment. As not all frames may characterize the…

计算机视觉与模式识别 · 计算机科学 2019-09-09 Jue Wang , Anoop Cherian

This paper presents a new framework for visual bag-of-words (BOW) refinement and reduction to overcome the drawbacks associated with the visual BOW model which has been widely used for image classification. Although very influential in the…

计算机视觉与模式识别 · 计算机科学 2017-07-04 Zhiwu Lu , Liwei Wang , Ji-Rong Wen

In this paper we tackle the problem of efficient video event detection. We argue that linear detection functions should be preferred in this regard due to their scalability and efficiency during estimation and evaluation. A popular approach…

计算机视觉与模式识别 · 计算机科学 2015-09-07 Iman Abbasnejad , Sridha Sridharan , Simon Denman , Clinton Fookes , Simon Lucey

Explicit concept space models have proven efficacy for text representation in many natural language and text mining applications. The idea is to embed textual structures into a semantic space of concepts which captures the main ideas,…

计算与语言 · 计算机科学 2018-12-21 Walid Shalaby , Wlodek Zadrozny

Video based action recognition is one of the important and challenging problems in computer vision research. Bag of Visual Words model (BoVW) with local features has become the most popular method and obtained the state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2014-05-20 Xiaojiang Peng , Limin Wang , Xingxing Wang , Yu Qiao

The bag-of-words (BOW) model is the common approach for classifying documents, where words are used as feature for training a classifier. This generally involves a huge number of features. Some techniques, such as Latent Semantic Analysis…

计算与语言 · 计算机科学 2015-04-13 Rémi Lebret , Ronan Collobert

In this paper we present a novel approach for extracting a Bag-of-Words (BoW) representation based on a Neural Network codebook. The conventional BoW model is based on a dictionary (codebook) built from elementary representations which are…

音频与语音处理 · 电气工程与系统科学 2019-07-12 Mohammed Senoussaoui , Patrick Cardinal , Alessandro Lameiras Koerich

Image Classification based on BOW (Bag-of-words) has broad application prospect in pattern recognition field but the shortcomings are existed because of single feature and low classification accuracy. To this end we combine three…

计算机视觉与模式识别 · 计算机科学 2015-11-06 Huilin Gao , Wenjie Chen , Lihua Dou

The Bag--of--Visual--Words (BoVW) is a visual description technique that aims at shortening the semantic gap by partitioning a low--level feature space into regions of the feature space that potentially correspond to visual concepts and by…

计算机视觉与模式识别 · 计算机科学 2017-03-17 Antonio Foncubierta-Rodríguez , Henning Müller , Adrien Depeursinge

Medical Image Retrieval is a challenging field in Visual information retrieval, due to the multi-dimensional and multi-modal context of the underlying content. Traditional models often fail to take the intrinsic characteristics of data into…

计算机视觉与模式识别 · 计算机科学 2020-07-21 Sowmya Kamath S , Karthik K

Representing videos by densely extracted local space-time features has recently become a popular approach for analysing actions. In this paper, we tackle the problem of categorising human actions by devising Bag of Words (BoW) models based…

计算机视觉与模式识别 · 计算机科学 2016-07-08 Masoud Faraki , Maziar Palhang , Conrad Sanderson

Convolutional Neural Networks (CNNs) are well established models capable of achieving state-of-the-art classification accuracy for various computer vision tasks. However, they are becoming increasingly larger, using millions of parameters,…

计算机视觉与模式识别 · 计算机科学 2017-07-27 Nikolaos Passalis , Anastasios Tefas

While low-level image features have proven to be effective representations for visual recognition tasks such as object recognition and scene classification, they are inadequate to capture complex semantic meaning required to solve…

多媒体 · 计算机科学 2014-06-17 Tim Althoff , Hyun Oh Song , Trevor Darrell
‹ 上一页 1 2 3 10 下一页 ›