中文
相关论文

相关论文: Log-Euclidean Bag of Words for Human Action Recogn…

200 篇论文

Human action recognition remains a challenging task due to the various sources of video data and large intra-class variations. It thus becomes one of the key issues in recent research to explore effective and robust representation to handle…

计算机视觉与模式识别 · 计算机科学 2015-11-17 Mengyi Liu , Ruiping Wang , Shiguang Shan , Xilin Chen

We present data-driven techniques to augment Bag of Words (BoW) models, which allow for more robust modeling and recognition of complex long-term activities, especially when the structure and topology of the activities are not known a…

计算机视觉与模式识别 · 计算机科学 2016-11-15 Vinay Bettadapura , Grant Schindler , Thomaz Plotz , Irfan Essa

The human action classification task is a widely researched topic and is still an open problem. Many state-of-the-arts approaches involve the usage of bag-of-video-words with spatio-temporal local features to construct characterizations for…

计算机视觉与模式识别 · 计算机科学 2016-10-18 Aznul Qalid Md Sabri , Jacques Boonaert , Erma Rahayu Mohd Faizal Abdullah , Ali Mohammed Mansoor

This paper proposes a semantic segmentation method for outdoor scenes captured by a surveillance camera. Our algorithm classifies each perceptually homogenous region as one of the predefined classes learned from a collection of manually…

计算机视觉与模式识别 · 计算机科学 2013-05-15 Wassim Bouachir , Atousa Torabi , Guillaume-Alexandre Bilodeau , Pascal Blais

Statistical classification of actions in videos is mostly performed by extracting relevant features, particularly covariance features, from image frames and studying time series associated with temporal evolutions of these features. A…

计算机视觉与模式识别 · 计算机科学 2015-04-13 Zhengwu Zhang , Jingyong Su , Eric Klassen , Huiling Le , Anuj Srivastava

We consider a family of structural descriptors for visual data, namely covariance descriptors (CovDs) that lie on a non-linear symmetric positive definite (SPD) manifold, a special type of Riemannian manifolds. We propose an improved…

计算机视觉与模式识别 · 计算机科学 2019-09-27 Kai-Xuan Chen , Xiao-Jun Wu , Jie-Yi Ren , Rui Wang , Josef Kittler

Video based action recognition is one of the important and challenging problems in computer vision research. Bag of Visual Words model (BoVW) with local features has become the most popular method and obtained the state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2014-05-20 Xiaojiang Peng , Limin Wang , Xingxing Wang , Yu Qiao

This work aims to present novel description methods for human action recognition. Generally, a video sequence can be represented as a collection of spatial temporal words by detecting space-time interest points and describing the unique…

人机交互 · 计算机科学 2011-01-04 Ruoyun Gao , Michael S. Lew , Ling Shao

We propose a new action and gesture recognition method based on spatio-temporal covariance descriptors and a weighted Riemannian locality preserving projection approach that takes into account the curved space formed by the descriptors. The…

计算机视觉与模式识别 · 计算机科学 2013-03-26 Andres Sanin , Conrad Sanderson , Mehrtash T. Harandi , Brian C. Lovell

Person re-identification is generally divided into two part: first how to represent a pedestrian by discriminative visual descriptors and second how to compare them by suitable distance metrics. Conventional methods isolate these two parts,…

计算机视觉与模式识别 · 计算机科学 2017-04-12 Lu Tian , Shengjin Wang

3D action recognition has broad applications in human-computer interaction and intelligent surveillance. However, recognizing similar actions remains challenging since previous literature fails to capture motion and shape cues effectively…

计算机视觉与模式识别 · 计算机科学 2017-12-08 Mengyuan Liu , Hong Liu , Chen Chen

The Bag--of--Visual--Words (BoVW) is a visual description technique that aims at shortening the semantic gap by partitioning a low--level feature space into regions of the feature space that potentially correspond to visual concepts and by…

计算机视觉与模式识别 · 计算机科学 2017-03-17 Antonio Foncubierta-Rodríguez , Henning Müller , Adrien Depeursinge

As research on action recognition matures, the focus is shifting away from categorizing basic task-oriented actions using hand-segmented video datasets to understanding complex goal-oriented daily human activities in real-world settings.…

计算机视觉与模式识别 · 计算机科学 2016-03-18 Hilde Kuehne , Juergen Gall , Thomas Serre

Representing texts as fixed-length vectors is central to many language processing tasks. Most traditional methods build text representations based on the simple Bag-of-Words (BoW) representation, which loses the rich semantic relations…

计算与语言 · 计算机科学 2017-07-19 Ruqing Zhang , Jiafeng Guo , Yanyan Lan , Jun Xu , Xueqi Cheng

Detection and classification of ships based on their silhouette profiles in natural imagery is an important undertaking in computer science. This problem can be viewed from a variety of perspectives, including security, traffic control, and…

计算机视觉与模式识别 · 计算机科学 2021-02-24 Sadegh Soleimani Pour , Ata Jodeiri , Hossein Rashidi , Seyed Mostafa Mirhassani , Hoda Kheradfallah , Hadi Seyedarabi

This article gives a survey for bag-of-words (BoW) or bag-of-features model in image retrieval system. In recent years, large-scale image retrieval shows significant potential in both industry applications and research problems. As local…

信息检索 · 计算机科学 2013-04-19 Jialu Liu

In this work\footnote {This work was supported in part by the National Science Foundation under grant IIS-1212948.}, we present a method to represent a video with a sequence of words, and learn the temporal sequencing of such words as the…

计算机视觉与模式识别 · 计算机科学 2019-06-18 Sangwoo Cho , Hassan Foroosh

Identifying human behaviors is a challenging research problem due to the complexity and variation of appearances and postures, the variation of camera settings, and view angles. In this paper, we try to address the problem of human behavior…

计算机视觉与模式识别 · 计算机科学 2019-03-08 Eissa Jaber Alreshidi , Mohammad Bilal

Recent work has explored methods for learning continuous vector space word representations reflecting the underlying semantics of words. Simple vector space arithmetic using cosine distances has been shown to capture certain types of…

计算与语言 · 计算机科学 2015-07-29 Sridhar Mahadevan , Sarath Chandar

In this paper, we present the Bag-of-Attributes (BoA) model for video representation aiming at video event retrieval. The BoA model is based on a semantic feature space for representing videos, resulting in high-level video feature vectors.…

信息检索 · 计算机科学 2020-12-29 Leonardo A. Duarte , Otávio A. B. Penatti , Jurandy Almeida
‹ 上一页 1 2 3 10 下一页 ›