English
Related papers

Related papers: Spectral Scalpel: Amplifying Adjacent Action Discr…

200 papers

Target speech separation refers to extracting the target speaker's speech from mixed signals. Despite the recent advances in deep learning based close-talk speech separation, the applications to real-world are still an open issue. Two main…

Sound · Computer Science 2020-01-03 Rongzhi Gu , Yuexian Zou

Hyperspectral image super-resolution has attained widespread prominence to enhance the spatial resolution of hyperspectral images. However, convolution-based methods have encountered challenges in harnessing the global spatial-spectral…

Image and Video Processing · Electrical Eng. & Systems 2023-11-30 Shi Chen , Lefei Zhang , Liangpei Zhang

Skeleton-based action recognition faces two longstanding challenges: the scarcity of labeled training samples and difficulty modeling short- and long-range temporal dependencies. To address these issues, we propose a unified framework,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-19 Feng Ding , Haisheng Fu , Soroush Oraki , Jie Liang

Graph Convolutional Networks (GCNs) have been widely used to model the high-order dynamic dependencies for skeleton-based action recognition. Most existing approaches do not explicitly embed the high-order spatio-temporal importance to…

Computer Vision and Pattern Recognition · Computer Science 2022-02-07 Lipeng Ke , Kuan-Chuan Peng , Siwei Lyu

Variations of human body skeletons may be considered as dynamic graphs, which are generic data representation for numerous real-world applications. In this paper, we propose a spatio-temporal graph convolution (STGC) approach for assembling…

Computer Vision and Pattern Recognition · Computer Science 2018-02-28 Chaolong Li , Zhen Cui , Wenming Zheng , Chunyan Xu , Jian Yang

In real-world scenarios, human actions often fall into a long-tailed distribution. It makes the existing skeleton-based action recognition works, which are mostly designed based on balanced datasets, suffer from a sharp performance…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Jiahang Zhang , Lilang Lin , Jiaying Liu

Skeleton-based human action recognition aims to classify human skeletal sequences, which are spatiotemporal representations of actions, into predefined categories. To reduce the reliance on costly annotations of skeletal sequences while…

Computer Vision and Pattern Recognition · Computer Science 2025-10-30 Zhigang Tu , Zhengbo Zhang , Jia Gong , Junsong Yuan , Bo Du

The foundation model has recently garnered significant attention due to its potential to revolutionize the field of visual representation learning in a self-supervised manner. While most foundation models are tailored to effectively process…

Computer Vision and Pattern Recognition · Computer Science 2024-02-13 Danfeng Hong , Bing Zhang , Xuyang Li , Yuxuan Li , Chenyu Li , Jing Yao , Naoto Yokoya , Hao Li , Pedram Ghamisi , Xiuping Jia , Antonio Plaza , Paolo Gamba , Jon Atli Benediktsson , Jocelyn Chanussot

Spiking Neural Networks (SNNs) are considered naturally suited for temporal processing, with membrane potential propagation widely regarded as the core temporal modeling mechanism. However, existing research lack analysis of its actual…

Neural and Evolutionary Computing · Computer Science 2025-12-08 Yiting Dong , Zhaofei Yu , Jianhao Ding , Zijie Xu , Tiejun Huang

Handling anomalies is a critical preprocessing step in multivariate time series prediction. However, existing approaches that separate anomaly preprocessing from model training for multivariate time series prediction encounter significant…

Machine Learning · Computer Science 2025-01-15 Yuanyuan Liang , Tianhao Zhang , Tingyu Xie

Inferring future activity information based on observed activity data is a crucial step to improve the accuracy of early activity prediction. Traditional methods based on generative adversarial networks(GAN) or joint learning frameworks can…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Tingyu Liu , Jun Huang , Chenyi Weng

Capitalizing on image-level pre-trained models for various downstream tasks has recently emerged with promising performance. However, the paradigm of "image pre-training followed by video fine-tuning" for high-dimensional video data…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Shu Yang , Zhiyuan Cai , Luyang Luo , Ning Ma , Shuchang Xu , Hao Chen

Spectral clustering requires the time-consuming decomposition of the Laplacian matrix of the similarity graph, thus limiting its applicability to large datasets. To improve the efficiency of spectral clustering, a top-down approach was…

Machine Learning · Computer Science 2024-12-19 Zhichang Xu , Zhiguo Long , Hua Meng

Adaptive stretching, where the post compression signal is iteratively stretched to maximize the correlation between the pre and post compression rf echo frames, has demonstrated superior performance compared to gradient based methods. At…

Tissues and Organs · Quantitative Biology 2022-10-25 Shaiban Ahmed , Rasheed Abid , S. Kaisar Alam

Forecasting epileptic seizures from multivariate EEG signals represents a critical challenge in healthcare time series prediction, requiring high sensitivity, low false alarm rates, and subject-specific adaptability. We present STAN, an…

Machine Learning · Computer Science 2026-03-04 Zan Li , Kyongmin Yeo , Wesley Gifford , Lara Marcuse , Madeline Fields , Bülent Yener

Scene Text Recognition (STR), the task of recognizing text against complex image backgrounds, is an active area of research. Current state-of-the-art (SOTA) methods still struggle to recognize text written in arbitrary shapes. In this…

Computer Vision and Pattern Recognition · Computer Science 2020-03-26 Ron Litman , Oron Anschel , Shahar Tsiper , Roee Litman , Shai Mazor , R. Manmatha

We propose an ultrasound speckle filtering method for not only preserving various edge features but also filtering tissue-dependent complex speckle noises in ultrasound images. The key idea is to detect these various edges using a phase…

Image and Video Processing · Electrical Eng. & Systems 2021-02-11 Kunqiang Mei , Bin Hu , Baowei Fei , Binjie Qin

Streaming recognition and segmentation of multi-party conversations with overlapping speech is crucial for the next generation of voice assistant applications. In this work we address its challenges discovered in the previous work on…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-12 Ilya Sklyar , Anna Piunova , Christian Osendorfer

Spatiotemporal dynamics is central to a wide range of applications from climatology, computer vision to neural sciences. From temporal observations taken on a high-dimensional vector of spatial locations, we seek to derive knowledge about…

Methodology · Statistics 2016-04-19 Lu Meng , Tian Zheng

This paper presents a novel soft tactile skin (STS) technology operating with sound waves. In this innovative approach, the sound waves generated by a speaker travel in channels embedded in a soft membrane and get modulated due to a…

Robotics · Computer Science 2024-03-01 Vishnu Rajendran S , Willow Mandil , Simon Parsons , Amir Ghalamzan E
‹ Prev 1 4 5 6 7 8 10 Next ›