English
Related papers

Related papers: ASC-SW: Atrous strip convolution network with slid…

200 papers

In image denoising networks, feature scaling is widely used to enlarge the receptive field size and reduce computational costs. This practice, however, also leads to the loss of high-frequency information and fails to consider within-scale…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Hao Shen , Zhong-Qiu Zhao , Wandi Zhang

Background and objective: In this paper, a modified U-Net based framework is presented, which leverages techniques from Squeeze-and-Excitation (SE) block, Atrous Spatial Pyramid Pooling (ASPP) and residual learning for accurate and robust…

Image and Video Processing · Electrical Eng. & Systems 2021-07-20 Jinke Wang , Peiqing Lv , Haiying Wang , Changfa Shi

Dynamic convolution achieves better performance for efficient CNNs at the cost of negligible FLOPs increase. However, the performance increase can not match the significantly expanded number of parameters, which is the main bottleneck in…

Computer Vision and Pattern Recognition · Computer Science 2023-05-29 Shwai He , Chenbo Jiang , Daize Dong , Liang Ding

The development of lightweight object detectors is essential due to the limited computation resources. To reduce the computation cost, how to generate redundant features plays a significant role. This paper proposes a new lightweight…

Computer Vision and Pattern Recognition · Computer Science 2021-07-13 Yu-Ming Zhang , Chun-Chieh Lee , Jun-Wei Hsieh , Kuo-Chin Fan

Drone-to-drone detection using visual feed has crucial applications, such as detecting drone collisions, detecting drone attacks, or coordinating flight with other drones. However, existing methods are computationally costly, follow…

Computer Vision and Pattern Recognition · Computer Science 2023-08-29 Tushar Sangam , Ishan Rajendrakumar Dave , Waqas Sultani , Mubarak Shah

Video super-resolution (VSR) is the task of restoring high-resolution frames from a sequence of low-resolution inputs. Different from single image super-resolution, VSR can utilize frames' temporal information to reconstruct results with…

Image and Video Processing · Electrical Eng. & Systems 2022-08-25 Wenyi Lian , Wenjing Lian

In this paper, we present Shift Convolution Network (ShiftConvNet) to provide matching capability between two feature maps for stereo estimation. The proposed method can speedily produce a highly accurate disparity map from stereo images. A…

Computer Vision and Pattern Recognition · Computer Science 2019-11-21 Jian Xie

Extracting features from a huge amount of data for object recognition is a challenging task. Convolution neural network can be used to meet the challenge, but it often requires a large number of computation resources. In this paper, a…

Image and Video Processing · Electrical Eng. & Systems 2018-05-08 Yunlong Ma , Chunyan Wang

Humans can accurately determine whether the object in hand has slipped or not by visual and tactile perception. However, it is still a challenge for robots to detect in-hand object slip through visuo-tactile fusion. To address this issue, a…

Robotics · Computer Science 2023-02-28 Junli Gao , Zhaoji Huang , Zhaonian Tang , Haitao Song , Wenyu Liang

This paper investigates a novel active-sensing-based obstacle avoidance paradigm for flying robots in dynamic environments. Instead of fusing multiple sensors to enlarge the field of view (FOV), we introduce an alternative approach that…

Robotics · Computer Science 2021-02-18 Gang Chen , Wei Dong , Xinjun Sheng , Xiangyang Zhu , Han Ding

In this work, we present STOPNet, a framework for 6-DoF object suction detection on production lines, with a focus on but not limited to transparent objects, which is an important and challenging problem in robotic systems and modern…

Robotics · Computer Science 2023-10-10 Yuxuan Kuang , Qin Han , Danshi Li , Qiyu Dai , Lian Ding , Dong Sun , Hanlin Zhao , He Wang

In the cascaded approach to spoken language translation (SLT), the ASR output is typically punctuated and segmented into sentences before being passed to MT, since the latter is typically trained on written text. However, erroneous…

Computation and Language · Computer Science 2022-10-19 Sukanta Sen , Ondřej Bojar , Barry Haddow

We propose a network for Congested Scene Recognition called CSRNet to provide a data-driven and deep learning method that can understand highly congested scenes and perform accurate count estimation as well as present high-quality density…

Computer Vision and Pattern Recognition · Computer Science 2018-04-12 Yuhong Li , Xiaofan Zhang , Deming Chen

With the rapid development of deep learning, a variety of change detection methods based on deep learning have emerged in recent years. However, these methods usually require a large number of training samples to train the network model, so…

Computer Vision and Pattern Recognition · Computer Science 2023-11-08 Weidong Yan , Pei Yan , Li Cao

To mitigate the effects of shadow fading and obstacle blocking, reconfigurable intelligent surface (RIS) has become a promising technology to improve the signal transmission quality of wireless communications by controlling the…

Information Theory · Computer Science 2021-11-10 Wangyang Xu , Jiancheng An , Yongjun Xu , Chongwen Huang , Lu Gan , Chau Yuen

Having precise perception of the environment is crucial for ensuring the secure and reliable functioning of autonomous driving systems. Radar object detection networks are one fundamental part of such systems. CNN-based object detectors…

Computer Vision and Pattern Recognition · Computer Science 2023-08-16 Marius Lippke , Maurice Quach , Sascha Braun , Daniel Köhler , Michael Ulrich , Bastian Bischoff , Wei Yap Tan

Recently, relying on convolutional neural networks (CNNs), many methods for salient object detection in optical remote sensing images (ORSI-SOD) are proposed. However, most methods ignore the huge parameters and computational cost brought…

Computer Vision and Pattern Recognition · Computer Science 2023-04-04 Gongyang Li , Zhi Liu , Xinpeng Zhang , Weisi Lin

3D neural networks have become prevalent for many 3D vision tasks including object detection, segmentation, registration, and various perception tasks for 3D inputs. However, due to the sparsity and irregularity of 3D data, custom 3D…

Computer Vision and Pattern Recognition · Computer Science 2022-04-11 Junha Lee , Christopher Choy , Jaesik Park

Robust efficient loop closure detection is essential for large-scale real-time SLAM. In this paper, we propose a novel unsupervised deep neural network architecture of a feature embedding for visual loop closure that is both reliable and…

Robotics · Computer Science 2018-05-28 Nate Merrill , Guoquan Huang

Existing salient object detection (SOD) methods mainly rely on U-shaped convolution neural networks (CNNs) with skip connections to combine the global contexts and local spatial details that are crucial for locating salient objects and…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Yu Qiu , Yun Liu , Le Zhang , Jing Xu