English
Related papers

Related papers: Image2PCI -- A Multitask Learning Framework for Es…

200 papers

Interpretation of deep learning remains a very challenging problem. Although the Class Activation Map (CAM) is widely used to interpret deep model predictions by highlighting object location, it fails to provide insight into the salient…

Computer Vision and Pattern Recognition · Computer Science 2024-05-30 Yuguang Yang , Runtang Guo , Sheng Wu , Yimi Wang , Juan Zhang , Xuan Gong , Baochang Zhang

In this paper, we present a novel deep image clustering approach termed PICI, which enforces the partial information discrimination and the cross-level interaction in a joint learning framework. In particular, we leverage a Transformer…

Computer Vision and Pattern Recognition · Computer Science 2024-01-25 Hai-Xin Zhang , Dong Huang , Hua-Bao Ling , Guang-Yu Zhang , Wei-jun Sun , Zi-hao Wen

Large-scale training have propelled significant progress in various sub-fields of AI such as computer vision and natural language processing. However, building robot learning systems at a comparable scale remains challenging. To develop…

Robotics · Computer Science 2023-02-17 Zhao Mandi , Homanga Bharadhwaj , Vincent Moens , Shuran Song , Aravind Rajeswaran , Vikash Kumar

Light inherently consists of multiple dimensions beyond intensity, including spectrum, polarization, etc. The coupling among these high-dimensional optical features provides a compressive characterization of intrinsic material properties.…

Optics · Physics 2025-05-02 Liheng Bian , Zhen Wang , Pengming Peng , Zhengyi Zhao , Rong Yan , Hanwen Xu , Jun Zhang

Inertial sensors are crucial for recognizing pedestrian activity. Recent advances in deep learning have greatly improved inertial sensing performance and robustness. Different domains and platforms use deep-learning techniques to enhance…

Machine Learning · Computer Science 2025-12-16 Zeev Yampolsky , Ofir Kruzel , Victoria Khalfin Fekson , Itzik Klein

We propose a new method to analyze the impact of errors in algorithms for multi-instance pose estimation and a principled benchmark that can be used to compare them. We define and characterize three classes of errors - localization,…

Computer Vision and Pattern Recognition · Computer Science 2017-08-08 Matteo Ruggero Ronchi , Pietro Perona

Satellites continuously generate massive volumes of data, particularly for Earth observation, including satellite image time series (SITS). However, most deep learning models are designed to process either entire images or complete time…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Leandro Stival , Ricardo da Silva Torres , Helio Pedrini

Sensor-based human activity segmentation and recognition are two important and challenging problems in many real-world applications and they have drawn increasing attention from the deep learning community in recent years. Most of the…

Computer Vision and Pattern Recognition · Computer Science 2023-03-21 Furong Duan , Tao Zhu , Jinqiang Wang , Liming Chen , Huansheng Ning , Yaping Wan

Recently years, the attempts on distilling mobile data into useful knowledge has been led to the deployment of machine learning algorithms at the network edge. Principal component analysis (PCA) is a classic technique for extracting the…

Information Theory · Computer Science 2022-04-04 Zezhong Zhang , Guangxu Zhu , Rui Wang , Vincent K. N. Lau , Kaibin Huang

Automated pavement crack detection and measurement are important road issues. Agencies have to guarantee the improvement of road safety. Conventional crack detection and measurement algorithms can be extremely time-consuming and low…

Computer Vision and Pattern Recognition · Computer Science 2020-02-11 Zhun Fan , Chong Li , Ying Chen , Paola Di Mascio , Xiaopeng Chen , Guijie Zhu , Giuseppe Loprencipe

Local structure such as context-specific independence (CSI) has received much attention in the probabilistic graphical model (PGM) literature, as it facilitates the modeling of large complex systems, as well as for reasoning with them. In…

Artificial Intelligence · Computer Science 2020-06-15 Yujia Shen , Arthur Choi , Adnan Darwiche

Among the various generative adversarial network (GAN)-based image inpainting methods, a coarse-to-fine network with a contextual attention module (CAM) has shown remarkable performance. However, owing to two stacked generative networks,…

Computer Vision and Pattern Recognition · Computer Science 2020-03-20 Yong-Goo Shin , Min-Cheol Sagong , Yoon-Jae Yeo , Seung-Wook Kim , Sung-Jea Ko

Controllable character image animation has a wide range of applications. Although existing studies have consistently improved performance, challenges persist in the field of character image animation, particularly concerning stability in…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Jingyun Xue , Hongfa Wang , Qi Tian , Yue Ma , Andong Wang , Zhiyuan Zhao , Shaobo Min , Wenzhe Zhao , Kaihao Zhang , Heung-Yeung Shum , Wei Liu , Mengyang Liu , Wenhan Luo

Monocular depth estimation and defocus estimation are two fundamental tasks in computer vision. Most existing methods treat depth estimation and defocus estimation as two separate tasks, ignoring the strong connection between them. In this…

Computer Vision and Pattern Recognition · Computer Science 2022-08-23 Renzhi He , Hualin Hong , Boya Fu , Fei Liu

Capabilities of inference and prediction are significant components of visual systems. In this paper, we address an important and challenging task of them: visual path prediction. Its goal is to infer the future path for a visual object in…

Computer Vision and Pattern Recognition · Computer Science 2016-12-16 Siyu Huang , Xi Li , Zhongfei Zhang , Zhouzhou He , Fei Wu , Wei Liu , Jinhui Tang , Yueting Zhuang

Deep learning models, specifically convolutional neural networks, have transformed the landscape of image classification by autonomously extracting features directly from raw pixel data. This article introduces an innovative image…

Image and Video Processing · Electrical Eng. & Systems 2024-12-19 Fatemeh Froughirad , Reza Bakhoda Eshtivani , Hamed Khajavi , Amir Rastgoo

Image-to-image (I2I) translation is a pixel-level mapping that requires a large number of paired training data and often suffers from the problems of high diversity and strong category bias in image scenes. In order to tackle these…

Computer Vision and Pattern Recognition · Computer Science 2019-04-22 Liqian Ma , Qianru Sun , Bernt Schiele , Luc Van Gool

For a number of tasks, such as 3D reconstruction, robotic interface, autonomous driving, etc., camera calibration is essential. In this study, we present a unique method for predicting intrinsic (principal point offset and focal length) and…

Computer Vision and Pattern Recognition · Computer Science 2022-12-26 Talha Hanif Butt , Murtaza Taj

Autonomous parking systems start with the detection of available parking slots. Parking slot detection performance has been dramatically improved by deep learning techniques. Deep learning-based object detection methods can be categorized…

Computer Vision and Pattern Recognition · Computer Science 2021-08-16 Quang Huy Bui , Jae Kyu Suhr

This paper presents parametric instance classification (PIC) for unsupervised visual feature learning. Unlike the state-of-the-art approaches which do instance discrimination in a dual-branch non-parametric fashion, PIC directly performs a…

Computer Vision and Pattern Recognition · Computer Science 2020-06-26 Yue Cao , Zhenda Xie , Bin Liu , Yutong Lin , Zheng Zhang , Han Hu