中文
相关论文

相关论文: On Coordinate Decoding for Keypoint Estimation Tas…

200 篇论文

Effectively representing heterogeneous tabular datasets for meta-learning purposes is still an open problem. Previous approaches rely on representations that are intended to be universal. This paper proposes two novel methods for tabular…

机器学习 · 计算机科学 2025-07-18 Antoni Zajko , Katarzyna Woźnica

Event camera is an asynchronous, high frequency vision sensor with low power consumption, which is suitable for human action understanding task. It is vital to encode the spatial-temporal information of event data properly and use standard…

计算机视觉与模式识别 · 计算机科学 2021-04-14 Chaoxing Huang

Several SLAM methods benefit from the use of semantic information. Most integrate photometric methods with high-level semantics such as object detection and semantic segmentation. We propose that adding a semantic segmentation decoder in a…

计算机视觉与模式识别 · 计算机科学 2022-11-03 Gabriel S. Gama , Nícolas S. Rosa , Valdir Grassi

The ability to accurately evaluate the performance of location determination systems is crucial for many applications. Typically, the performance of such systems is obtained by comparing ground truth locations with estimated locations.…

信号处理 · 电气工程与系统科学 2021-06-28 Chen Gu , Ahmed Shokry , Moustafa Youssef

We introduce a novel method for robust and accurate 3D object pose estimation from a single color image under large occlusions. Following recent approaches, we first predict the 2D projections of 3D points related to the target object and…

计算机视觉与模式识别 · 计算机科学 2018-07-27 Markus Oberweger , Mahdi Rad , Vincent Lepetit

In recent years, huge progress has been made on learning neural implicit representations from multi-view images for 3D reconstruction. As an additional input complementing coordinates, using sinusoidal functions as positional encodings…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Sijia Jiang , Jing Hua , Zhizhong Han

Estimating the 3D pose of a hand from a 2D image is a well-studied problem and a requirement for several real-life applications such as virtual reality, augmented reality, and hand gesture recognition. Currently, reasonable estimations can…

计算机视觉与模式识别 · 计算机科学 2022-05-10 Danilo Avola , Luigi Cinque , Alessio Fagioli , Gian Luca Foresti , Adriano Fragomeni , Daniele Pannone

The 2D heatmap-based approaches have dominated Human Pose Estimation (HPE) for years due to high performance. However, the long-standing quantization error problem in the 2D heatmap-based methods leads to several well-known drawbacks: 1)…

计算机视觉与模式识别 · 计算机科学 2022-07-06 Yanjie Li , Sen Yang , Peidong Liu , Shoukui Zhang , Yunxiao Wang , Zhicheng Wang , Wankou Yang , Shu-Tao Xia

The task of room layout estimation is to locate the wall-floor, wall-ceiling, and wall-wall boundaries. Most recent methods solve this problem based on edge/keypoint detection or semantic segmentation. However, these approaches have shown…

计算机视觉与模式识别 · 计算机科学 2020-08-17 Weidong Zhang , Wei Zhang , Yinda Zhang

While invaluable for many computer vision applications, decomposing a natural image into intrinsic reflectance and shading layers represents a challenging, underdetermined inverse problem. As opposed to strict reliance on conventional…

计算机视觉与模式识别 · 计算机科学 2018-09-03 Qingnan Fan , Jiaolong Yang , Gang Hua , Baoquan Chen , David Wipf

Encoding 3D points is one of the primary steps in learning-based implicit scene representation. Using features that gather information from neighbors with multi-resolution grids has proven to be the best geometric encoder for this task.…

计算机视觉与模式识别 · 计算机科学 2024-02-13 Arihant Gaur , G. Dias Pais , Pedro Miraldo

We introduce a technique for 3D human keypoint estimation that directly models the notion of spatial uncertainty of a keypoint. Our technique employs a principled approach to modelling spatial uncertainty inspired from techniques in robust…

计算机视觉与模式识别 · 计算机科学 2020-12-22 Francis Williams , Or Litany , Avneesh Sud , Kevin Swersky , Andrea Tagliasacchi

We rethink the role of positional encoding in 3D representation learning and fine-tuning. We argue that using positional encoding in point Transformer-based methods serves to aggregate multi-scale features of point clouds. Additionally, we…

计算机视觉与模式识别 · 计算机科学 2025-09-25 Shaochen Zhang , Zekun Qi , Runpei Dong , Xiuxiu Bai , Xing Wei

The practical application of deep neural networks are still limited by their lack of transparency. One of the efforts to provide explanation for decisions made by artificial intelligence (AI) is the use of saliency or heat maps highlighting…

计算机视觉与模式识别 · 计算机科学 2020-10-20 Erico Tjoa , Guan Cuntai

In this paper, we propose a deep neural network approach for mapping the 2D pixel coordinates in an image to the corresponding Red-Green-Blue (RGB) color values. The neural network is termed CocoNet, i.e. coordinates-to-color network.…

计算机视觉与模式识别 · 计算机科学 2018-11-26 Paul Andrei Bricman , Radu Tudor Ionescu

Quantum cryptography via key distribution mechanisms that utilize quantum entanglement between sender-receiver pairs will form the basis of future large-scale quantum networks. A key engineering challenge in such networks will be the…

信息论 · 计算机科学 2015-05-14 Yixuan Xie , Jun Li , Robert Malaney , Jinhong Yuan

Semantic keypoints provide concise abstractions for a variety of visual understanding tasks. Existing methods define semantic keypoints separately for each category with a fixed number of semantic labels in fixed indices. As a result, this…

计算机视觉与模式识别 · 计算机科学 2018-07-27 Xingyi Zhou , Arjun Karpur , Linjie Luo , Qixing Huang

We present HashEncoding, a novel autoencoding architecture that leverages a non-parametric multiscale coordinate hash function to facilitate a per-pixel decoder without convolutions. By leveraging the space-folding behaviour of hashing…

计算机视觉与模式识别 · 计算机科学 2022-11-30 Lukas Zhornyak , Zhengjie Xu , Haoran Tang , Jianbo Shi

The problem of counterfactual visual explanations is considered. A new family of discriminant explanations is introduced. These produce heatmaps that attribute high scores to image regions informative of a classifier prediction but not of a…

计算机视觉与模式识别 · 计算机科学 2020-04-17 Pei Wang , Nuno Vasconcelos

Modern 3D semantic scene graph estimation methods utilize ground truth 3D annotations to accurately predict target objects, predicates, and relationships. In the absence of given 3D ground truth representations, we explore leveraging only…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Qi Xun Yeo , Yanyan Li , Gim Hee Lee