中文
相关论文

相关论文: U-Net with ResNet Backbone for Garment Landmarking…

200 篇论文

Fueled by the power of deep learning techniques and implicit shape learning, recent advances in single-image human digitalization have reached unprecedented accuracy and could recover fine-grained surface details such as garment wrinkles.…

计算机视觉与模式识别 · 计算机科学 2022-03-30 Heming Zhu , Lingteng Qiu , Yuda Qiu , Xiaoguang Han

Landmark localization in images and videos is a classic problem solved in various ways. Nowadays, with deep networks prevailing throughout machine learning, there are revamped interests in pushing facial landmark detection technologies to…

计算机视觉与模式识别 · 计算机科学 2019-08-16 Joseph P Robinson , Yuncheng Li , Ning Zhang , Yun Fu , and Sergey Tulyakov

In this paper we present our work on developing an automated system for land cover classification. This system takes a multiband satellite image of an area as input and outputs the land cover map of the area at the same resolution as the…

计算机视觉与模式识别 · 计算机科学 2020-10-14 Vasilis Pollatos , Loukas Kouvaras , Eleni Charou

Learning anatomical segmentation from heterogeneous labels in multi-center datasets is a common situation encountered in clinical scenarios, where certain anatomical structures are only annotated in images coming from particular medical…

图像与视频处理 · 电气工程与系统科学 2023-09-06 Nicolás Gaggion , Maria Vakalopoulou , Diego H. Milone , Enzo Ferrante

A common practice in transfer learning is to initialize the downstream model weights by pre-training on a data-abundant upstream task. In object detection specifically, the feature backbone is typically initialized with Imagenet classifier…

计算机视觉与模式识别 · 计算机科学 2022-06-28 Cristina Vasconcelos , Vighnesh Birodkar , Vincent Dumoulin

This paper introduces a novel segmentation framework that integrates a classifier network with a reverse HRNet architecture for efficient image segmentation. Our approach utilizes a ResNet-50 backbone, pretrained in a semi-supervised…

计算机视觉与模式识别 · 计算机科学 2024-02-12 Anupam Gupta , Ashok Krishnamurthy , Lisa Singh

Lane is critical in the vision navigation system of the intelligent vehicle. Naturally, lane is a traffic sign with high-level semantics, whereas it owns the specific local pattern which needs detailed low-level features to localize…

计算机视觉与模式识别 · 计算机科学 2022-03-22 Tu Zheng , Yifei Huang , Yang Liu , Wenjian Tang , Zheng Yang , Deng Cai , Xiaofei He

Residual network (ResNet) and densely connected network (DenseNet) have significantly improved the training efficiency and performance of deep convolutional neural networks (DCNNs) mainly for object classification tasks. In this paper, we…

图像与视频处理 · 电气工程与系统科学 2020-04-29 Mina Jafari , Dorothee Auer , Susan Francis , Jonathan Garibaldi , Xin Chen

Building extraction is an essential component of study in the science of remote sensing, and applications for building extraction heavily rely on semantic segmentation of high-resolution remote sensing imagery. Semantic information…

计算机视觉与模式识别 · 计算机科学 2023-10-12 Tareque Bashar Ovi , Nomaiya Bashree , Protik Mukherjee , Shakil Mosharrof , Masuma Anjum Parthima

Accurate and high precision of the indoor positioning is as important as ensuring reliable navigation in outdoor environments. Using the state-of-the-art deep learning models provides better reliability and accuracy to navigate and monitor…

信号处理 · 电气工程与系统科学 2025-08-19 Muhammad Ammad , Paul Schwarzbach , Michael Schultz , Oliver Michler

In recent years, deep learning has emerged as a promising technique for medical image analysis. However, this application domain is likely to suffer from a limited availability of large public datasets and annotations. A common solution to…

计算机视觉与模式识别 · 计算机科学 2025-06-25 Roberto Di Via , Matteo Santacesaria , Francesca Odone , Vito Paolo Pastore

This work presents a flexible system to reconstruct 3D models of objects captured with an RGB-D sensor. A major advantage of the method is that our reconstruction pipeline allows the user to acquire a full 3D model of the object. This is…

计算机视觉与模式识别 · 计算机科学 2015-05-22 Aitor Aldoma , Johann Prankl , Alexander Svejda , Markus Vincze

We consider image classification with estimated depth. This problem falls into the domain of transfer learning, since we are using a model trained on a set of depth images to generate depth maps (additional features) for use in another…

计算机视觉与模式识别 · 计算机科学 2017-09-22 Yihui He

The rapid progress of text-to-image diffusion models raises significant concerns regarding the unauthorized reproduction of trademarked content. While prior work targets general concepts (e.g., styles, celebrities), it fails to address…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Dawid Malarz , Filip Manjak , Maciej Zięba , Przemysław Spurek , Artur Kasymov

Capturing visual image with a hyperspectral camera has been successfully applied to many areas due to its narrow-band imaging technology. Hyperspectral reconstruction from RGB images denotes a reverse process of hyperspectral imaging by…

图像与视频处理 · 电气工程与系统科学 2020-05-12 Yuzhi Zhao , Lai-Man Po , Qiong Yan , Wei Liu , Tingyu Lin

This paper presents a novel approach for landmark recognition in images that we've successfully deployed at Mail ru. This method enables us to recognize famous places, buildings, monuments, and other landmarks in user photos. The main…

计算机视觉与模式识别 · 计算机科学 2019-08-30 Andrei Boiarov , Eduard Tyantov

This paper tackles the task of category-level pose estimation for garments. With a near infinite degree of freedom, a garment's full configuration (i.e., poses) is often described by the per-vertex 3D locations of its entire 3D surface.…

计算机视觉与模式识别 · 计算机科学 2021-08-17 Cheng Chi , Shuran Song

We propose a reversible face de-identification method for low resolution video data, where landmark-based techniques cannot be reliably used. Our solution is able to generate a photo realistic de-identified stream that meets the data…

计算机视觉与模式识别 · 计算机科学 2020-07-10 Hugo Proença

We introduce PhysXNet, a learning-based approach to predict the dynamics of deformable clothes given 3D skeleton motion sequences of humans wearing these clothes. The proposed model is adaptable to a large variety of garments and changing…

计算机视觉与模式识别 · 计算机科学 2021-11-16 Jordi Sanchez-Riera , Albert Pumarola , Francesc Moreno-Noguer

We present here, a novel network architecture called MergeNet for discovering small obstacles for on-road scenes in the context of autonomous driving. The basis of the architecture rests on the central consideration of training with less…

计算机视觉与模式识别 · 计算机科学 2018-03-20 Krishnam Gupta , Syed Ashar Javed , Vineet Gandhi , K. Madhava Krishna