中文
相关论文

相关论文: ICDAR 2021 Competition on Historical Map Segmentat…

200 篇论文

In this report, we present our findings from benchmarking experiments for information extraction on historical handwritten marriage records Esposalles from IEHHR - ICDAR 2017 robust reading competition. The information extraction is modeled…

计算机视觉与模式识别 · 计算机科学 2018-07-18 Animesh Prasad , Hervé Déjean , Jean-Luc Meunier , Max Weidemann , Johannes Michael , Gundram Leifert

Semantic segmentation benefits robotics related applications especially autonomous driving. Most of the research on semantic segmentation is only on increasing the accuracy of segmentation models with little attention to computationally…

计算机视觉与模式识别 · 计算机科学 2020-05-19 Mennatullah Siam , Mostafa Gamal , Moemen Abdel-Razek , Senthil Yogamani , Martin Jagersand

The presented algorithms for segmentation and tracking follow a 3-step approach where we detect, track and finally segment nuclei. In the preprocessing phase, we detect centroids of the cell nuclei using a convolutional neural network (CNN)…

计算机视觉与模式识别 · 计算机科学 2019-04-16 Dennis Eschweiler , Johannes Stegmaier

At the heart of all automated driving systems is the ability to sense the surroundings, e.g., through semantic segmentation of LiDAR sequences, which experienced a remarkable progress due to the release of large datasets such as…

计算机视觉与模式识别 · 计算机科学 2022-01-21 Kunyu Peng , Juncong Fei , Kailun Yang , Alina Roitberg , Jiaming Zhang , Frank Bieder , Philipp Heidenreich , Christoph Stiller , Rainer Stiefelhagen

In this work we predict vehicle speed and steering angle given camera image frames. Our key contribution is using an external pre-trained neural network for segmentation. We augment the raw images with their segmentation masks and mirror…

计算机视觉与模式识别 · 计算机科学 2019-10-24 Antonia Lovjer , Minsu Yeom , Benedikt D. Schifferer , Iddo Drori

The traditional object retrieval task aims to learn a discriminative feature representation with intra-similarity and inter-dissimilarity, which supposes that the objects in an image are manually or automatically pre-cropped exactly.…

计算机视觉与模式识别 · 计算机科学 2020-09-04 Lei Zhang , Zhenwei He , Yi Yang , Liang Wang , Xinbo Gao

Accurately and efficiently extracting building footprints from a wide range of remote sensed imagery remains a challenge due to their complex structure, variety of scales and diverse appearances. Existing convolutional neural network…

计算机视觉与模式识别 · 计算机科学 2020-10-01 Qing Zhu , Cheng Liao , Han Hu , Xiaoming Mei , Haifeng Li

Iris segmentation and localization in non-cooperative environment is challenging due to illumination variations, long distances, moving subjects and limited user cooperation, etc. Traditional methods often suffer from poor performance when…

计算机视觉与模式识别 · 计算机科学 2019-06-20 Caiyong Wang , Yuhao Zhu , Yunfan Liu , Ran He , Zhenan Sun

The binary segmentation of roads in very high resolution (VHR) remote sensing images (RSIs) has always been a challenging task due to factors such as occlusions (caused by shadows, trees, buildings, etc.) and the intra-class variances of…

计算机视觉与模式识别 · 计算机科学 2021-06-30 Lei Ding , Lorenzo Bruzzone

This paper presents the details of the Audio-Visual Scene Classification task in the DCASE 2021 Challenge (Task 1 Subtask B). The task is concerned with classification using audio and video modalities, using a dataset of synchronized…

音频与语音处理 · 电气工程与系统科学 2021-07-21 Shanshan Wang , Toni Heittola , Annamaria Mesaros , Tuomas Virtanen

Video activity Recognition has recently gained a lot of momentum with the release of massive Kinetics (400 and 600) data. Architectures such as I3D and C3D networks have shown state-of-the-art performances for activity recognition. The one…

计算机视觉与模式识别 · 计算机科学 2019-03-19 Manjot Bilkhu , Hammababdullah Ayyubi

In this paper, we have worked on interpretability, trust, and understanding of the decisions made by models in the form of classification tasks. The task is divided into 3 subtasks. The first task consists of determining Binary Sexism…

计算与语言 · 计算机科学 2023-04-11 Debashish Roy , Manish Shrivastava

This paper describes a methodology to produce a 7-classes land cover map of urban areas from very high resolution images and limited noisy labeled data. The objective is to make a segmentation map of a large area (a french department) with…

Reticular structures form the backbone of major infrastructure like bridges, pylons, and airports, but their inspection and maintenance are costly and hazardous, often requiring human intervention. While prior research has focused on fault…

Estimating building footprint maps from geospatial data is of paramount importance in urban planning, development, disaster management, and various other applications. Deep learning methodologies have gained prominence in building…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Anuja Vats , David Völgyes , Martijn Vermeer , Marius Pedersen , Kiran Raja , Daniele S. M. Fantin , Jacob Alexander Hay

Recently, text detection and recognition in natural scenes are becoming increasing popular in the computer vision community as well as the document analysis community. However, majority of the existing ideas, algorithms and systems are…

计算机视觉与模式识别 · 计算机科学 2015-06-11 Xinyu Zhou , Shuchang Zhou , Cong Yao , Zhimin Cao , Qi Yin

We report the findings of a month-long online competition in which participants developed algorithms for augmenting the digital version of patent documents published by the United States Patent and Trademark Office (USPTO). The goal was to…

计算机视觉与模式识别 · 计算机科学 2016-03-11 Christoph Riedl , Richard Zanibbi , Marti A. Hearst , Siyu Zhu , Michael Menietti , Jason Crusan , Ivan Metelsky , Karim R. Lakhani

Automated medical image segmentation plays an important role in many clinical applications, which however is a very challenging task, due to complex background texture, lack of clear boundary and significant shape and texture variation…

图像与视频处理 · 电气工程与系统科学 2020-10-26 Qikui Zhu , Liang Li , Jiangnan Hao , Yunfei Zha , Yan Zhang , Yanxiang Cheng , Fei Liao , Pingxiang Li

The focus of this paper is using a convolutional machine learning model with a modified U-Net structure for creating land cover classification mapping based on satellite imagery. The aim of the research is to train and test convolutional…

计算机视觉与模式识别 · 计算机科学 2020-03-09 Priit Ulmas , Innar Liiv

Deep learning has largely reduced the need for manual feature selection in image segmentation. Nevertheless, network architecture optimization and hyperparameter tuning are mostly manual and time consuming. Although there are increasing…

图像与视频处理 · 电气工程与系统科学 2019-09-16 Ken C. L. Wong , Mehdi Moradi