English
Related papers

Related papers: Pix2Poly: A Sequence Prediction Method for End-to-…

200 papers

This paper considers the problem of extracting building footprints from satellite imagery -- a task that is critical for many urban planning and decision-making applications. While recent advancements in deep learning have made great…

Computer Vision and Pattern Recognition · Computer Science 2023-11-07 Muhammad Ahmad Waseem , Muhammad Tahir , Zubair Khalid , Momin Uppal

Accurately and efficiently extracting building footprints from a wide range of remote sensed imagery remains a challenge due to their complex structure, variety of scales and diverse appearances. Existing convolutional neural network…

Computer Vision and Pattern Recognition · Computer Science 2020-10-01 Qing Zhu , Cheng Liao , Han Hu , Xiaoming Mei , Haifeng Li

The growing demand for detailed building roof data has driven the development of automated extraction methods to overcome the inefficiencies of traditional approaches, particularly in handling complex variations in building geometries.…

Computer Vision and Pattern Recognition · Computer Science 2025-03-17 Chaikal Amrullah , Daniel Panangian , Ksenia Bittner

Topological features such as persistence diagrams and their functional approximations like persistence images (PIs) have been showing substantial promise for machine learning and computer vision applications. This is greatly attributed to…

Computer Vision and Pattern Recognition · Computer Science 2020-05-26 Anirudh Som , Hongjun Choi , Karthikeyan Natesan Ramamurthy , Matthew Buman , Pavan Turaga

Recently, Transformer-based text detection techniques have sought to predict polygons by encoding the coordinates of individual boundary vertices using distinct query features. However, this approach incurs a significant memory overhead and…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Xuyang Chen , Dong Wang , Konrad Schindler , Mingwei Sun , Yongliang Wang , Nicolo Savioli , Liqiu Meng

We propose DeepV2D, an end-to-end deep learning architecture for predicting depth from video. DeepV2D combines the representation ability of neural networks with the geometric principles governing image formation. We compose a collection of…

Computer Vision and Pattern Recognition · Computer Science 2020-04-29 Zachary Teed , Jia Deng

Polygonal road outline extraction from high-resolution aerial images is an important task in large-scale topographic mapping, where roads are represented as vectorized polygons, capturing essential geometric features with minimal vertex…

Computer Vision and Pattern Recognition · Computer Science 2025-04-30 Weiqin Jiao , Hao Cheng , George Vosselman , Claudio Persello

Interpretability of neural networks and their underlying theoretical behavior remain an open field of study even after the great success of their practical applications, particularly with the emergence of deep learning. In this work,…

Machine Learning · Statistics 2023-11-16 Pablo Morala , Jenny Alexandra Cifuentes , Rosa E. Lillo , Iñaki Ucar

Building footprint maps are vital to many remote sensing applications, such as 3D building modeling, urban planning, and disaster management. Due to the complexity of buildings, the accurate and reliable generation of the building footprint…

Computer Vision and Pattern Recognition · Computer Science 2020-12-02 Qingyu Li , Yilei Shi , Xin Huang , Xiao Xiang Zhu

Exploring contextual information in the local region is important for shape understanding and analysis. Existing studies often employ hand-crafted or explicit ways to encode contextual information of local regions. However, it is hard to…

Computer Vision and Pattern Recognition · Computer Science 2018-11-16 Xinhai Liu , Zhizhong Han , Yu-Shen Liu , Matthias Zwicker

Predicting a binary mask for an object is more accurate but also more computationally expensive than a bounding box. Polygonal masks as developed in CenterPoly can be a good compromise. In this paper, we improve over CenterPoly by enhancing…

Computer Vision and Pattern Recognition · Computer Science 2023-05-10 Katia Jodogne-Del Litto , Guillaume-Alexandre Bilodeau

The long-coveted task of reconstructing 3D geometry from images is still a standing problem. In this paper, we build on the power of neural networks and introduce Pix2Vex, a network trained to convert camera-captured images into 3D…

Computer Vision and Pattern Recognition · Computer Science 2019-05-28 Felix Petersen , Amit H. Bermano , Oliver Deussen , Daniel Cohen-Or

Good quality reconstruction and comprehension of a scene rely on 3D estimation methods. The 3D information was usually obtained from images by stereo-photogrammetry, but deep learning has recently provided us with excellent results for…

Computer Vision and Pattern Recognition · Computer Science 2021-08-02 Rémy Leroy , Pauline Trouvé-Peloux , Frédéric Champagnat , Bertrand Le Saux , Marcela Carvalho

We introduce P2P-NET, a general-purpose deep neural network which learns geometric transformations between point-based shape representations from two domains, e.g., meso-skeletons and surfaces, partial and complete scans, etc. The…

Graphics · Computer Science 2018-05-16 Kangxue Yin , Hui Huang , Daniel Cohen-Or , Hao Zhang

Effectively parsing the facade is essential to 3D building reconstruction, which is an important computer vision problem with a large amount of applications in high precision map for navigation, computer aided design, and city generation…

Computer Vision and Pattern Recognition · Computer Science 2021-06-03 Hantang Liu , Wentong Li , Jianke Zhu

Road network and building footprint extraction is essential for many applications such as updating maps, traffic regulations, city planning, ride-hailing, disaster response \textit{etc}. Mapping road networks is currently both expensive and…

Computer Vision and Pattern Recognition · Computer Science 2020-10-15 An Tran , Ali Zonoozi , Jagannadan Varadarajan , Hannes Kruppa

We propose a decentralised "local2global" approach to graph representation learning, that one can a-priori use to scale any embedding technique. Our local2global approach proceeds by first dividing the input graph into overlapping subgraphs…

Machine Learning · Computer Science 2021-07-27 Lucas G. S. Jeub , Giovanni Colavizza , Xiaowen Dong , Marya Bazzi , Mihai Cucuringu

Extracting lane topology from perspective views (PV) is crucial for planning and control in autonomous driving. This approach extracts potential drivable trajectories for self-driving vehicles without relying on high-definition (HD) maps.…

Computer Vision and Pattern Recognition · Computer Science 2025-07-02 Yiming Yang , Yueru Luo , Bingkun He , Erlong Li , Zhipeng Cao , Chao Zheng , Shuqi Mei , Zhen Li

A key step in any scanning-based asset creation workflow is to convert unordered point clouds to a surface. Classical methods (e.g., Poisson reconstruction) start to degrade in the presence of noisy and partial scans. Hence, deep learning…

Computer Vision and Pattern Recognition · Computer Science 2024-02-14 Philipp Erler , Paul Guerrero , Stefan Ohrhallinger , Michael Wimmer , Niloy J. Mitra

The accurate and automatic extraction of roads from satellite imagery is critical for applications in navigation and urban planning, significantly reducing the need for manual annotation. Many existing methods decompose this task into…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Zhengyang Wei , Renzhi Jing , Yiyi He , Jenny Suckale