中文
相关论文

相关论文: CubiCasa5K: A Dataset and an Improved Multi-Task M…

200 篇论文

We present a new annotated microscopic cellular image dataset to improve the effectiveness of machine learning methods for cellular image analysis. Cell counting is an important step in cell analysis. Typically, domain experts manually…

图像与视频处理 · 电气工程与系统科学 2024-11-20 Abdurahman Ali Mohammed , Catherine Fonder , Donald S. Sakaguchi , Wallapak Tavanapong , Surya K. Mallapragada , Azeez Idris

Missions to small celestial bodies rely heavily on optical feature tracking for characterization of and relative navigation around the target body. While deep learning has led to great advancements in feature detection and description,…

天体物理仪器与方法 · 物理学 2023-01-16 Travis Driver , Katherine Skinner , Mehregan Dor , Panagiotis Tsiotras

Indoor scene understanding is central to applications such as robot navigation and human companion assistance. Over the last years, data-driven deep neural networks have outperformed many traditional approaches thanks to their…

计算机视觉与模式识别 · 计算机科学 2017-07-04 Yinda Zhang , Shuran Song , Ersin Yumer , Manolis Savva , Joon-Young Lee , Hailin Jin , Thomas Funkhouser

Urban waste management remains a critical challenge for the development of smart cities. Despite the growing number of litter detection datasets, the problem of monitoring overflowing waste containers, particularly from images captured by…

计算机视觉与模式识别 · 计算机科学 2025-11-21 Diogo J. Paulo , João Martins , Hugo Proença , João C. Neves

In the past few years, computer vision and pattern recognition systems have been becoming increasingly more powerful, expanding the range of automatic tasks enabled by machine vision. Here we show that computer analysis of building images…

计算机视觉与模式识别 · 计算机科学 2023-06-22 Fan Wei , Yuan Li , Lior Shamir

Benchmark datasets in computer vision often contain off-topic images, near duplicates, and label errors, leading to inaccurate estimates of model performance. In this paper, we revisit the task of data cleaning and formalize it as either a…

When humans have to solve everyday tasks, they simply pick the objects that are most suitable. While the question which object should one use for a specific task sounds trivial for humans, it is very difficult to answer for robots or other…

计算机视觉与模式识别 · 计算机科学 2019-04-08 Johann Sawatzky , Yaser Souri , Christian Grund , Juergen Gall

Since a building's floorplan remains consistent over time and is inherently robust to changes in visual appearance, visual Floorplan Localization (FLoc) has received increasing attention from researchers. However, as a compact and…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Bolei Chen , Shengsheng Yan , Yongzheng Cui , Jiaxu Kang , Ping Zhong , Jianxin Wang

Fire scene datasets are crucial for training robust computer vision models, particularly in tasks such as fire early warning and emergency rescue operations. However, among the currently available fire-related data, there is a significant…

计算机视觉与模式识别 · 计算机科学 2026-04-07 Haozhou Zhai , Yanzhe Gao , Tianjiang Hu

For the last few decades, several major subfields of artificial intelligence including computer vision, graphics, and robotics have progressed largely independently from each other. Recently, however, the community has realized that…

计算机视觉与模式识别 · 计算机科学 2022-06-06 Yiyi Liao , Jun Xie , Andreas Geiger

Object pose estimation is crucial for robotic applications and augmented reality. Beyond instance level 6D object pose estimation methods, estimating category-level pose and shape has become a promising trend. As such, a new research field…

计算机视觉与模式识别 · 计算机科学 2022-05-19 Pengyuan Wang , HyunJun Jung , Yitong Li , Siyuan Shen , Rahul Parthasarathy Srikanth , Lorenzo Garattoni , Sven Meier , Nassir Navab , Benjamin Busam

Architectural floor plans are widely available priors which contain not only geometry but also the semantic information of the environment, yet existing localization methods largely ignore this semantic information. To address this, we…

计算机视觉与模式识别 · 计算机科学 2026-04-29 Muhammad Shaheer , Miguel Fernandez-Cortizas , Asier Bikandi-Noya , Holger Voos , Jose Luis Sanchez-Lopez

We address the problem of learning a single model for person re-identification, attribute classification, body part segmentation, and pose estimation. With predictions for these tasks we gain a more holistic understanding of persons, which…

计算机视觉与模式识别 · 计算机科学 2020-11-10 Kilian Pfeiffer , Alexander Hermans , István Sárándi , Mark Weber , Bastian Leibe

With the enhancement of remote sensing image resolution and the rapid advancement of deep learning, land cover mapping is transitioning from pixel-level segmentation to object-based vector modeling. This shift demands more from deep…

计算机视觉与模式识别 · 计算机科学 2025-08-25 Yu Meng , Ligao Deng , Zhihao Xi , Jiansheng Chen , Jingbo Chen , Anzhi Yue , Diyou Liu , Kai Li , Chenhao Wang , Kaiyu Li , Yupeng Deng , Xian Sun

We address 2D floorplan reconstruction from 3D scans. Existing approaches typically employ heuristically designed multi-stage pipelines. Instead, we formulate floorplan reconstruction as a single-stage structured prediction task: find a…

计算机视觉与模式识别 · 计算机科学 2023-03-29 Yuanwen Yue , Theodora Kontogianni , Konrad Schindler , Francis Engelmann

In this work, we introduce VQA 360, a novel task of visual question answering on 360 images. Unlike a normal field-of-view image, a 360 image captures the entire visual content around the optical center of a camera, demanding more…

计算机视觉与模式识别 · 计算机科学 2020-01-13 Shih-Han Chou , Wei-Lun Chao , Wei-Sheng Lai , Min Sun , Ming-Hsuan Yang

The Open Images Dataset contains approximately 9 million images and is a widely accepted dataset for computer vision research. As is common practice for large datasets, the annotations are not exhaustive, with bounding boxes and attribute…

计算机视觉与模式识别 · 计算机科学 2021-05-07 Candice Schumann , Susanna Ricco , Utsav Prabhu , Vittorio Ferrari , Caroline Pantofaru

Training deep neural networks requires datasets with a large number of annotated examples. The collection and annotation of these datasets is not only extremely expensive but also faces legal and privacy problems. These factors are a…

计算机视觉与模式识别 · 计算机科学 2025-01-17 Christoph Reinders , Frederik Schubert , Bodo Rosenhahn

Active learning has emerged as a promising approach to reduce the substantial annotation burden in 3D object detection tasks, spurring several initiatives in outdoor environments. However, its application in indoor environments remains…

计算机视觉与模式识别 · 计算机科学 2025-03-21 Jiangyi Wang , Na Zhao

Large collections of geo-referenced panoramic images are freely available for cities across the globe, as well as detailed maps with location and meta-data on a great variety of urban objects. They provide a potentially rich source of…

计算机视觉与模式识别 · 计算机科学 2022-09-01 Inske Groenen , Stevan Rudinac , Marcel Worring