中文
相关论文

相关论文: CroBIM-U: Uncertainty-Driven Referring Remote Sens…

200 篇论文

Reliable localization is critical for robot navigation in complex indoor environments. In this paper, we propose an uncertainty-aware localization method that enhances the reliability of localization outputs without modifying the prediction…

机器人学 · 计算机科学 2025-04-23 Hye-Min Won , Jieun Lee , Jiyong Oh

Recently, a surge in scientific publications suspected of image manipulation has led to numerous retractions, bringing the issue of image integrity into sharp focus. Although research on forensic detectors for image plagiarism and image…

计算机视觉与模式识别 · 计算机科学 2024-04-19 Xun Lin , Wenzhong Tang , Haoran Wang , Yizhong Liu , Yakun Ju , Shuai Wang , Zitong Yu

Despite advances in generic object detection, there remains a performance gap in detecting small objects compared to normal-scale objects. We reveal that conventional object localization methods suffer from gradient instability in small…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Huixin Sun , Yanjing Li , Linlin Yang , Xianbin Cao , Baochang Zhang

Few shot segmentation (FSS) aims to learn pixel-level classification of a target object in a query image using only a few annotated support samples. This is challenging as it requires modeling appearance variations of target objects and the…

计算机视觉与模式识别 · 计算机科学 2021-10-19 Soopil Kim , Philip Chikontwe , Sang Hyun Park

Guided ultrasonic wave localization uses spatially distributed multistatic sensor arrays and generalized beamforming strategies to detect and locate damage across a structure. The propagation channel is often very complex. Methods can…

信号处理 · 电气工程与系统科学 2024-10-30 Ishan D. Khurjekar , Joel B. Harley

The usage of convolutional neural networks (CNNs) for unsupervised image segmentation was investigated in this study. In the proposed approach, label prediction and network parameter learning are alternately iterated to meet the following…

计算机视觉与模式识别 · 计算机科学 2020-08-26 Wonjik Kim , Asako Kanezaki , Masayuki Tanaka

Efficient intravascular access in trauma and critical care significantly impacts patient outcomes. However, the availability of skilled medical personnel in austere environments is often limited. Autonomous robotic ultrasound systems can…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Rohini Banerjee , Cecilia G. Morales , Artur Dubrawski

GUI grounding, which localizes interface elements from screenshots given natural language queries, remains challenging for small icons and dense layouts. Test-time zoom-in methods improve localization by cropping and re-running inference at…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Fei Tang , Bofan Chen , Zhengxi Lu , Tongbo Chen , Songqin Nong , Tao Jiang , Wenhao Xu , Weiming Lu , Jun Xiao , Yueting Zhuang , Yongliang Shen

This paper proposes a depth estimation method using radar-image fusion by addressing the uncertain vertical directions of sparse radar measurements. In prior radar-image fusion work, image features are merged with the uncertain sparse…

计算机视觉与模式识别 · 计算机科学 2024-03-26 Masaya Kotani , Takeru Oba , Norimichi Ukita

Deep neural networks (DNNs) have witnessed great successes in semantic segmentation, which requires a large number of labeled data for training. We present a novel learning framework called Uncertainty guided Cross-head Co-training (UCC)…

计算机视觉与模式识别 · 计算机科学 2023-02-24 Jiashuo Fan , Bin Gao , Huan Jin , Lihui Jiang

Neural Radiance Field (NeRF)-based segmentation methods focus on object semantics and rely solely on RGB data, lacking intrinsic material properties. This limitation restricts accurate material perception, which is crucial for robotics,…

图像与视频处理 · 电气工程与系统科学 2025-08-07 Fabian Perez , Sara Rojas , Carlos Hinojosa , Hoover Rueda-Chacón , Bernard Ghanem

Recent advancements in codebook-based real image super-resolution (SR) have shown promising results in real-world applications. The core idea involves matching high-quality image features from a codebook based on low-resolution (LR) image…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Weilei Wen , Tianyi Zhang , Qianqian Zhao , Zhaohui Zheng , Chunle Guo , Xiuli Shao , Chongyi Li

Most of the existing blind image Super-Resolution (SR) methods assume that the blur kernels are space-invariant. However, the blur involved in real applications are usually space-variant due to object motion, out-of-focus, etc., resulting…

计算机视觉与模式识别 · 计算机科学 2023-04-10 Xuhai Chen , Jiangning Zhang , Chao Xu , Yabiao Wang , Chengjie Wang , Yong Liu

Super resolution techniques can enhance the spatial resolution of remote sensing images, enabling more efficient large scale earth observation applications. While single image SR methods enhance low resolution images, they neglect valuable…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Ce Wang , Wanjie Sun

Deep learning relies heavily on data augmentation to mitigate limited data, especially in medical imaging. Recent multimodal learning integrates text and images for segmentation, known as referring or text-guided image segmentation.…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Shurong Chai , Rahul Kumar JAIN , Rui Xu , Shaocong Mo , Ruibo Hou , Shiyu Teng , Jiaqing Liu , Lanfen Lin , Yen-Wei Chen

Semantic-aware 3D reconstruction from sparse, unposed images remains challenging for feed-forward 3D Gaussian Splatting (3DGS). Existing methods often predict an over-complete set of Gaussian primitives under sparse-view supervision,…

计算机视觉与模式识别 · 计算机科学 2026-03-19 Guibiao Liao , Qian Ren , Kaimin Liao , Hua Wang , Zhi Chen , Luchao Wang , Yaohua Tang

Image-to-image translation is an ill-posed problem as unique one-to-one mapping may not exist between the source and target images. Learning-based methods proposed in this context often evaluate the performance on test data that is similar…

图像与视频处理 · 电气工程与系统科学 2021-10-08 Uddeshya Upadhyay , Viswanath P. Sudarshan , Suyash P. Awate

Referring Image Segmentation (RIS) is a fundamental vision-language task that outputs object masks based on text descriptions. Many works have achieved considerable progress for RIS, including different fusion method designs. In this work,…

计算机视觉与模式识别 · 计算机科学 2023-07-25 Jianzong Wu , Xiangtai Li , Xia Li , Henghui Ding , Yunhai Tong , Dacheng Tao

Fluorescence microscopy images contain several channels, each indicating a marker staining the sample. Since many different marker combinations are utilized in practice, it has been challenging to apply deep learning based segmentation…

计算机视觉与模式识别 · 计算机科学 2021-01-28 Alvaro Gomariz , Raphael Egli , Tiziano Portenier , César Nombela-Arrieta , Orcun Goksel

Confidence alone is often misleading in hyperspectral image classification, as models tend to mistake high predictive scores for correctness while lacking awareness of uncertainty. This leads to confirmation bias, especially under sparse…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Muzhou Yang , Wuzhou Quan , Mingqiang Wei