中文
相关论文

相关论文: PanORama: Multiview Consistent Panoptic Segmentati…

200 篇论文

A major focus of clinical imaging workflow is disease diagnosis and management, leading to medical imaging datasets strongly tied to specific clinical objectives. This scenario has led to the prevailing practice of developing task-specific…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Yunhe Gao , Zhuowei Li , Di Liu , Mu Zhou , Shaoting Zhang , Dimitris N. Metaxas

Segmenting objects in an environment is a crucial task for autonomous driving and robotics, as it enables a better understanding of the surroundings of each agent. Although camera sensors provide rich visual details, they are vulnerable to…

计算机视觉与模式识别 · 计算机科学 2025-05-07 Huawei Sun , Bora Kunter Sahin , Georg Stettinger , Maximilian Bernhard , Matthias Schubert , Robert Wille

Panoramic images can broaden the Field of View (FoV), occlusion-aware prediction can deepen the understanding of the scene, and domain adaptation can transfer across viewing domains. In this work, we introduce a novel task, Occlusion-Aware…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Yihong Cao , Jiaming Zhang , Hao Shi , Kunyu Peng , Yuhongxuan Zhang , Hui Zhang , Rainer Stiefelhagen , Kailun Yang

Intraoperative optical imaging is essential for surgical precision and patient safety, but current systems present anatomical and fluorescence information separately, causing delays and increasing cognitive load. A unified system for…

Scene rearrangement, like table tidying, is a challenging task in robotic manipulation due to the complexity of predicting diverse object arrangements. Web-scale trained generative models such as Stable Diffusion can aid by generating…

机器人学 · 计算机科学 2024-12-03 Shutong Jin , Ruiyu Wang , Kuangyi Chen , Florian T. Pokorny

Immersive scene generation, notably panorama creation, benefits significantly from the adaptation of large pre-trained text-to-image (T2I) models for multi-view image generation. Due to the high cost of acquiring multi-view images,…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Aoming Liu , Zhong Li , Zhang Chen , Nannan Li , Yi Xu , Bryan A. Plummer

High-quality panoramic images with a Field of View (FoV) of 360{\deg} are essential for contemporary panoramic computer vision tasks. However, conventional imaging systems come with sophisticated lens designs and heavy optical components.…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Qi Jiang , Shaohua Gao , Yao Gao , Kailun Yang , Zhonghua Yi , Hao Shi , Lei Sun , Kaiwei Wang

Multimodal large laboratory models (MLLMs) still struggle with spatial understanding under the dominant perspective-image paradigm, which inherits the narrow field of view of human-like perception. For navigation, robotic search, and 3D…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Changpeng Wang , Xin Lin , Junhan Liu , Yuheng Liu , Zhen Wang , Donglian Qi , Yunfeng Yan , Xi Chen

Panoptic segmentation, which combines instance and semantic segmentation, has gained a lot of attention in autonomous vehicles, due to its comprehensive representation of the scene. This task can be applied for cameras and LiDAR sensors,…

计算机视觉与模式识别 · 计算机科学 2024-12-31 Fardin Ayar , Ehsan Javanmardi , Manabu Tsukada , Mahdi Javanmardi , Mohammad Rahmati

Mammography is crucial for breast cancer surveillance and early diagnosis. However, analyzing mammography images is a demanding task for radiologists, who often review hundreds of mammograms daily, leading to overdiagnosis and…

计算机视觉与模式识别 · 计算机科学 2024-07-22 Kun Zhao , Jakub Prokop , Javier Montalt Tordera , Sadegh Mohammadi

Enhancing visual odometry by exploiting sparse depth measurements from LiDAR is a promising solution for improving tracking accuracy of an odometry. Most existing works utilize a monocular pinhole camera, yet could suffer from poor…

机器人学 · 计算机科学 2025-09-16 Qirui Hu , Zikang Yuan , Tianle Xu , Xiaoxiang Wang , Jinni Geng , Xin Yang

Universal Image Segmentation is not a new concept. Past attempts to unify image segmentation in the last decades include scene parsing, panoptic segmentation, and, more recently, new panoptic architectures. However, such panoptic…

计算机视觉与模式识别 · 计算机科学 2023-01-02 Jitesh Jain , Jiachen Li , MangTik Chiu , Ali Hassani , Nikita Orlov , Humphrey Shi

Unsupervised panoptic segmentation aims to partition an image into semantically meaningful regions and distinct object instances without training on manually annotated data. In contrast to prior work on unsupervised panoptic scene…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Oliver Hahn , Christoph Reich , Nikita Araslanov , Daniel Cremers , Christian Rupprecht , Stefan Roth

Recently, the first foundation model developed specifically for image segmentation tasks was developed, termed the "Segment Anything Model" (SAM). SAM can segment objects in input imagery based on cheap input prompts, such as one (or more)…

计算机视觉与模式识别 · 计算机科学 2023-11-10 Simiao Ren , Francesco Luzi , Saad Lahrichi , Kaleb Kassaw , Leslie M. Collins , Kyle Bradbury , Jordan M. Malof

Wound image segmentation is a critical component for the clinical diagnosis and in-time treatment of wounds. Recently, deep learning has become the mainstream methodology for wound image segmentation. However, the pre-processing of the…

图像与视频处理 · 电气工程与系统科学 2022-07-13 Honghui Liu , Changjian Wang , Kele Xu , Fangzhao Li , Ming Feng , Yuxing Peng , Hongjun He

Photoacoustic tomography (PAT) offers optical contrast, whereas magnetic resonance imaging (MRI) excels in imaging soft tissue and organ anatomy. The fusion of PAT with MRI holds promising application prospects due to their complementary…

图像与视频处理 · 电气工程与系统科学 2025-03-20 Yutian Zhong , Jinchuan He , Zhichao Liang , Shuangyang Zhang , Qianjin Feng , Lijun Lu , Li Qi

Unsupervised video-based surgical instrument segmentation has the potential to accelerate the adoption of robot-assisted procedures by reducing the reliance on manual annotations. However, the generally low quality of optical flow in…

计算机视觉与模式识别 · 计算机科学 2025-03-27 Yang Liu , Peiran Wu , Jiayu Huo , Gongyu Zhang , Zhen Yuan , Christos Bergeles , Rachel Sparks , Prokar Dasgupta , Alejandro Granados , Sebastien Ourselin

We present a novel end-to-end single-shot method that segments countable object instances (things) as well as background regions (stuff) into a non-overlapping panoptic segmentation at almost video frame rate. Current state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2020-09-01 Mark Weber , Jonathon Luiten , Bastian Leibe

In this paper, we propose PanoViT, a panorama vision transformer to estimate the room layout from a single panoramic image. Compared to CNN models, our PanoViT is more proficient in learning global information from the panoramic image for…

计算机视觉与模式识别 · 计算机科学 2022-12-26 Weichao Shen , Yuan Dong , Zonghao Chen , Zhengyi Zhao , Yang Gao , Zhu Liu

Accurate 3D scene representation and panoptic understanding are essential for applications such as virtual reality, robotics, and autonomous driving. However, challenges persist with existing methods, including precise 2D-to-3D mapping,…

计算机视觉与模式识别 · 计算机科学 2024-10-08 Shenghao Li