English
Related papers

Related papers: PPS-Ctrl: Controllable Sim-to-Real Translation for…

200 papers

We propose GaussCtrl, a text-driven method to edit a 3D scene reconstructed by the 3D Gaussian Splatting (3DGS). Our method first renders a collection of images by using the 3DGS and edits them by using a pre-trained 2D diffusion model…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Jing Wu , Jia-Wang Bian , Xinghui Li , Guangrun Wang , Ian Reid , Philip Torr , Victor Adrian Prisacariu

The field of image synthesis has made tremendous strides forward in the last years. Besides defining the desired output image with text-prompts, an intuitive approach is to additionally use spatial guidance in form of an image, such as a…

Computer Vision and Pattern Recognition · Computer Science 2024-08-13 Denis Zavadski , Johann-Friedrich Feiden , Carsten Rother

Real-time estimation of actual object depth is an essential module for various autonomous system tasks such as 3D reconstruction, scene understanding and condition assessment. During the last decade of machine learning, extensive deployment…

Computer Vision and Pattern Recognition · Computer Science 2022-06-09 Christoph Angermann , Matthias Schwab , Markus Haltmeier , Christian Laubichler , Steinbjörn Jónsson

Monocular depth estimation is a challenging task in complex compositions depicting multiple objects of diverse scales. Albeit the recent great progress thanks to the deep convolutional neural networks (CNNs), the state-of-the-art monocular…

Computer Vision and Pattern Recognition · Computer Science 2017-08-09 Bo Li , Yuchao Dai , Mingyi He

Convolutional Neural Networks (CNNs) are propelling advances in a range of different computer vision tasks such as object detection and object segmentation. Their success has motivated research in applications of such models for medical…

Computer Vision and Pattern Recognition · Computer Science 2020-10-19 Kristoffer Wickstrøm , Michael Kampffmeyer , Robert Jenssen

Synthesizing high-quality images from low-field MRI holds significant potential. Low-field MRI is cheaper, more accessible, and safer, but suffers from low resolution and poor signal-to-noise ratio. This synthesis process can reduce…

Computer Vision and Pattern Recognition · Computer Science 2025-10-16 Zhenxuan Zhang , Peiyuan Jing , Zi Wang , Ula Briski , Coraline Beitone , Yue Yang , Yinzhe Wu , Fanwen Wang , Liutao Yang , Jiahao Huang , Zhifan Gao , Zhaolin Chen , Kh Tohidul Islam , Guang Yang , Peter J. Lally

Text-conditioned generative models for volumetric medical imaging provide semantic control but lack explicit anatomical guidance, often resulting in outputs that are spatially ambiguous or anatomically inconsistent. In contrast,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Daniele Molino , Camillo Maria Caruso , Paolo Soda , Valerio Guarrasi

End-to-end deep-learning networks recently demonstrated extremely good perfor- mance for stereo matching. However, existing networks are difficult to use for practical applications since (1) they are memory-hungry and unable to process even…

Computer Vision and Pattern Recognition · Computer Science 2018-07-17 Stepan Tulyakov , Anton Ivanov , Francois Fleuret

In medical imaging, chromosome straightening plays a significant role in the pathological study of chromosomes and in the development of cytogenetic maps. Whereas different approaches exist for the straightening task, typically geometric…

Computer Vision and Pattern Recognition · Computer Science 2021-10-20 Sifan Song , Daiyun Huang , Yalun Hu , Chunxiao Yang , Jia Meng , Fei Ma , Frans Coenen , Jiaming Zhang , Jionglong Su

Photon-Counting Computed Tomography (PCCT) is a novel imaging modality that simultaneously acquires volumetric data at multiple X-ray energy levels, generating separate volumes that capture energy-dependent attenuation properties.…

Human-Computer Interaction · Computer Science 2025-08-21 Mohit Sharma , Emma Nilsson , Martin Falk , Talha Bin Masood , Lee Jollans , Anders Persson , Tino Ebbers , Ingrid Hotz

This paper addresses the problem of photometric stereo for non-Lambertian surfaces. Existing approaches often adopt simplified reflectance models to make the problem more tractable, but this greatly hinders their applications on real-world…

Computer Vision and Pattern Recognition · Computer Science 2018-07-24 Guanying Chen , Kai Han , Kwan-Yee K. Wong

Large-scale synthetic datasets are beneficial to stereo matching but usually introduce known domain bias. Although unsupervised image-to-image translation networks represented by CycleGAN show great potential in dealing with domain gap, it…

Computer Vision and Pattern Recognition · Computer Science 2020-05-06 Rui Liu , Chengxi Yang , Wenxiu Sun , Xiaogang Wang , Hongsheng Li

Neural implicit scene representations have recently shown encouraging results in dense visual SLAM. However, existing methods produce low-quality scene reconstruction and low-accuracy localization performance when scaling up to large indoor…

Computer Vision and Pattern Recognition · Computer Science 2025-05-28 Tianchen Deng , Guole Shen , Tong Qin , Jianyu Wang , Wentao Zhao , Jingchuan Wang , Danwei Wang , Weidong Chen

Commercial depth sensors usually generate noisy and missing depths, especially on specular and transparent objects, which poses critical issues to downstream depth or point cloud-based tasks. To mitigate this problem, we propose a powerful…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Qiyu Dai , Jiyao Zhang , Qiwei Li , Tianhao Wu , Hao Dong , Ziyuan Liu , Ping Tan , He Wang

Capturing and labeling camera images in the real world is an expensive task, whereas synthesizing labeled images in a simulation environment is easy for collecting large-scale image data. However, learning from only synthetic images may not…

Computer Vision and Pattern Recognition · Computer Science 2018-07-06 Tadanobu Inoue , Subhajit Chaudhury , Giovanni De Magistris , Sakyasingha Dasgupta

In the last few years, large improvements in image clustering have been driven by the recent advances in deep learning. However, due to the architectural complexity of deep neural networks, there is no mathematical theory that explains the…

Computer Vision and Pattern Recognition · Computer Science 2021-08-30 Angel Villar-Corrales , Veniamin I. Morgenshtern

Cardiac contraction is a rapid, coordinated process that unfolds across three-dimensional tissue on millisecond timescales. Traditional optical imaging is often inadequate for capturing dynamic cellular structure in the beating heart…

Computer Vision and Pattern Recognition · Computer Science 2025-11-06 Yi Gong , Xinyuan Zhang , Jichen Chai , Yichen Ding , Yifei Lou

Pretraining robust vision or multimodal foundation models (e.g., CLIP) relies on large-scale datasets that may be noisy, potentially misaligned, and have long-tail distributions. Previous works have shown promising results in augmenting…

Computer Vision and Pattern Recognition · Computer Science 2024-10-17 Qingqing Cao , Mahyar Najibi , Sachin Mehta

The perception of transparent objects for grasp and manipulation remains a major challenge, because existing robotic grasp methods which heavily rely on depth maps are not suitable for transparent objects due to their unique visual…

Computer Vision and Pattern Recognition · Computer Science 2024-05-27 Yifan Zhou , Wanli Peng , Zhongyu Yang , He Liu , Yi Sun

Pansharpening is a crucial task in remote sensing, enabling the generation of high-resolution multispectral images by fusing low-resolution multispectral data with high-resolution panchromatic images. This paper provides a comprehensive…

Image and Video Processing · Electrical Eng. & Systems 2024-12-09 Mahek Kantharia , Neeraj Badal , Zankhana Shah