中文
相关论文

相关论文: The First Controllable Bokeh Rendering Challenge a…

200 篇论文

Introduction: We describe the foundation of PETRIC, an image reconstruction challenge to minimise the computational runtime of related algorithms for Positron Emission Tomography (PET). Purpose: Although several similar challenges are…

Tactile representation learning (TRL) equips robots with the ability to leverage touch information, boosting performance in tasks such as environment perception and object manipulation. However, the heterogeneity of tactile sensors results…

机器人学 · 计算机科学 2023-05-02 Ben Zandonati , Ruohan Wang , Ruihan Gao , Yan Wu

In the current landscape of biometrics and surveillance, the ability to accurately recognize faces in uncontrolled settings is paramount. The Watchlist Challenge addresses this critical need by focusing on face detection and open-set…

In this paper, we target the adaptive source driven 3D scene editing task by proposing a CustomNeRF model that unifies a text description or a reference image as the editing prompt. However, obtaining desired editing results conformed with…

计算机视觉与模式识别 · 计算机科学 2023-12-05 Runze He , Shaofei Huang , Xuecheng Nie , Tianrui Hui , Luoqi Liu , Jiao Dai , Jizhong Han , Guanbin Li , Si Liu

Relighting, which synthesizes a novel view under a given lighting condition (unseen in training time), is a must feature for immersive photo-realistic experience. However, real-time relighting is challenging due to high computation cost of…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Euntae Choi , Vincent Carpentier , Seunghun Shin , Sungjoo Yoo

Iteration of training and evaluating a machine learning model is an important process to improve its performance. However, while teachable interfaces enable blind users to train and test an object recognizer with photos taken in their…

Object detection is a comprehensively studied problem in autonomous driving. However, it has been relatively less explored in the case of fisheye cameras. The strong radial distortion breaks the translation invariance inductive bias of…

计算机视觉与模式识别 · 计算机科学 2022-06-28 Saravanabalagi Ramachandran , Ganesh Sistu , Varun Ravi Kumar , John McDonald , Senthil Yogamani

Sketching is a powerful artistic technique for capturing essential visual information about real-world objects and has increasingly attracted attention in image synthesis research. However, the field lacks a unified benchmark to evaluate…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Xingyue Lin , Xingjian Hu , Shuai Peng , Jianhua Zhu , Liangcai Gao

Cognitive training for sustained attention and working memory is vital across domains relying on robust mental capacity such as education or rehabilitation. Adaptive systems are essential, dynamically matching difficulty to user ability to…

人机交互 · 计算机科学 2026-02-19 Dominik Szczepaniak , Monika Harvey , Fani Deligianni

We introduce INQUIRE, a text-to-image retrieval benchmark designed to challenge multimodal vision-language models on expert-level queries. INQUIRE includes iNaturalist 2024 (iNat24), a new dataset of five million natural world images, along…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Edward Vendrow , Omiros Pantazis , Alexander Shepard , Gabriel Brostow , Kate E. Jones , Oisin Mac Aodha , Sara Beery , Grant Van Horn

This paper reviews the AIM 2019 challenge on real world super-resolution. It focuses on the participating methods and final results. The challenge addresses the real world setting, where paired true high and low-resolution images are…

Standard datasets often present limitations, particularly due to the fixed nature of input data sensors, which makes it difficult to compare methods that actively adjust sensor parameters to suit environmental conditions. This is the case…

机器人学 · 计算机科学 2025-06-24 Olivier Gamache , Jean-Michel Fortin , Matěj Boxan , François Pomerleau , Philippe Giguère

We consider the problem of obtaining image quality representations in a self-supervised manner. We use prediction of distortion type and degree as an auxiliary task to learn features from an unlabeled image dataset containing a mixture of…

计算机视觉与模式识别 · 计算机科学 2022-06-30 Pavan C. Madhusudana , Neil Birkbeck , Yilin Wang , Balu Adsumilli , Alan C. Bovik

We present a three-stage progressive shadow-removal pipeline for the CVPR2026 NTIRE WSRD+ challenge. Built on OmniSR, our method treats deshadowing as iterative direct refinement, where later stages correct residual artefacts left by…

计算机视觉与模式识别 · 计算机科学 2026-04-22 Lorenzo Beltrame , Jules Salzinger , Filip Svoboda , Jasmin Lampert , Phillipp Fanta-Jende , Radu Timofte , Marco Körner

Existing neural human rendering methods struggle with a single image input due to the lack of information in invisible areas and the depth ambiguity of pixels in visible areas. In this regard, we propose Monocular Neural Human Renderer…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Hongsuk Choi , Gyeongsik Moon , Matthieu Armando , Vincent Leroy , Kyoung Mu Lee , Gregory Rogez

We consider the problem of realistic bokeh rendering from a single all-in-focus image. Bokeh rendering mimics aesthetic shallow depth-of-field (DoF) in professional photography, but these visual effects generated by existing methods suffer…

计算机视觉与模式识别 · 计算机科学 2023-06-09 Xianrui Luo , Juewen Peng , Ke Xian , Zijin Wu , Zhiguo Cao

Professional photo editing remains challenging, requiring extensive knowledge of imaging pipelines and significant expertise. While recent deep learning approaches, particularly style transfer methods, have attempted to automate this…

图像与视频处理 · 电气工程与系统科学 2025-12-11 Omar Elezabi , Marcos V. Conde , Zongwei Wu , Radu Timofte

The ninth AI City Challenge continues to advance real-world applications of computer vision and AI in transportation, industrial automation, and public safety. The 2025 edition featured four tracks and saw a 17% increase in participation,…

This paper presents the report of the URVIS 2026 challenge on adverse-to-extreme panoptic segmentation. As the first challenge of its kind, it attracted 17 registered participants and 47 submissions, with 4 teams reaching the final phase.…