中文
相关论文

相关论文: RUBIK: A Structured Benchmark for Image Matching a…

200 篇论文

Benchmarking optimization algorithms is fundamental for the advancement of computational intelligence. However, widely adopted artificial test suites exhibit limited correspondence with the diversity and complexity of real-world engineering…

计算工程、金融与科学 · 计算机科学 2026-04-17 Stefan Ivić , Siniša Družeta , Luka Grbčić

As AI-generated images proliferate across digital platforms, reliable detection methods have become critical for combating misinformation and maintaining content authenticity. While numerous deepfake detection methods have been proposed,…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Simiao Ren , Yuchen Zhou , Xingyu Shen , Kidus Zewde , Tommy Duong , George Huang , Hatsanai , Tiangratanakul , Tsang , Ng , En Wei , Jiayu Xue

Accurate camera calibration is a well-known and widely used task in computer vision that has been researched for decades. However, the standard approach based on checkerboard calibration patterns has some drawbacks that limit its…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Peer Stelldinger , Nils Schönherr , Justus Biermann

The conventional methods for estimating camera poses and scene structures from severely blurry or low resolution images often result in failure. The off-the-shelf deblurring or super-resolution methods may show visually pleasing results.…

计算机视觉与模式识别 · 计算机科学 2017-09-19 Haesol Park , Kyoung Mu Lee

With the advantage of high mobility, Unmanned Aerial Vehicles (UAVs) are used to fuel numerous important applications in computer vision, delivering more efficiency and convenience than surveillance cameras with fixed camera angle, scale…

计算机视觉与模式识别 · 计算机科学 2018-04-03 Dawei Du , Yuankai Qi , Hongyang Yu , Yifan Yang , Kaiwen Duan , Guorong Li , Weigang Zhang , Qingming Huang , Qi Tian

Benchmarks for robot manipulation are crucial to measuring progress in the field, yet there are few benchmarks that demonstrate critical manipulation skills, possess standardized metrics, and can be attempted by a wide array of robot…

机器人学 · 计算机科学 2022-02-16 Boling Yang , Patrick E. Lancaster , Siddhartha S. Srinivasa , Joshua R. Smith

We present a comprehensive study and evaluation of existing single image compression artifacts removal algorithms, using a new 4K resolution benchmark including diversified foreground objects and background scenes with rich structures,…

图像与视频处理 · 电气工程与系统科学 2020-08-26 Jiaying Liu , Dong Liu , Wenhan Yang , Sifeng Xia , Xiaoshuai Zhang , Yuanying Dai

We propose a new benchmark evaluating the performance of multimodal large language models on rebus puzzles. The dataset covers 333 original examples of image-based wordplay, cluing 13 categories such as movies, composers, major cities, and…

Research on automated image enhancement has gained momentum in recent years, partially due to the need for easy-to-use tools for enhancing pictures captured by ubiquitous cameras on mobile devices. Many of the existing leading methods…

计算机视觉与模式识别 · 计算机科学 2017-04-06 Parag S. Chandakkar , Baoxin Li

Data-efficient image classification using deep neural networks in settings, where only small amounts of labeled data are available, has been an active research area in the recent past. However, an objective comparison between published…

计算机视觉与模式识别 · 计算机科学 2021-08-31 Lorenzo Brigato , Björn Barz , Luca Iocchi , Joachim Denzler

Composed Image Retrieval (CIR) has made significant progress, yet current benchmarks are limited to single ground-truth answers and lack the annotations needed to evaluate false positive avoidance, robustness and multi-image reasoning. We…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Rohan Mahadev , Joyce Yuan , Patrick Poirson , David Xue , Hao-Yu Wu , Dmitry Kislyuk

Camera pose tracking attracts much interest both from academic and industrial communities, of which the methods based on planar markers are easy to be implemented. However, most of the existing methods need to identify multiple points in…

计算机视觉与模式识别 · 计算机科学 2019-07-25 Fulin Tang , Yihong Wu

This letter presents a novel method to estimate the relative poses between RGB-D cameras with minimal overlapping fields of view in a panoramic RGB-D camera system. This calibration problem is relevant to applications such as indoor 3D…

图像与视频处理 · 电气工程与系统科学 2018-09-11 Hang Liu , Hengyu Li , Xiahua Liu , Jun Luo , Shaorong Xie , Yu Sun

Text-image composed retrieval aims to retrieve the target image through the composed query, which is specified in the form of an image plus some text that describes desired modifications to the input image. It has recently attracted…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Shitong Sun , Jindong Gu , Shaogang Gong

We introduce Cube Bench, a Rubik's-cube benchmark for evaluating spatial and sequential reasoning in multimodal large language models (MLLMs). The benchmark decomposes performance into five skills: (i) reconstructing cube faces from images…

计算与语言 · 计算机科学 2025-12-24 Dhruv Anand , Ehsan Shareghi

In recent years there has been significant improvement in the capability of Visual Place Recognition (VPR) methods, building on the success of both hand-crafted and learnt visual features, temporal filtering and usage of semantic scene…

计算机视觉与模式识别 · 计算机科学 2019-05-01 Mubariz Zaffar , Ahmad Khaliq , Shoaib Ehsan , Michael Milford , Klaus McDonald-Maier

Recent advancements in bird's eye view (BEV) representations have shown remarkable promise for in-vehicle 3D perception. However, while these methods have achieved impressive results on standard benchmarks, their robustness in varied…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Shaoyuan Xie , Lingdong Kong , Wenwei Zhang , Jiawei Ren , Liang Pan , Kai Chen , Ziwei Liu

Single-view RGB model-based object pose estimation methods achieve strong generalization but are fundamentally limited by depth ambiguity, clutter, and occlusions. Multi-view pose estimation methods have the potential to solve these issues,…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Anna Šárová Mikeštíková , Médéric Fourmy , Martin Cífka , Josef Sivic , Vladimir Petrik

Visible images offer rich texture details, while infrared images emphasize salient targets. Fusing these complementary modalities enhances scene understanding, particularly for advanced vision tasks under challenging conditions. Recently,…

计算机视觉与模式识别 · 计算机科学 2025-06-10 Beining Xu , Junxian Li

Background: Pose estimation of rigid objects is a practical challenge in optical metrology and computer vision. This paper presents a novel stochastic-geometrical modeling framework for object pose estimation based on observing multiple…

计算机视觉与模式识别 · 计算机科学 2024-06-04 Wolfgang Hoegele