English
Related papers

Related papers: Towards Smart Point-and-Shoot Photography

200 papers

Segment Anything Model (SAM) has gained significant recognition in the field of semantic segmentation due to its versatile capabilities and impressive performance. Despite its success, SAM faces two primary limitations: (1) it relies…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Yuchen Li , Li Zhang , Youwei Liang , Pengtao Xie

We introduce new linear mathematical formulations to calculate the focal length of a camera in an active platform. Through mathematical derivations, we show that the focal lengths in each direction can be estimated using only one point…

Computer Vision and Pattern Recognition · Computer Science 2018-06-12 Mehdi Faraji , Anup Basu

There has been a lot of recent research on improving the efficiency of fine-tuning foundation models. In this paper, we propose a novel efficient fine-tuning method that allows the input image size of Segment Anything Model (SAM) to be…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Sota Kato , Hinako Mitsuoka , Kazuhiro Hotta

Segment Anything (SAM) provides an unprecedented foundation for human segmentation, but may struggle under occlusion, where keypoints may be partially or fully invisible. We adapt SAM 2.1 for pose-guided segmentation with minimal encoder…

Computer Vision and Pattern Recognition · Computer Science 2026-01-19 Constantin Kolomiiets , Miroslav Purkrabek , Jiri Matas

Adding more cameras to SLAM systems improves robustness and accuracy but complicates the design of the visual front-end significantly. Thus, most systems in the literature are tailored for specific camera configurations. In this work, we…

Robotics · Computer Science 2021-01-01 Juichung Kuo , Manasi Muglikar , Zichao Zhang , Davide Scaramuzza

Image composition is one of the most important applications in image processing. However, the inharmonious appearance between the spliced region and background degrade the quality of the image. Thus, we address the problem of Image…

Computer Vision and Pattern Recognition · Computer Science 2020-04-22 Xiaodong Cun , Chi-Man Pun

Scene-aware Adaptive Compressive Sensing (ACS) has attracted significant interest due to its promising capability for efficient and high-fidelity acquisition of scene images. ACS typically prescribes adaptive sampling allocation (ASA) based…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Zhifu Tian , Tao Hu , Chaoyang Niu , Di Wu , Shu Wang

Image-to-point cloud registration aims to determine the relative camera pose between an RGB image and a reference point cloud, serving as a general solution for locating 3D objects from 2D observations. Matching individual points with…

Computer Vision and Pattern Recognition · Computer Science 2024-01-19 Gongxin Yao , Yixin Xuan , Yiwei Chen , Yu Pan

Absolute pose regressor (APR) networks are trained to estimate the pose of the camera given a captured image. They compute latent image representations from which the camera position and orientation are regressed. APRs provide a different…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Yoli Shavit , Yosi Keller

Surgical image segmentation is highly challenging, primarily due to scarcity of annotated data. Generalist prompted segmentation models like the Segment-Anything Model (SAM) can help tackle this task, but because they require image-specific…

Computer Vision and Pattern Recognition · Computer Science 2025-07-16 Aditya Murali , Farahdiba Zarin , Adrien Meyer , Pietro Mascagni , Didier Mutter , Nicolas Padoy

Few-shot semantic segmentation (FSS) aims to segment novel classes in query images using only a small annotated support set. While prior research has mainly focused on improving decoders, the encoder's limited ability to extract meaningful…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Pasquale De Marinis , Gennaro Vessio , Giovanna Castellano

Robust and accurate localization is an essential component for robotic navigation and autonomous driving. The use of cameras for localization with high definition map (HD Map) provides an affordable localization sensor set. Existing methods…

Computer Vision and Pattern Recognition · Computer Science 2021-07-07 Chengcheng Guo , Minjie Lin , Heyang Guo , Pengpeng Liang , Erkang Cheng

Accurate localization and 3D maps are increasingly needed for various artificial intelligence based IoT applications such as augmented reality, intelligent transportation, crowd monitoring, robotics, etc. This article proposes a novel…

Robotics · Computer Science 2021-03-23 Max Jwo Lem Lee , Li-Ta Hsu

Shape and pose estimation is a critical perception problem for a self-driving car to fully understand its surrounding environment. One fundamental challenge in solving this problem is the incomplete sensor signal (e.g., LiDAR scans),…

Robotics · Computer Science 2022-07-05 Josephine Monica , Wei-Lun Chao , Mark Campbell

Machine-learning excels in many areas with well-defined goals. However, a clear goal is usually not available in art forms, such as photography. The success of a photograph is measured by its aesthetic value, a very subjective concept. This…

Computer Vision and Pattern Recognition · Computer Science 2017-07-13 Hui Fang , Meng Zhang

Direct imaging of Earth-like exoplanets requires high contrast imaging capability and high angular resolution. Primary mirror segmentation is a key technological solution for large-aperture telescopes because it opens the path toward…

Instrumentation and Methods for Astrophysics · Physics 2016-06-10 Pierre Janin-Potiron , Patrice Martinez , Pierre Baudoz , Marcel Carbillet

Semantic 2D maps are commonly used by humans and machines for navigation purposes, whether it's walking or driving. However, these maps have limitations: they lack detail, often contain inaccuracies, and are difficult to create and…

Computer Vision and Pattern Recognition · Computer Science 2023-11-02 Paul-Edouard Sarlin , Eduard Trulls , Marc Pollefeys , Jan Hosang , Simon Lynen

We present CPO, a fast and robust algorithm that localizes a 2D panorama with respect to a 3D point cloud of a scene possibly containing changes. To robustly handle scene changes, our approach deviates from conventional feature point…

Computer Vision and Pattern Recognition · Computer Science 2024-02-05 Junho Kim , Hojun Jang , Changwoon Choi , Young Min Kim

We introduce CUPS, a novel method for learning sequence-to-sequence 3D human shapes and poses from RGB videos with uncertainty quantification. To improve on top of prior work, we develop a method to generate and score multiple hypotheses…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Harry Zhang , Luca Carlone

Perception systems, especially cameras, are the eyes of automated driving systems. Ensuring that they function reliably and robustly is therefore an important building block in the automation of vehicles. There are various approaches to…

Computer Vision and Pattern Recognition · Computer Science 2024-07-15 Philipp Rigoll , Laurenz Adolph , Lennart Ries , Eric Sax