English
Related papers

Related papers: DINOv3 Visual Representations for Blueberry Percep…

200 papers

Fruit harvesting poses a significant labor and financial burden for the industry, highlighting the critical need for advancements in robotic harvesting solutions. Machine vision-based fruit detection has been recognized as a crucial…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Jiajia Li , Kyle Lammers , Xunyuan Yin , Xiang Yin , Long He , Renfu Lu , Zhaojian Li

The ability to accurately represent and localise relevant objects is essential for robots to carry out tasks effectively. Traditional approaches, where robots simply capture an image, process that image to take an action, and then forget…

Robotics · Computer Science 2023-11-29 David Rapado Rincon , Eldert J. van Henten , Gert Kootstra

This study presents a vision-guided robotic control system for automated fruit tree pruning applications. Traditional pruning practices are labor-intensive and limit agricultural efficiency and scalability, highlighting the need for…

Robotics · Computer Science 2025-06-09 Dawood Ahmed , Basit Muhammad Imran , Martin Churuvija , Manoj Karkee

Agricultural domains are being transformed by recent advances in AI and computer vision that support quantitative visual evaluation. Using aerial and ground imaging over a time series, we develop a framework for characterizing the ripening…

Computer Vision and Pattern Recognition · Computer Science 2024-12-16 Faith Johnson , Ryan Meegan , Jack Lowry , Peter Oudemans , Kristin Dana

3D Move To See (3DMTS) is a mutli-perspective visual servoing method for unstructured and occluded environments, like that encountered in robotic crop harvesting. This paper presents a deep learning method, Deep-3DMTS for creating a…

Robotics · Computer Science 2019-08-07 Paul Zapotezny-Anderson , Chris Lehnert

We present a zero-shot segmentation approach for agricultural imagery that leverages Plantnet, a large-scale plant classification model, in conjunction with its DinoV2 backbone and the Segment Anything Model (SAM). Rather than collecting…

Computer Vision and Pattern Recognition · Computer Science 2025-10-15 Simon Ravé , Jean-Christophe Lombardo , Pejman Rasti , Alexis Joly , David Rousseau

Blueberry detection in natural environments remains challenging due to variable lighting, occlusions, and motion blur due to environmental factors and imaging devices. Deep learning-based object detectors promise to address these…

Computer Vision and Pattern Recognition · Computer Science 2025-10-07 Xinyang Mu , Yuzhen Lu , Boyang Deng

This paper addresses the challenge of developing a multi-arm quadrupedal robot capable of efficiently harvesting fruit in complex, natural environments. To overcome the inherent limitations of traditional bimanual manipulation, we introduce…

Robotics · Computer Science 2025-05-14 Zhichao Liu , Jingzong Zhou , Konstantinos Karydis

Detecting diverse objects within complex indoor 3D point clouds presents significant challenges for robotic perception, particularly with varied object shapes, clutter, and the co-existence of static and dynamic elements where traditional…

Robotics · Computer Science 2025-07-24 Haichuan Li , Changda Tian , Panos Trahanias , Tomi Westerlund

Fruit ripeness estimation models have for decades depended on spectral index features or colour-based features, such as mean, standard deviation, skewness, colour moments, and/or histograms for learning traits of fruit ripeness. Recently,…

Computer Vision and Pattern Recognition · Computer Science 2024-06-03 Chollette C. Olisah , Ben Trewhella , Bo Li , Melvyn L. Smith , Benjamin Winstone , E. Charles Whitfield , Felicidad Fernández Fernández , Harriet Duncalfe

Purpose: Depth estimation in robotic surgery is vital in 3D reconstruction, surgical navigation and augmented reality visualization. Although the foundation model exhibits outstanding performance in many vision tasks, including depth…

Computer Vision and Pattern Recognition · Computer Science 2024-01-15 Beilei Cui , Mobarakol Islam , Long Bai , Hongliang Ren

Anomaly detection and classification in medical imaging are critical for early diagnosis but remain challenging due to limited annotated data, class imbalance, and the high cost of expert labeling. Emerging vision foundation models such as…

Image and Video Processing · Electrical Eng. & Systems 2025-09-17 Fazle Rafsani , Jay Shah , Catherine D. Chong , Todd J. Schwedt , Teresa Wu

Multiple Object Tracking (MOT) is a computer vision task that has been employed in a variety of sectors. Some common limitations in MOT are varying object appearances, occlusions, or crowded scenes. To address these challenges, machine…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Niels G. Faber , Seyed Sahand Mohammadi Ziabari , Fatemeh Karimi Nejadasl

Bird's-Eye-View (BEV) representation has emerged as a mainstream paradigm for multi-view 3D object detection, demonstrating impressive perceptual capabilities. However, existing methods overlook the geometric quality of BEV representation,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Jinqing Zhang , Yanan Zhang , Yunlong Qi , Zehua Fu , Qingjie Liu , Yunhong Wang

Training real-world neural network models to achieve high performance and generalizability typically requires a substantial amount of labeled data, spanning a broad range of variation. This data-labeling process can be both labor and cost…

Computer Vision and Pattern Recognition · Computer Science 2021-08-31 Zhenghao Fei , Alex Olenskyj , Brian N. Bailey , Mason Earles

Benchmarking 3D spatial understanding of foundation models is essential for real-world applications such as robotics and autonomous driving. Existing evaluations often rely on downstream fine-tuning with linear heads or task-specific…

Computer Vision and Pattern Recognition · Computer Science 2026-01-19 Valentina Lilova , Toyesh Chakravorty , Julian I. Bibo , Emma Boccaletti , Brandon Li , Lívia Baxová , Cees G. M. Snoek , Mohammadreza Salehi

This paper presents a comprehensive review of ground agricultural robotic systems and applications with special focus on harvesting that span research and commercial products and results, as well as their enabling technologies. The majority…

Despite the remarkable success of large-scale pre-trained image representation models (i.e., vision encoders) across various vision tasks, they are predominantly trained on 2D image data and therefore often fail to capture 3D spatial…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Byungwoo Jeon , Dongyoung Kim , Huiwon Jang , Insoo Kim , Jinwoo Shin

Deep learning-based automatic medical image segmentation plays a critical role in clinical diagnosis and treatment planning but remains challenging in few-shot scenarios due to the scarcity of annotated training data. Recently,…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Guoping Xu , Jayaram K. Udupa , Weiguo Lu , You Zhang

Accurate measurement of eyelid parameters such as Margin Reflex Distances (MRD1, MRD2) and Levator Function (LF) is critical in oculoplastic diagnostics but remains limited by manual, inconsistent methods. This study evaluates deep learning…

Machine Learning · Computer Science 2025-04-02 Chun-Hung Chen