中文
相关论文

相关论文: MOO: A Multi-view Oriented Observations Dataset fo…

200 篇论文

In recent years, language-guided open-set aerial object detection has gained significant attention due to its better alignment with real-world application needs. However, due to limited datasets, most existing language-guided methods…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Guoting Wei , Yu Liu , Xia Yuan , Xizhe Xue , Linlin Guo , Yifan Yang , Chunxia Zhao , Zongwen Bai , Haokui Zhang , Rong Xiao

We present DogMo, a large-scale multi-view RGB-D video dataset capturing diverse canine movements for the task of motion recovery from images. DogMo comprises 1.2k motion sequences collected from 10 unique dogs, offering rich variation in…

计算机视觉与模式识别 · 计算机科学 2025-10-29 Zan Wang , Siyu Chen , Luya Mo , Xinfeng Gao , Yuxin Shen , Lebin Ding , Wei Liang

Re-identification (re-ID) is currently investigated as a closed-world image retrieval task, and evaluated by retrieval based metrics. The algorithms return ranking lists to users, but cannot tell which images are the true target. In…

计算机视觉与模式识别 · 计算机科学 2020-11-24 Zheng Wang , Xin Yuan , Toshihiko Yamasaki , Yutian Lin , Xin Xu , Wenjun Zeng

Wildlife ReID involves utilizing visual technology to identify specific individuals of wild animals in different scenarios, holding significant importance for wildlife conservation, ecological research, and environmental monitoring.…

计算机视觉与模式识别 · 计算机科学 2024-10-28 Chenyue Li , Shuoyi Chen , Mang Ye

Person re-identification (Re-ID) aims to match the same pedestrian in a large gallery with different cameras and views. Enhancing the robustness of the extracted feature representations is a main challenge in Re-ID. Existing methods usually…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Chao Yuan , Tianyi Zhang , Guanglin Niu

Detecting faces in overhead images remains a significant challenge due to extreme scale variations and environmental clutter. To address this, we created the BirdsEye-RU dataset, a comprehensive collection of 2,978 images containing over…

We present the HOH (Human-Object-Human) Handover Dataset, a large object count dataset with 136 objects, to accelerate data-driven research on handover studies, human-robot handover implementation, and artificial intelligence (AI) on…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Noah Wiederhold , Ava Megyeri , DiMaggio Paris , Sean Banerjee , Natasha Kholgade Banerjee

Vision Transformers (ViTs) have excelled in vehicle re-identification (ReID) tasks. However, non-square aspect ratios of image or video input might significantly affect the re-identification performance. To address this issue, we propose a…

计算机视觉与模式识别 · 计算机科学 2024-07-11 Mei Qiu , Lauren Christopher , Lingxi Li

For people affected by blindness and low vision (BLV), safe and independent navigation remains a major challenge, impacting over 2.2 billion individuals worldwide. Although multimodal large language models (MLLMs) offer new opportunities…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Junhyeok Kim , Jaewoo Park , Junhee Park , Sangeyl Lee , Jiwan Chung , Jisung Kim , Ji Hoon Joung , Youngjae Yu

Camera-based animal re-identification (Animal Re-ID) can support wildlife monitoring and precision livestock management in large outdoor environments with limited wireless connectivity. In these settings, inference must run directly on…

计算机视觉与模式识别 · 计算机科学 2025-12-10 Yubo Chen , Di Zhao , Yun Sing Koh , Talia Xu

In this paper, we studied the problem of localizing a generic set of keypoints across multiple quadruped or four-legged animal species from images. Due to the lack of large scale animal keypoint dataset with ground truth annotations, we…

计算机视觉与模式识别 · 计算机科学 2021-09-01 Prianka Banik , Lin Li , Xishuang Dong

We present the ALTO dataset, a vision-focused dataset for the development and benchmarking of Visual Place Recognition and Localization methods for Unmanned Aerial Vehicles. The dataset is composed of two long (approximately 150km and…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Ivan Cisneros , Peng Yin , Ji Zhang , Howie Choset , Sebastian Scherer

Applications of unmanned aerial vehicle (UAV) in logistics, agricultural automation, urban management, and emergency response are highly dependent on oriented object detection (OOD) to enhance visual perception. Although existing datasets…

计算机视觉与模式识别 · 计算机科学 2025-04-29 Kai Ye , Haidi Tang , Bowen Liu , Pingyang Dai , Liujuan Cao , Rongrong Ji

Object detection has witnessed significant progress by relying on large, manually annotated datasets. Annotating such datasets is highly time consuming and expensive, which motivates the development of weakly supervised and few-shot object…

计算机视觉与模式识别 · 计算机科学 2020-08-27 Carlo Biffi , Steven McDonagh , Philip Torr , Ales Leonardis , Sarah Parisot

We introduce YOLO11-JDE, a fast and accurate multi-object tracking (MOT) solution that combines real-time object detection with self-supervised Re-Identification (Re-ID). By incorporating a dedicated Re-ID branch into YOLO11s, our model…

计算机视觉与模式识别 · 计算机科学 2025-05-28 Iñaki Erregue , Kamal Nasrollahi , Sergio Escalera

This paper extends the popular task of multi-object tracking to multi-object tracking and segmentation (MOTS). Towards this goal, we create dense pixel-level annotations for two existing tracking datasets using a semi-automatic annotation…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Paul Voigtlaender , Michael Krause , Aljosa Osep , Jonathon Luiten , Berin Balachandar Gnana Sekar , Andreas Geiger , Bastian Leibe

The development and implementation of visual-inertial odometry (VIO) has focused on structured environments, but interest in localization in off-road environments is growing. In this paper, we present the RELLIS Off-road Odometry Analysis…

机器人学 · 计算机科学 2022-04-08 George Chustz , Srikanth Saripalli

Weeds are one of the major reasons for crop yield loss but current weeding practices fail to manage weeds in an efficient and targeted manner. Effective weed management is especially important for crops with high worldwide production such…

计算机视觉与模式识别 · 计算机科学 2025-02-19 Ekin Celikkan , Timo Kunzmann , Yertay Yeskaliyev , Sibylle Itzerott , Nadja Klein , Martin Herold

Aerial-Ground Re-Identification (AG-ReID) is constrained by the viewpoint-domain gap, as drastic viewpoint disparities occlude or distort discriminative features, making cross-viewpoint image retrieval challenging. While existing methods…

计算机视觉与模式识别 · 计算机科学 2026-04-30 William Grolleau , Astrid Sabourin , Guillaume Lapouge , Catherine Achard

In recent years, the development of deep learning approaches for the task of person re-identification led to impressive results. However, this comes with a limitation for industrial and practical real-world applications. Firstly, most of…

计算机视觉与模式识别 · 计算机科学 2024-09-09 Federico Cunico , Marco Cristani