中文
相关论文

相关论文: FSMC-Pose: Frequency and Spatial Fusion with Multi…

200 篇论文

Place recognition is a challenging task in computer vision, crucial for enabling autonomous vehicles and robots to navigate previously visited environments. While significant progress has been made in learnable multimodal methods that…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Alexander Melekhin , Dmitry Yudin , Ilia Petryashin , Vitaly Bezuglyj

Monocular depth foundation models generalize well across scenes, yet they are typically optimized with uniform pixel-wise objectives that do not distinguish user-specified or task-relevant target regions from the surrounding context. We…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Yuxin Du , Tao Lin , Zile Zhong , Runting Li , Xiyao Chen , Jiting Liu , Chenglin Liu , Ying-Cong Chen , Yuqian Fu , Bo Zhao

Accurate body dimension and weight measurements are critical for optimizing poultry management, health assessment, and economic efficiency. This study introduces an innovative deep learning-based model leveraging multimodal data-2D RGB…

计算机视觉与模式识别 · 计算机科学 2025-05-12 Wenbo Xiao , Qiannan Han , Gang Shu , Guiping Liang , Hongyan Zhang , Song Wang , Zhihao Xu , Weican Wan , Chuang Li , Guitao Jiang , Yi Xiao

Cattle farming is one of the important and profitable agricultural industries. Employing intelligent automated precision livestock farming systems that can count animals, track the animals and their poses will raise productivity and…

计算机视觉与模式识别 · 计算机科学 2023-12-15 Kian Eng Ong , Sivaji Retta , Ramarajulu Srinivasan , Shawn Tan , Jun Liu

Accurate remote sensing-based crop yield prediction remains a fundamental challenging task due to complex spatial patterns, heterogeneous spectral characteristics, and dynamic agricultural conditions. Existing methods often suffer from…

计算机视觉与模式识别 · 计算机科学 2025-07-09 Juli Zhang , Zeyu Yan , Jing Zhang , Qiguang Miao , Quan Wang

Multi-source stationary computed tomography (MSS-CT) offers significant advantages in medical and industrial applications due to its gantry-less scan architecture and/or capability of simultaneous multi-source emission. However, the lack of…

医学物理 · 物理学 2025-01-20 Yingxian Xia , Zhiqiang Chen , Li Zhang , Yuxiang Xing , Hewei Gao

Image deblurring aims to reconstruct a latent sharp image from its corresponding blurred one. Although existing methods have achieved good performance, most of them operate exclusively in either the spatial domain or the frequency domain,…

计算机视觉与模式识别 · 计算机科学 2025-02-21 Hu Gao , Depeng Dang

We present a contact-based phenotyping robot platform that can autonomously insert nitrate sensors into cornstalks to proactively monitor macronutrient levels in crops. This task is challenging because inserting such sensors requires…

机器人学 · 计算机科学 2023-11-08 Moonyoung Lee , Aaron Berger , Dominic Guri , Kevin Zhang , Lisa Coffee , George Kantor , Oliver Kroemer

In recent years, human pose estimation has made significant progress through the implementation of deep learning techniques. However, these techniques still face limitations when confronted with challenging scenarios, including occlusion,…

计算机视觉与模式识别 · 计算机科学 2023-11-10 Sihan Gao , Jing Zhu , Xiaoxuan Zhuang , Zhaoyue Wang , Qijin Li

Recent subject-driven image customization excels in fidelity, yet fine-grained instance-level spatial control remains an elusive challenge, hindering real-world applications. This limitation stems from two factors: a scarcity of scalable,…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Junjie Hu , Tianyang Han , Kai Ma , Jialin Gao , Song Yang , Xianhua He , Junfeng Luo , Xiaoming Wei , Wenqiang Zhang

Wheat plays a critical role in global food security, making it one of the most extensively studied crops. Accurate identification and measurement of key characteristics of wheat heads are essential for breeders to select varieties for…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Yasashwini Sai Gowri P , Karthik Seemakurthy , Andrews Agyemang Opoku , Sita Devi Bharatula

Convolutional Neural Networks (CNNs) have drawn researchers' attention to identifying cattle using muzzle images. However, CNNs often fail to capture long-range dependencies within the complex patterns of the muzzle. The transformers handle…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Rabin Dulal , Lihong Zheng , Muhammad Ashad Kabir

Accurate and scalable quantification of animal pose and appearance is crucial for studying behavior. Current 3D pose estimation techniques, such as keypoint- and mesh-based techniques, often face challenges including limited…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Jack Goffinet , Youngjo Min , Carlo Tomasi , David E. Carlson

With the rapid advancement of real-time deepfake generation techniques, forged content is becoming increasingly realistic and widespread across applications like video conferencing and social media. Although state-of-the-art detectors…

计算机视觉与模式识别 · 计算机科学 2025-08-29 Libo Lv , Tianyi Wang , Mengxiao Huang , Ruixia Liu , Yinglong Wang

The stability and reliability of wireless data transmission in vehicular networks face significant challenges due to the high dynamics of path loss caused by the complexity of rapidly changing environments. This paper proposes a multi-modal…

信号处理 · 电气工程与系统科学 2024-12-11 Kai Wang , Li Yu , Jianhua Zhang , Yixuan Tian , Eryu Guo , Guangyi Liu

This study introduces RicEns-Net, a novel Deep Ensemble model designed to predict crop yields by integrating diverse data sources through multimodal data fusion techniques. The research focuses specifically on the use of synthetic aperture…

图像与视频处理 · 电气工程与系统科学 2025-02-11 Akshay Dagadu Yewle , Laman Mirzayeva , Oktay Karakuş

Category-level object pose estimation, which predicts the pose of objects within a known category without prior knowledge of individual instances, is essential in applications like warehouse automation and manufacturing. Existing methods…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Yifan Yang , Peili Song , Enfan Lan , Dong Liu , Jingtai Liu

Multi-modal collaborative perception calls for great attention to enhancing the safety of autonomous driving. However, current multi-modal approaches remain a ``local fusion to communication'' sequence, which fuses multi-modal data locally…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Kang Yang , Peng Wang , Lantao Li , Tianci Bu , Chen Sun , Deying Li , Yongcai Wang

We study multi-dataset training (MDT) for pose estimation, where skeletal heterogeneity presents a unique challenge that existing methods have yet to address. In traditional domains, \eg regression and classification, MDT typically relies…

计算机视觉与模式识别 · 计算机科学 2025-05-26 Uyoung Jeong , Jonathan Freer , Seungryul Baek , Hyung Jin Chang , Kwang In Kim

Parameter-efficient fine-tuning (PEFT) in multimodal tracking reveals a concerning trend where recent performance gains are often achieved at the cost of inflated parameter budgets, which fundamentally erodes PEFT's efficiency promise. In…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Junbin Su , Ziteng Xue , Shihui Zhang , Kun Chen , Weiming Hu , Zhipeng Zhang