中文
相关论文

相关论文: Delineate Anything Flow: Fast, Country-Level Field…

200 篇论文

Accurate and timely crop yield prediction is crucial for global food security and modern agricultural management. Traditional methods often lack the scalability and granularity required for precision farming. This paper introduces FARM:…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Shayan Nejadshamsi , Yuanyuan Zhang , Shadi Zaki , Brock Porth , Lysa Porth , Vahab Khoshdel

In industrial anomaly detection, model efficiency and mobile-friendliness become the primary concerns in real-world applications. Simultaneously, the impressive generalization capabilities of Segment Anything (SAM) have garnered broad…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Chenghao Li , Lei Qi , Xin Geng

Precise segmentation of Unmanned Aerial Vehicle (UAV)-captured images plays a vital role in tasks such as crop yield estimation and plant health assessment in banana plantations. By identifying and classifying planted areas, crop area can…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Ang He , Ximei Wu , Xing Xu , Jing Chen , Xiaobin Guo , Sheng Xu

This work presents Depth Anything, a highly practical solution for robust monocular depth estimation. Without pursuing novel technical modules, we aim to build a simple yet powerful foundation model dealing with any images under any…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Lihe Yang , Bingyi Kang , Zilong Huang , Xiaogang Xu , Jiashi Feng , Hengshuang Zhao

Bird image segmentation remains a challenging task in computer vision due to extreme pose diversity, complex plumage patterns, and variable lighting conditions. This paper presents a dual-pipeline framework for binary bird image…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Abhinav Munagala

Land-use and land cover (LULC) analysis is critical in remote sensing, with wide-ranging applications across diverse fields such as agriculture, utilities, and urban planning. However, automating LULC map generation using machine learning…

计算机视觉与模式识别 · 计算机科学 2024-12-18 Sparsh Pekhale , Rakshith Sathish , Sathisha Basavaraju , Divya Sharma

Motion segmentation from a single moving camera presents a significant challenge in the field of computer vision. This challenge is compounded by the unknown camera movements and the lack of depth information of the scene. While deep…

计算机视觉与模式识别 · 计算机科学 2024-06-28 Yuxiang Huang , Yuhao Chen , John Zelek

Explicitly disentangling style and content in vision models remains challenging due to their semantic overlap and the subjectivity of human perception. Existing methods propose separation through generative or discriminative objectives, but…

计算机视觉与模式识别 · 计算机科学 2025-08-06 Pingchuan Ma , Xiaopei Yang , Yusong Li , Ming Gui , Felix Krause , Johannes Schusterbauer , Björn Ommer

Data heterogeneity hinders clinical deployment of medical image analysis models, and generative data augmentation helps mitigate this issue. However, recent diffusion-based methods that synthesize image-mask pairs often ignore distribution…

图像与视频处理 · 电气工程与系统科学 2026-04-06 Jie Yang , Ziqi Ye , Aihua Ke , Jian Luo , Bo Cai , Xiaosong Wang

In this paper, we address two critical challenges in the domain of flood detection: the computational expense of large-scale time series change detection and the lack of interpretable decision-making processes on explainable AI (XAI). To…

计算机视觉与模式识别 · 计算机科学 2024-05-14 Ziyang Zhang , Plamen Angelov , Dmitry Kangin , Nicolas Longépé

Accurate and timely crop yield estimation is critical for global food security, agricultural policy, and farm management. The Copernicus Sentinel-2 satellite constellation, with high spatial, temporal, and spectral resolution, has…

图像与视频处理 · 电气工程与系统科学 2026-03-26 Mohammadreza Narimani , Alireza Pourreza , Ali Moghimi , Parastoo Farajpoor

Graph generative models are essential across diverse scientific domains by capturing complex distributions over relational data. Among them, graph diffusion models achieve superior performance but face inefficient sampling and limited…

机器学习 · 计算机科学 2025-06-17 Yiming Qin , Manuel Madeira , Dorina Thanou , Pascal Frossard

Dense image correspondence is central to many applications, such as visual odometry, 3D reconstruction, object association, and re-identification. Historically, dense correspondence has been tackled separately for wide-baseline scenarios…

计算机视觉与模式识别 · 计算机科学 2026-02-11 Yuchen Zhang , Nikhil Keetha , Chenwei Lyu , Bhuvan Jhamb , Yutian Chen , Yuheng Qiu , Jay Karhade , Shreyas Jha , Yaoyu Hu , Deva Ramanan , Sebastian Scherer , Wenshan Wang

Efficiently predicting motion plans directly from vision remains a fundamental challenge in robotics, where planning typically requires explicit goal specification and task-specific design. Recent vision-language-action (VLA) models infer…

Accurate per-branch 3D reconstruction is a prerequisite for autonomous UAV-based tree pruning; however, dense disparity maps from modern stereo matchers often remain too noisy for individual branch analysis in complex forest canopies. This…

图像与视频处理 · 电气工程与系统科学 2026-02-25 Yida Lin , Bing Xue , Mengjie Zhang , Sam Schofield , Richard Green

This paper studies the problem of Line Segment Detection (LSD) for the characterization of line geometry in images, with the aim of learning a domain-agnostic robust LSD model that works well for any natural images. With the focus of…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Zeran Ke , Bin Tan , Xianwei Zheng , Yujun Shen , Tianfu Wu , Nan Xue

In autonomous driving scenarios, the collected LiDAR point clouds can be challenged by occlusion and long-range sparsity, limiting the perception of autonomous driving systems. Scene completion methods can infer the missing parts of…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Andrea Matteazzi , Dietmar Tutsch

We propose a new technique for computing dense scene flow from two handheld videos with wide camera baselines and different photometric properties due to different sensors or camera settings like exposure and white balance. Our technique…

计算机视觉与模式识别 · 计算机科学 2016-09-19 Christian Richardt , Hyeongwoo Kim , Levi Valgaerts , Christian Theobalt

The proliferation of smartphones and other mobile devices provides a unique opportunity to make Advanced Driver Assistance Systems (ADAS) accessible to everyone in the form of an application empowered by low-cost Machine/Deep Learning…

计算机视觉与模式识别 · 计算机科学 2024-10-28 Muhammad Zaeem Shahzad , Muhammad Abdullah Hanif , Muhammad Shafique

Scene flow is the dense 3D reconstruction of motion and geometry of a scene. Most state-of-the-art methods use a pair of stereo images as input for full scene reconstruction. These methods depend a lot on the quality of the RGB images and…

计算机视觉与模式识别 · 计算机科学 2020-08-20 Rishav , Ramy Battrawy , René Schuster , Oliver Wasenmüller , Didier Stricker