English
Related papers

Related papers: MoonAnything: A Vision Benchmark with Large-Scale …

200 papers

In an era marked by renewed interest in lunar exploration and the prospect of establishing a sustainable human presence on the Moon, innovative approaches supporting mission preparation and astronaut training are imperative. To this end,…

Instrumentation and Methods for Astrophysics · Physics 2024-10-23 Giacomo Franchini , Brenno Tuberga , Marcello Chiaberge

Understanding multi-image, multi-turn scenarios is a critical yet underexplored capability for Large Vision-Language Models (LVLMs). Existing benchmarks predominantly focus on static or horizontal comparisons -- e.g., spotting visual…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Wenbo Lyu , Yingjun Du , Jinglin Zhao , Xianton Zhen , Ling Shao

Recent advancements in audio-video joint generation models have demonstrated impressive capabilities in content creation. However, generating high-fidelity human-centric videos in complex, real-world physical scenes remains a significant…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Lei Zhu , Xing Cai , Yingjie Chen , Yiheng Li , Binxin Yang , Hao Liu , Jie Chen , Chen Li , Jing LYu

Mars exploration requires precise and reliable terrain models to ensure safe rover navigation across its unpredictable and often hazardous landscapes. Stereoscopic vision serves a critical role in the rover's perception, allowing scene…

Instrumentation and Methods for Astrophysics · Physics 2025-09-09 Yan-Shan Lu , Miguel Arana-Catania , Saurabh Upadhyay , Leonard Felicetti

Robotic systems demand accurate and comprehensive 3D environment perception, requiring simultaneous capture of photo-realistic appearance (optical), precise layout shape (geometric), and open-vocabulary scene understanding (semantic).…

Robotics · Computer Science 2025-09-10 Yinan Deng , Yufeng Yue , Jianyu Dou , Jingyu Zhao , Jiahui Wang , Yujie Tang , Yi Yang , Mengyin Fu

Food volume estimation is an essential step in the pipeline of dietary assessment and demands the precise depth estimation of the food surface and table plane. Existing methods based on computer vision require either multi-image input or…

Computer Vision and Pattern Recognition · Computer Science 2020-08-04 Ya Lu , Thomai Stathopoulou , Stavroula Mougiakakou

A hallmark of advanced artificial intelligence is the capacity to progress from passive visual perception to the strategic modification of visual information to facilitate complex reasoning. This advanced capability, however, remains…

Computer Vision and Pattern Recognition · Computer Science 2025-11-19 Jingkun Ma , Runzhe Zhan , Yang Li , Di Sun , Hou Pong Chan , Lidia S. Chao , Derek F. Wong

Context: With no conclusive detection to date, the search for exomoons, satellites of planets orbiting other stars, remains a formidable challenge. Detecting these objects, compiling a population-level sample and constraining their…

This work presents PanMatch, a versatile foundation model for robust correspondence matching. Unlike previous methods that rely on task-specific architectures and domain-specific fine-tuning to support tasks like stereo matching, optical…

Computer Vision and Pattern Recognition · Computer Science 2025-07-14 Yongjian Zhang , Longguang Wang , Kunhong Li , Ye Zhang , Yun Wang , Liang Lin , Yulan Guo

At present, large multimodal models (LMMs) have exhibited impressive generalization capabilities in understanding and generating visual signals. However, they currently still lack sufficient capability to perceive low-level visual quality…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Zhipeng Huang , Zhizheng Zhang , Yiting Lu , Zheng-Jun Zha , Zhibo Chen , Baining Guo

Large Multimodal Models (LMMs) has demonstrated capabilities across various domains, but comprehensive benchmarks for agricultural remote sensing (RS) remain scarce. Existing benchmarks designed for agricultural RS scenarios exhibit notable…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Qingmei Li , Yang Zhang , Zurong Mai , Yuhang Chen , Shuohong Lou , Henglian Huang , Jiarui Zhang , Zhiwei Zhang , Yibin Wen , Weijia Li , Haohuan Fu , Jianxi Huang , Juepeng Zheng

The lack of labeled datasets in 3D vision for surgical scenes inhibits the development of robust 3D reconstruction algorithms in the medical domain. Despite the popularity of Neural Radiance Fields and 3D Gaussian Splatting in the general…

Computer Vision and Pattern Recognition · Computer Science 2025-03-03 John J. Han , Jie Ying Wu

Equation discovery from data is a central challenge in machine learning for science, which requires the recovery of concise symbolic expressions that govern complex physical and geometric phenomena. Recent large language model (LLM)…

Machine Learning · Computer Science 2026-03-04 Sanchit Kabra , Shobhnik Kriplani , Parshin Shojaee , Chandan K. Reddy

Light-field cameras play a vital role for rich 3-D information retrieval in narrow range depth sensing applications. The key obstacle in composing light-fields from exposures taken by a plenoptic camera is to computationally calibrate,…

Image and Video Processing · Electrical Eng. & Systems 2021-07-27 Christopher Hahne , Amar Aggoun

Spectral unmixing is an important and challenging problem in hyperspectral data processing. This topic has been extensively studied and a variety of unmixing algorithms have been proposed in the literature. However, the lack of publicly…

Computer Vision and Pattern Recognition · Computer Science 2019-02-25 Min Zhao , Jie Chen , Zhe He

Earth observation foundation models have shown strong generalization across multiple Earth observation tasks, but their robustness under real-world perturbations remains underexplored. To bridge this gap, we introduce REOBench, the first…

Computer Vision and Pattern Recognition · Computer Science 2025-10-24 Xiang Li , Yong Tao , Siyuan Zhang , Siwei Liu , Zhitong Xiong , Chunbo Luo , Lu Liu , Mykola Pechenizkiy , Xiao Xiang Zhu , Tianjin Huang

Large Atomistic Models (LAMs) have undergone remarkable progress recently, emerging as universal or fundamental representations of the potential energy surface defined by the first-principles calculations of atomistic systems. However, our…

Multimodal Large Language Models (MLLMs) show reasoning promise, yet their visual perception is a critical bottleneck. Strikingly, MLLMs can produce correct answers even while misinterpreting crucial visual elements, masking these…

Computer Vision and Pattern Recognition · Computer Science 2025-12-11 Aditya Kanade , Tanuja Ganu

Large Multimodal Models (LMMs) exhibit major shortfalls when interpreting images and, by some measures, have poorer spatial cognition than small children or animals. Despite this, they attain high scores on many popular visual benchmarks,…

Feature matching is one of the most fundamental and active research areas in computer vision. A comprehensive evaluation of feature matchers is necessary, since it would advance both the development of this field and also high-level…

Computer Vision and Pattern Recognition · Computer Science 2018-08-08 JiaWang Bian , Ruihan Yang , Yun Liu , Le Zhang , Ming-Ming Cheng , Ian Reid , WenHai Wu
‹ Prev 1 4 5 6 7 8 10 Next ›