English
Related papers

Related papers: CrossVIT-augmented Geospatial-Intelligence Visuali…

200 papers

The production of food, feed, fiber, and fuel is a key task of agriculture, which has to cope with many challenges in the upcoming decades, e.g., a higher demand, climate change, lack of workers, and the availability of arable land. Vision…

Computer Vision and Pattern Recognition · Computer Science 2024-07-25 Jan Weyler , Federico Magistri , Elias Marks , Yue Linn Chong , Matteo Sodano , Gianmarco Roggiolani , Nived Chebrolu , Cyrill Stachniss , Jens Behley

The growing homelessness crisis in the U.S. presents complex social, economic, and public health challenges, straining shelters, healthcare, and social services while limiting effective interventions. Traditional assessment methods struggle…

Computers and Society · Computer Science 2025-03-24 Julia Gersey , Rose Allegrette , Joshua Lian , Zawad Munshi , Aarti Phatke

Semantic segmentation has a broad range of applications in a variety of domains including land coverage analysis, autonomous driving, and medical image analysis. Convolutional neural networks (CNN) and Vision Transformers (ViTs) provide the…

Computer Vision and Pattern Recognition · Computer Science 2023-05-08 Hans Thisanke , Chamli Deshan , Kavindu Chamith , Sachith Seneviratne , Rajith Vidanaarachchi , Damayanthi Herath

Assigning a label to each pixel in an image, namely semantic segmentation, has been an important task in computer vision, and has applications in autonomous driving, robotic navigation, localization, and scene understanding. Fully…

Computer Vision and Pattern Recognition · Computer Science 2019-05-22 Sercan Türkmen , Janne Heikkilä

The strong temporal consistency of surveillance video enables compelling compression performance with traditional methods, but downstream vision applications operate on decoded image frames with a high data rate. Since it is not…

Multimedia · Computer Science 2024-02-09 Andrew C. Freeman , Ketan Mayer-Patel , Montek Singh

Recently, a surge of interest in visual transformers is to reduce the computational cost by limiting the calculation of self-attention to a local window. Most current work uses a fixed single-scale window for modeling by default, ignoring…

Computer Vision and Pattern Recognition · Computer Science 2022-04-11 Pengzhen Ren , Changlin Li , Guangrun Wang , Yun Xiao , Qing Du , Xiaodan Liang , Xiaojun Chang

Modern methods for data visualization via dimensionality reduction, such as t-SNE, usually have performance issues that prohibit their application to large amounts of high-dimensional data. In this work, we propose NCVis -- a…

Machine Learning · Statistics 2020-01-31 Aleksandr Artemenkov , Maxim Panov

Traversability estimation is critical for enabling robots to navigate across diverse terrains and environments. While recent self-supervised learning methods achieve promising results, they often fail to capture the characteristics of…

Robotics · Computer Science 2025-08-26 Zipeng Fang , Yanbo Wang , Lei Zhao , Weidong Chen

Smart City applications such as intelligent traffic routing or accident prevention rely on computer vision methods for exact vehicle localization and tracking. Due to the scarcity of accurately labeled data, detecting and tracking vehicles…

Computer Vision and Pattern Recognition · Computer Science 2022-08-31 Fabian Herzog , Junpeng Chen , Torben Teepe , Johannes Gilg , Stefan Hörmann , Gerhard Rigoll

Accurate and robust navigation in unstructured environments requires fusing data from multiple sensors. Such fusion ensures that the robot is better aware of its surroundings, including areas of the environment that are not immediately…

Robotics · Computer Science 2024-03-12 Mateus Valverde Gasparino , Arun Narenthiran Sivakumar , Girish Chowdhary

The study of complex many-body systems via analysis of the trajectories of the units that dynamically move and interact within them is a non-trivial task. The workflow for extracting meaningful information from the raw trajectory data is…

Materials Science · Physics 2025-10-31 Simone Martino , Matteo Becchi , Andrew Tarzia , Daniele Rapetti , Giovanni M. Pavan

Attention is sparse in vision transformers. We observe the final prediction in vision transformers is only based on a subset of most informative tokens, which is sufficient for accurate image recognition. Based on this observation, we…

Computer Vision and Pattern Recognition · Computer Science 2021-10-27 Yongming Rao , Wenliang Zhao , Benlin Liu , Jiwen Lu , Jie Zhou , Cho-Jui Hsieh

Using an ego-centric camera to do localization and tracking is highly needed for urban navigation and indoor assistive system when GPS is not available or not accurate enough. The traditional hand-designed feature tracking and estimation…

Computer Vision and Pattern Recognition · Computer Science 2018-12-04 Liang Yang , Hao Jiang , Jizhong Xiao , Zhouyuan Huo

Spatial and temporal interactions are central and fundamental in many activities in our world. A common problem faced when visualizing this type of data is how to provide an overview that helps users navigate efficiently. Traditional…

Human-Computer Interaction · Computer Science 2023-02-28 Giovani Valdrighi , Nivan Ferreira , Jorge Poco

Omnidirectional depth estimation enables efficient 3D perception over a full 360-degree range. However, in real-world applications such as autonomous driving and robotics, achieving real-time performance and robust cross-scene…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Ming Li , Xiong Yang , Chaofan Wu , Jiaheng Li , Pinzhi Wang , Xuejiao Hu , Sidan Du , Yang Li

Although vision Transformers have achieved excellent performance as backbone models in many vision tasks, most of them intend to capture global relations of all tokens in an image or a window, which disrupts the inherent spatial and local…

Computer Vision and Pattern Recognition · Computer Science 2021-12-28 Gang Li , Di Xu , Xing Cheng , Lingyu Si , Changwen Zheng

In this paper, we present the development of a sensing system with the capability to compute multispectral point clouds in real-time. The proposed multi-eye sensor system effectively registers information from the visible, (long-wave)…

Robotics · Computer Science 2021-08-20 Tanhao Zhang , Luyin Hu , Lu Li , David Navarro-Alarcon

Generating 3D city models rapidly is crucial for many applications. Monocular height estimation is one of the most efficient and timely ways to obtain large-scale geometric information. However, existing works focus primarily on training…

Computer Vision and Pattern Recognition · Computer Science 2023-09-22 Zhitong Xiong , Wei Huang , Jingtao Hu , Xiao Xiang Zhu

As urban populations grow, cities are becoming more complex, driving the deployment of interconnected sensing systems to realize the vision of smart cities. These systems aim to improve safety, mobility, and quality of life through…

Networking and Internet Architecture · Computer Science 2025-01-14 Navid Salami Pargoo , Mahshid Ghasemi , Shuren Xia , Mehmet Kerem Turkcan , Taqiya Ehsan , Chengbo Zang , Yuan Sun , Javad Ghaderi , Gil Zussman , Zoran Kostic , Jorge Ortiz

We introduce VISUALCENT, a unified human pose and instance segmentation framework to address generalizability and scalability limitations to multi person visual human analysis. VISUALCENT leverages centroid based bottom up keypoint…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Niaz Ahmad , Youngmoon Lee , Guanghui Wang