English
Related papers

Related papers: SVOM/VT: Real-Time Onboard Data Processing

200 papers

Whether by processing videos with fixed resolution from start to end or incorporating pooling and down-scaling strategies, existing video transformers process the whole video content throughout the network without specially handling the…

Computer Vision and Pattern Recognition · Computer Science 2023-04-25 Chenbin Pan , Rui Hou , Hanchao Yu , Qifan Wang , Senem Velipasalar , Madian Khabsa

We present GASV, a novel Python-based software package specifically designed for the analysis of Very Long Baseline Interferometry (VLBI) data. Developed with ease of installation and user friendliness in mind, GASV supports both pipeline…

Instrumentation and Methods for Astrophysics · Physics 2026-02-05 Dang Yao , Yuan-wei Wu , Xu-hai Yang

Vision Transformer (ViT) models have recently emerged as powerful and versatile models for various visual tasks. Recently, a work called PMF has achieved promising results in few-shot image classification by utilizing pre-trained vision…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Junjie Zhu , Yiying Li , Chunping Qiu , Ke Yang , Naiyang Guan , Xiaodong Yi

This paper describes how advanced deep learning based computer vision algorithms are applied to enable real-time on-board sensor processing for small UAVs. Four use cases are considered: target detection, classification and localization,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-03 Alessandro Palmas , Pietro Andronico

In visual planning (VP), an agent learns to plan goal-directed behavior from observations of a dynamical system obtained offline, e.g., images obtained from self-supervised robot interaction. Most previous works on VP approached the problem…

Artificial Intelligence · Computer Science 2020-02-28 Kara Liu , Thanard Kurutach , Christine Tung , Pieter Abbeel , Aviv Tamar

Particle Image Velocimetry (PIV) is a method of im-aging and analysing fields of flows. The PIV tech-niques compute and display all the motion vectors of the field in a resulting image. Speeds more than thou-sand vectors per second can be…

Hardware Architecture · Computer Science 2008-07-24 Alain Aubert , Nathalie Bochard , Virginie Fresse

We present the VVV-SkZ_pipeline, a DAOPHOT-based photometric pipeline, created to perform PSF-fitting photometry of "VISTA Variables in the V\'ia L\'actea" (VVV) ESO Public Survey data. The pipeline replaces the user avoiding repetitive…

Instrumentation and Methods for Astrophysics · Physics 2013-03-11 Francesco Mauro , Christian Moni Bidin , André-Nicolas Chené , Doug Geisler , Javier Alonso-García , Jura Borissova , Giovanni Carraro

Since the discovery of RRATs, interest in single pulse radio searches has increased dramatically. Due to the large data volumes generated by these searches, especially in planned surveys for future radio telescopes, such searches have to be…

Instrumentation and Methods for Astrophysics · Physics 2014-02-03 Alessio Magro

Fast radio bursts (FRBs) are extremely bright, millisecond duration cosmic transients of unknown origin. The growing number of wide-field and high-time-resolution radio surveys, particularly with next-generation facilities such as the SKA…

Modern GPUs come with dedicated hardware to perform ray/triangle intersections and bounding volume hierarchy (BVH) traversal. While the primary use case for this hardware is photorealistic 3D computer graphics, with careful algorithm design…

Vision-Language Models (VLMs) have demonstrated strong performance on multimodal reasoning tasks, but their deployment remains challenging due to high inference latency and computational cost, particularly when processing high-resolution…

Computer Vision and Pattern Recognition · Computer Science 2025-12-25 Putu Indah Githa Cahyani , Komang David Dananjaya Suartana , Novanto Yudistira

Vision language models (VLMs) demonstrate strong capabilities in jointly processing visual and textual data. However, they often incur substantial computational overhead due to redundant visual information, particularly in long-form video…

Machine Learning · Computer Science 2025-04-25 Yudong Liu , Jingwei Sun , Yueqian Lin , Jingyang Zhang , Ming Yin , Qinsi Wang , Jianyi Zhang , Hai Li , Yiran Chen

Vision Transformers (ViTs) have emerged as the backbone of many segmentation models, consistently achieving state-of-the-art (SOTA) performance. However, their success comes at a significant computational cost. Image token pruning is one of…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Hanning Chen , Yang Ni , Wenjun Huang , Yezi Liu , SungHeon Jeong , Fei Wen , Nathaniel Bastian , Hugo Latapie , Mohsen Imani

We consider how to most efficiently leverage teleoperator time to collect data for learning robust image-based value functions and policies for sparse reward robotic tasks. To accomplish this goal, we modify the process of data collection…

Robotics · Computer Science 2022-10-06 David Brandfonbrener , Stephen Tu , Avi Singh , Stefan Welker , Chad Boodoo , Nikolai Matni , Jake Varley

The Australian SKA Pathfinder (ASKAP) is a next generation radio telescope currently under construction in Western Australia. The fast survey speed and wide field of view make it an ideal instrument for blind transients searches. The ASKAP…

Instrumentation and Methods for Astrophysics · Physics 2012-01-17 Jay Banyer , Tara Murphy , the VAST Collaboration

Visual place recognition methods struggle with occlusions and partial visual overlaps. We propose a novel visual place recognition approach based on overlap prediction, called VOP, shifting from traditional reliance on global image…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 Tong Wei , Philipp Lindenberger , Jiri Matas , Daniel Barath

A software-defined optical receiver is implemented on an off-the-shelf commercial graphics processing unit (GPU). The receiver provides real-time signal processing functionality to process 1 GBaud minimum phase (MP) 4-, 8-, 16-, 32-, 64-,…

Vision Transformers (ViTs) have demonstrated remarkable capabilities in learning representations, but their performance is compromised when applied to unseen domains. Previous methods either engage in prompt learning during the training…

Computer Vision and Pattern Recognition · Computer Science 2024-09-11 Yunbei Zhang , Akshay Mehra , Jihun Hamm

In recent years, Transformer has achieved good results in Natural Language Processing (NLP) and has also started to expand into Computer Vision (CV). Excellent models such as the Vision Transformer and Swin Transformer have emerged. At the…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Wei Hu , Dian Xu , Zimeng Fan , Fang Liu , Yanxiang He

The primary challenge in computer vision is precisely calculating the pose of 6D objects, however many current approaches are still fragile and have trouble generalizing from synthetic data to real-world situations with fluctuating…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Md Selim Sarowar , Sungho Kim
‹ Prev 1 4 5 6 7 8 10 Next ›