中文
相关论文

相关论文: MTSGL: Multi-Task Structure Guided Learning for Ro…

200 篇论文

Autonomous spacecraft control via Shielded Deep Reinforcement Learning (SDRL) has become a rapidly growing research area. However, the construction of shields and the definition of tasking remains informal, resulting in policies with no…

机器学习 · 计算机科学 2024-03-15 Robert Reed , Hanspeter Schaub , Morteza Lahijanian

Array synthetic aperture radar (Array-SAR), also known as tomographic SAR (TomoSAR), has demonstrated significant potential for high-quality 3D mapping, particularly in urban areas.While deep learning (DL) methods have recently shown…

图像与视频处理 · 电气工程与系统科学 2024-12-24 Yu Ren , Xu Zhan , Yunqiao Hu , Xiangdong Ma , Liang Liu , Mou Wang , Jun Shi , Shunjun Wei , Tianjiao Zeng , Xiaoling Zhang

Skeleton-aware sign language recognition (SLR) has gained popularity due to its ability to remain unaffected by background information and its lower computational requirements. Current methods utilize spatial graph modules and temporal…

计算机视觉与模式识别 · 计算机科学 2024-03-20 Lianyu Hu , Liqing Gao , Zekang Liu , Wei Feng

At present, the Synthetic Aperture Radar (SAR) image classification method based on convolution neural network (CNN) has faced some problems such as poor noise resistance and generalization ability. Spiking neural network (SNN) is one of…

计算机视觉与模式识别 · 计算机科学 2021-06-16 Jiankun Chen , Xiaolan Qiu , Chibiao Ding , Yirong Wu

Low-rank tensor representation (LRTR) has emerged as a powerful tool for multi-dimensional data processing. However, classical LRTR-based methods face two critical limitations: (1) they typically assume that the holistic data is low-rank,…

计算机视觉与模式识别 · 计算机科学 2025-08-21 Zhizhou Wang , Jianli Wang , Ruijing Zheng , Zhenyu Wu

Deep generative models (DGMs) have shown promise in image generation. However, most of the existing work learn the model by simply optimizing a divergence between the marginal distributions of the model and the data, and often fail to…

机器学习 · 计算机科学 2019-06-11 Kun Xu , Chongxuan Li , Jun Zhu , Bo Zhang

Whole-slide images (WSIs) are critical for cancer diagnosis due to their ultra-high resolution and rich semantic content. However, their massive size and the limited availability of fine-grained annotations pose substantial challenges for…

计算机视觉与模式识别 · 计算机科学 2025-06-30 Daoxi Cao , Hangbei Cheng , Yijin Li , Ruolin Zhou , Xuehan Zhang , Xinyi Li , Binwei Li , Xuancheng Gu , Jianan Zhang , Xueyu Liu , Yongfei Wu

Understanding urban dynamics and promoting sustainable development requires comprehensive insights about buildings. While geospatial artificial intelligence has advanced the extraction of such details from Earth observational data, existing…

计算机视觉与模式识别 · 计算机科学 2023-10-31 Zhen Qian , Min Chen , Zhuo Sun , Fan Zhang , Qingsong Xu , Jinzhao Guo , Zhiwei Xie , Zhixin Zhang

Synthetic Aperture Radar (SAR) is a crucial remote sensing technology, enabling all-weather, day-and-night observation with strong surface penetration for precise and continuous environmental monitoring and analysis. However, SAR image…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Yimin Wei , Aoran Xiao , Yexian Ren , Yuting Zhu , Hongruixuan Chen , Junshi Xia , Naoto Yokoya

Self-supervised learning (SSL) is a standard approach for representation learning in aerial imagery. Existing methods enforce invariance between augmented views, which works well when augmentations preserve semantic content. However, aerial…

计算机视觉与模式识别 · 计算机科学 2026-04-24 Wadii Boulila , Adel Ammar , Bilel Benjdira , Maha Driss

Air traffic trajectory recognition has gained significant interest within the air traffic management community, particularly for fundamental tasks such as classification and clustering. This paper introduces Aircraft Trajectory…

机器学习 · 计算机科学 2024-10-23 Thaweerath Phisannupawong , Joshua Julian Damanik , Han-Lim Choi

Spatial understanding remains a weakness of Large Vision-Language Models (LVLMs). Existing supervised fine-tuning (SFT) and recent reinforcement learning with verifiable rewards (RLVR) pipelines depend on costly supervision, specialized…

计算机视觉与模式识别 · 计算机科学 2025-11-26 Yuhong Liu , Beichen Zhang , Yuhang Zang , Yuhang Cao , Long Xing , Xiaoyi Dong , Haodong Duan , Dahua Lin , Jiaqi Wang

Recent advancements in deep learning have greatly enhanced 3D object recognition, but most models are limited to closed-set scenarios, unable to handle unknown samples in real-world applications. Open-set recognition (OSR) addresses this…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Jinfeng Xu , Xianzhi Li , Yuan Tang , Xu Han , Qiao Yu , Yixue Hao , Long Hu , Min Chen

A machine can understand human activities, and the meaning of signs can help overcome the communication barriers between the inaudible and ordinary people. Sign Language Recognition (SLR) is a fascinating research area and a crucial task…

计算机视觉与模式识别 · 计算机科学 2024-09-02 M. Madhiarasan , Partha Pratim Roy

Similarity analysis is one of the crucial steps in most fMRI studies. Representational Similarity Analysis (RSA) can measure similarities of neural signatures generated by different cognitive states. This paper develops Deep…

图像与视频处理 · 电气工程与系统科学 2020-10-06 Muhammad Yousefnezhad , Jeffrey Sawalha , Alessandro Selvitella , Daoqiang Zhang

Despite impressive advancements in Visual-Language Models (VLMs) for multi-modal tasks, their reliance on RGB inputs limits precise spatial understanding. Existing methods for integrating spatial cues, such as point clouds or depth, either…

计算机视觉与模式识别 · 计算机科学 2025-10-27 Yang Liu , Ming Ma , Xiaomin Yu , Pengxiang Ding , Han Zhao , Mingyang Sun , Siteng Huang , Donglin Wang

Existing color-guided depth super-resolution (DSR) approaches require paired RGB-D data as training samples where the RGB image is used as structural guidance to recover the degraded depth map due to their geometrical similarity. However,…

计算机视觉与模式识别 · 计算机科学 2021-03-25 Baoli Sun , Xinchen Ye , Baopu Li , Haojie Li , Zhihui Wang , Rui Xu

Multi-label Recognition (MLR) involves assigning multiple labels to each data instance in an image, offering advantages over single-label classification in complex scenarios. However, it faces the challenge of annotating all relevant…

机器学习 · 计算机科学 2025-06-03 Ruhui Zhang , Hezhe Qiao , Pengcheng Xu , Mingsheng Shang , Lin Chen

This paper proposes a new architecture - Attentive Tensor Product Learning (ATPL) - to represent grammatical structures in deep learning models. ATPL is a new architecture to bridge this gap by exploiting Tensor Product Representations…

计算与语言 · 计算机科学 2018-11-30 Qiuyuan Huang , Li Deng , Dapeng Wu , Chang Liu , Xiaodong He

The rise of unmanned aerial vehicle (UAV) operations, as well as the vulnerability of the UAVs' sensors, has led to the need for proper monitoring systems for detecting any abnormal behavior of the UAV. This work addresses this problem by…

系统与控制 · 电气工程与系统科学 2023-10-18 Antreas Palamas , Nicolas Souli , Tania Panayiotou , Panayiotis Kolios , Georgios Ellinas