中文
相关论文

相关论文: DSRC: Learning Density-insensitive and Semantic-aw…

200 篇论文

Multi-agent collaborative perception (MCP) has recently attracted much attention. It includes three key processes: communication for sharing, collaboration for integration, and reconstruction for different downstream tasks. Existing methods…

计算机视觉与模式识别 · 计算机科学 2023-03-23 Tianhang Wang , Guang Chen , Kai Chen , Zhengfa Liu , Bo Zhang , Alois Knoll , Changjun Jiang

Vision and touch are two of the important sensing modalities for humans and they offer complementary information for sensing the environment. Robots could also benefit from such multi-modal sensing ability. In this paper, addressing for the…

机器人学 · 计算机科学 2018-03-14 Shan Luo , Wenzhen Yuan , Edward Adelson , Anthony G. Cohn , Raul Fuentes

Big progress has been achieved in domain adaptation in decades. Existing works are always based on an ideal assumption that testing target domain are i.i.d. with training target domains. However, due to unpredictable corruptions (e.g.,…

计算机视觉与模式识别 · 计算机科学 2021-04-22 Yifan Xu , Kekai Sheng , Weiming Dong , Baoyuan Wu , Changsheng Xu , Bao-Gang Hu

We address the problem of inferring self-supervised dense semantic correspondences between objects in multi-object scenes. The method introduces learning of class-aware dense object descriptors by providing either unsupervised discrete…

机器人学 · 计算机科学 2021-10-06 Denis Hadjivelichkov , Dimitrios Kanoulas

Learning generative models directly from corrupted observations is a long standing challenge across natural and scientific domains. We introduce Restoration Score Distillation (RSD), a unified framework for learning high fidelity, one step…

机器学习 · 计算机科学 2026-03-19 Yasi Zhang , Tianyu Chen , Zhendong Wang , Ying Nian Wu , Mingyuan Zhou , Oscar Leong

The paper introduces a novel, holistic approach for robust Screen-Camera Communication (SCC), where video content on a screen is visually encoded in a human-imperceptible fashion and decoded by a camera capturing images of such screen…

计算机视觉与模式识别 · 计算机科学 2021-05-13 Vu Tran , Gihan Jayatilaka , Ashwin Ashok , Archan Misra

We present an efficient document representation learning framework, Document Vector through Corruption (Doc2VecC). Doc2VecC represents each document as a simple average of word embeddings. It ensures a representation generated as such…

计算与语言 · 计算机科学 2017-07-11 Minmin Chen

Cross-modal learning has become a fundamental paradigm for integrating heterogeneous information sources such as images, text, and structured attributes. However, multimodal representations often suffer from modality dominance, redundant…

The rapid evolution of generative adversarial networks (GANs) and diffusion models has made synthetic media increasingly realistic, raising societal concerns around misinformation, identity fraud, and digital trust. Existing deepfake…

计算机视觉与模式识别 · 计算机科学 2025-11-03 Sales Aribe

Collaborative perception (CP) enhances scene understanding through multi-agent information sharing. While LiDAR-centric systems offer precise geometry, high costs and performance degradation in adverse weather necessitate multi-modal…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Xiaokai Bai , Lianqing Zheng , Runwei Guan , Siyuan Cao , Huiliang Shen

This paper presents CORE, a conceptually simple, effective and communication-efficient model for multi-agent cooperative perception. It addresses the task from a novel perspective of cooperative reconstruction, based on two key insights: 1)…

计算机视觉与模式识别 · 计算机科学 2023-07-26 Binglu Wang , Lei Zhang , Zhaozhong Wang , Yongqiang Zhao , Tianfei Zhou

Machine Reading Comprehension (MRC) is an important testbed for evaluating models' natural language understanding (NLU) ability. There has been rapid progress in this area, with new models achieving impressive performance on various…

计算与语言 · 计算机科学 2021-05-27 Chenglei Si , Ziqing Yang , Yiming Cui , Wentao Ma , Ting Liu , Shijin Wang

Deep reinforcement learning (DRL) has achieved remarkable progress in online path planning tasks for multi-UAV systems. However, existing DRL-based methods often suffer from performance degradation when tackling unseen scenarios, since the…

机器人学 · 计算机科学 2024-07-16 Jiafan Zhuang , Zihao Xia , Gaofei Han , Boxi Wang , Wenji Li , Dongliang Wang , Zhifeng Hao , Ruichu Cai , Zhun Fan

The state-of-the-art deep neural networks are vulnerable to common corruptions (e.g., input data degradations, distortions, and disturbances caused by weather changes, system error, and processing). While much progress has been made in…

计算机视觉与模式识别 · 计算机科学 2022-08-23 Chenyu Yi , Siyuan Yang , Haoliang Li , Yap-peng Tan , Alex Kot

In this paper, we present D2C-SR, a novel framework for the task of real-world image super-resolution. As an ill-posed problem, the key challenge in super-resolution related tasks is there can be multiple predictions for a given…

计算机视觉与模式识别 · 计算机科学 2022-07-21 Youwei Li , Haibin Huang , Lanpeng Jia , Haoqiang Fan , Shuaicheng Liu

The LiDAR-based multi-agent and single-agent perception has shown promising performance in environmental understanding for robots and automated vehicles. However, there is no existing method that simultaneously solves both multi-agent and…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Haochen Yang , Baolu Li , Lei Li , Delin Ren , Jiacheng Guo , Minghai Qin , Tianyun Zhang , Hongkai Yu

Bias is a common problem inherent in recommender systems, which is entangled with users' preferences and poses a great challenge to unbiased learning. For debiasing tasks, the doubly robust (DR) method and its variants show superior…

信息检索 · 计算机科学 2023-03-03 Haoxuan Li , Yan Lyu , Chunyuan Zheng , Peng Wu

In this paper, we propose a new unsupervised feature learning framework, namely Deep Sparse Coding (DeepSC), that extends sparse coding to a multi-layer architecture for visual object recognition tasks. The main innovation of the framework…

机器学习 · 计算机科学 2013-12-23 Yunlong He , Koray Kavukcuoglu , Yun Wang , Arthur Szlam , Yanjun Qi

Autonomous robotic manipulation in clutter is challenging. A large variety of objects must be perceived in complex scenes, where they are partially occluded and embedded among many distractors, often in restricted spaces. To tackle these…

计算机视觉与模式识别 · 计算机科学 2018-10-03 Max Schwarz , Anton Milan , Arul Selvam Periyasamy , Sven Behnke