中文
相关论文

相关论文: BGM: Background Mixup for X-ray Prohibited Items D…

200 篇论文

Faithfully reconstructing 3D geometry and generating novel views of scenes are critical tasks in 3D computer vision. Despite the widespread use of image augmentations across computer vision applications, their potential remains…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Juan C. Pérez , Sara Rojas , Jesus Zarzar , Bernard Ghanem

In this paper we propose a novel image representation called face X-ray for detecting forgery in face images. The face X-ray of an input face image is a greyscale image that reveals whether the input image can be decomposed into the…

计算机视觉与模式识别 · 计算机科学 2020-04-21 Lingzhi Li , Jianmin Bao , Ting Zhang , Hao Yang , Dong Chen , Fang Wen , Baining Guo

In the current paradigm of image captioning, deep learning models are trained to generate text from image embeddings of latent features. We challenge the assumption that fine-tuning of large, bespoke models is required to improve model…

计算机视觉与模式识别 · 计算机科学 2025-10-08 Steven Song , Anirudh Subramanyam , Irene Madejski , Robert L. Grossman

In video analysis, background models have many applications such as background/foreground separation, change detection, anomaly detection, tracking, and more. However, while learning such a model in a video captured by a static camera is a…

计算机视觉与模式识别 · 计算机科学 2022-09-19 Guy Erez , Ron Shapira Weber , Oren Freifeld

Biological signals of interest in high-dimensional data are often masked by dominant variation shared across conditions. This variation, arising from baseline biological structure or technical effects, can prevent standard dimensionality…

机器学习 · 计算机科学 2026-02-27 Yixuan Li , Archer Y. Yang , Yue Li

Many recent semi-supervised learning (SSL) studies build teacher-student architecture and train the student network by the generated supervisory signal from the teacher. Data augmentation strategy plays a significant role in the SSL…

计算机视觉与模式识别 · 计算机科学 2022-03-16 JongMok Kim , Jooyoung Jang , Seunghyeon Seo , Jisoo Jeong , Jongkeun Na , Nojun Kwak

Enhancing RAW images captured under low light conditions is a challenging task. Recent deep learning based RAW enhancement methods have shifted from using real paired data to relying on synthetic datasets. These synthetic datasets are…

图像与视频处理 · 电气工程与系统科学 2025-09-11 Juntai Zeng

Learning-based image harmonization techniques are usually trained to undo synthetic random global transformations applied to a masked foreground in a single ground truth photo. This simulated data does not model many of the important…

计算机视觉与模式识别 · 计算机科学 2023-03-02 Ke Wang , Michaël Gharbi , He Zhang , Zhihao Xia , Eli Shechtman

In this paper, a color edge detection strategy based on collaborative filtering combined with multiscale gradient fusion is proposed. The block-matching and 3D (BM3D) filter are used to enhance the sparse representation in the transform…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Zhuoyue Wang , Yiyi Tao , Danqing Ma , Jiajing Chen

Neural networks are prone to overfitting and memorizing data patterns. To avoid over-fitting and enhance their generalization and performance, various methods have been suggested in the literature, including dropout, regularization, label…

计算机视觉与模式识别 · 计算机科学 2023-02-08 Humza Naveed , Saeed Anwar , Munawar Hayat , Kashif Javed , Ajmal Mian

Recently, a number of image-mixing-based augmentation techniques have been introduced to improve the generalization of deep neural networks. In these techniques, two or more randomly selected natural images are mixed together to generate an…

计算机视觉与模式识别 · 计算机科学 2024-05-27 Khawar Islam , Muhammad Zaigham Zaheer , Arif Mahmood , Karthik Nandakumar

Raman spectroscopy has attracted interest as a non-invasive optical technique to study the composition and structure of a wide range of materials at the microscopic level. The intrinsic fluorescence background can be orders of magnitude…

材料科学 · 物理学 2015-10-28 P. J. Cadusch , M. M. Hlaing , S. A. Wade , S. L. McArthur , P. R. Stoddart

Camouflaged objects that blend into natural scenes pose significant challenges for deep-learning models to detect and synthesize. While camouflaged object detection is a crucial task in computer vision with diverse real-world applications,…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Haichao Zhang , Can Qin , Yu Yin , Yun Fu

Multimodal synthetic data generation is crucial in domains such as autonomous driving, robotics, augmented/virtual reality, and retail. We propose a novel approach, GenMM, for jointly editing RGB videos and LiDAR scans by inserting…

计算机视觉与模式识别 · 计算机科学 2024-06-18 Bharat Singh , Viveka Kulharia , Luyu Yang , Avinash Ravichandran , Ambrish Tyagi , Ashish Shrivastava

Data augmentation is widely used to enhance generalization in visual classification tasks. However, traditional methods struggle when source and target domains differ, as in domain adaptation, due to their inability to address domain gaps.…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Khawar Islam , Muhammad Zaigham Zaheer , Arif Mahmood , Karthik Nandakumar , Naveed Akhtar

Background and aim: Surgical telepresence using augmented perception has been applied, but mixed reality is still being researched and is only theoretical. The aim of this work is to propose a solution to improve the visualization in the…

多媒体 · 计算机科学 2021-10-19 Nirakar Puri , Abeer Alsadoon , P. W. C. Prasad , Nada Alsalami , Tarik A. Rashid

Ground Penetrating Radar (GPR) has been widely used to estimate the healthy operation of some urban roads and underground facilities. When identifying subsurface anomalies by GPR in an area, the obtained data could be unbalanced, and the…

计算机视觉与模式识别 · 计算机科学 2022-11-23 Xiren Zhou , Shikang Liu , Ao Chen , Yizhan Fan , Huanhuan Chen

Future high-sensitivity measurements of the cosmic microwave background (CMB) anisotropies and energy spectrum will be limited by our understanding and modeling of foregrounds. Not only does more information need to be gathered and…

宇宙学与河外天体物理 · 物理学 2017-09-20 Jens Chluba , J. Colin Hill , Maximilian H. Abitbol

In light of the success of contrastive learning in the image domain, current self-supervised video representation learning methods usually employ contrastive loss to facilitate video representation learning. When naively pulling two…

计算机视觉与模式识别 · 计算机科学 2022-03-15 Shuangrui Ding , Maomao Li , Tianyu Yang , Rui Qian , Haohang Xu , Qingyi Chen , Jue Wang , Hongkai Xiong

Vision-language Navigation (VLN) tasks require an agent to navigate step-by-step while perceiving the visual observations and comprehending a natural language instruction. Large data bias, which is caused by the disparity ratio between the…

计算机视觉与模式识别 · 计算机科学 2021-11-02 Chong Liu , Fengda Zhu , Xiaojun Chang , Xiaodan Liang , Zongyuan Ge , Yi-Dong Shen