English
Related papers

Related papers: SA-CycleGAN-2.5D: Self-Attention CycleGAN with Tri…

200 papers

We propose two new techniques for training Generative Adversarial Networks (GANs). Our objectives are to alleviate mode collapse in GAN and improve the quality of the generated samples. First, we propose neighbor embedding, a manifold…

Computer Vision and Pattern Recognition · Computer Science 2018-11-06 Ngoc-Trung Tran , Tuan-Anh Bui , Ngai-Man Cheung

Training (source) domain bias affects state-of-the-art object detectors, such as Faster R-CNN, when applied to new (target) domains. To alleviate this problem, researchers proposed various domain adaptation methods to improve object…

Computer Vision and Pattern Recognition · Computer Science 2021-01-22 Petru Soviany , Radu Tudor Ionescu , Paolo Rota , Nicu Sebe

Medical semantic-mask synthesis boosts data augmentation and analysis, yet most GAN-based approaches still produce one-to-one images and lack spatial consistency in complex scans. To address this, we propose AnatoMaskGAN, a novel synthesis…

Image and Video Processing · Electrical Eng. & Systems 2025-08-18 Zonglin Wu , Yule Xue , Qianxiang Hu , Yaoyao Feng , Yuqi Ma , Shanxiong Chen

The performance of most speaker diarization systems with x-vector embeddings is both vulnerable to noisy environments and lacks domain robustness. Earlier work on speaker diarization using generative adversarial network (GAN) with an…

Audio and Speech Processing · Electrical Eng. & Systems 2020-07-21 Monisankha Pal , Manoj Kumar , Raghuveer Peri , Tae Jin Park , So Hyun Kim , Catherine Lord , Somer Bishop , Shrikanth Narayanan

Precise analysis of nanoparticles for characterization in electron microscopy images is essential for advancing nanomaterial development. Yet it remains challenging due to the time-consuming nature of manual methods and the shortcomings of…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Anindya Pal , Varun Ajith , Saumik Bhattacharya , Sayantari Ghosh

We introduce a new local sparse attention layer that preserves two-dimensional geometry and locality. We show that by just replacing the dense attention layer of SAGAN with our construction, we obtain very significant FID, Inception score…

Machine Learning · Computer Science 2019-12-03 Giannis Daras , Augustus Odena , Han Zhang , Alexandros G. Dimakis

3D GAN inversion aims to project a single image into the latent space of a 3D Generative Adversarial Network (GAN), thereby achieving 3D geometry reconstruction. While there exist encoders that achieve good results in 3D GAN inversion, they…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Bahri Batuhan Bilecen , Ahmet Berke Gokmen , Aysegul Dundar

Modern optical microscopes are fully motorised; however, transforming them into truly smart systems requires real-time adjustment of acquisition settings in response to detected objects and dynamic biological events. At the core are…

Acquisition differences across sites, scanners, and protocols in dMRI introduce variability that complicates structural connectome analysis. This motivates deep learning models that can represent high-dimensional connectomes in a…

Medical image segmentation requires models that preserve fine anatomical boundaries while remaining practical for clinical deployment. Transformers capture long-range dependencies but incur quadratic attention cost, whereas CNNs are…

Computer Vision and Pattern Recognition · Computer Science 2026-05-04 Hongbo Zheng , Afshin Bozorgpour , Dorit Merhof , Minjia Zhang

Drug combinations are essential in cancer therapy, leveraging synergistic drug-drug interactions (DDI) to enhance efficacy and combat resistance. However, the vast combinatorial space makes experimental screening impractical, and existing…

Machine Learning · Computer Science 2026-03-26 Yuxuan Nie , Yutong Song , Jinjie Yang , Yupeng Song , Yujue Zhou , Hong Peng

Although deep convolutional networks have been widely studied for head and neck (HN) organs at risk (OAR) segmentation, their use for routine clinical treatment planning is limited by a lack of robustness to imaging artifacts, low soft…

Computer Vision and Pattern Recognition · Computer Science 2021-03-01 Harini Veeraraghavan , Jue Jiang , Sharif Elguindi , Sean L. Berry , Ifeanyirochukwu Onochie , Aditya Apte , Laura Cervino , Joseph O. Deasy

There are considerable interests in automatic stroke lesion segmentation on magnetic resonance (MR) images in the medical imaging field, as stroke is an important cerebrovascular disease. Although deep learning-based models have been…

Image and Video Processing · Electrical Eng. & Systems 2023-03-07 Weiyi Yu , Zhizhong Huang , Junping Zhang , Hongming Shan

Training a model to perform a task typically requires a large amount of data from the domains in which the task will be applied. However, it is often the case that data are abundant in some domains but scarce in others. Domain adaptation…

Machine Learning · Computer Science 2019-01-25 Ehsan Hosseini-Asl , Yingbo Zhou , Caiming Xiong , Richard Socher

Multiplex brightfield imaging offers the advantage of simultaneously analyzing multiple biomarkers on a single slide, as opposed to single biomarker labeling on multiple consecutive slides. To accurately analyze multiple biomarkers…

Image and Video Processing · Electrical Eng. & Systems 2024-08-16 Satarupa Mukherjee , Jim Martin , Yao Nie

Diversity in data is critical for the successful training of deep learning models. Leveraged by a recurrent generative adversarial network, we propose the CT-SGAN model that generates large-scale 3D synthetic CT-scan volumes ($\geq…

Image and Video Processing · Electrical Eng. & Systems 2021-11-08 Ahmad Pesaranghader , Yiping Wang , Mohammad Havaei

In this paper, we focus on the semantic image synthesis task that aims at transferring semantic label maps to photo-realistic images. Existing methods lack effective semantic constraints to preserve the semantic information and ignore the…

Computer Vision and Pattern Recognition · Computer Science 2020-09-01 Hao Tang , Song Bai , Nicu Sebe

Compressed Sensing MRI (CS-MRI) has provided theoretical foundations upon which the time-consuming MRI acquisition process can be accelerated. However, it primarily relies on iterative numerical solvers which still hinders their adaptation…

Computer Vision and Pattern Recognition · Computer Science 2018-06-12 Tran Minh Quan , Thanh Nguyen-Duc , Won-Ki Jeong

Computed tomography (CT) is widely used in screening, diagnosis, and image-guided therapy for both clinical and research purposes. Since CT involves ionizing radiation, an overarching thrust of related technical research is development of…

Image and Video Processing · Electrical Eng. & Systems 2019-06-25 Chenyu You , Guang Li , Yi Zhang , Xiaoliu Zhang , Hongming Shan , Shenghong Ju , Zhen Zhao , Zhuiyang Zhang , Wenxiang Cong , Michael W. Vannier , Punam K. Saha , Ge Wang

Recently, 3D GANs based on 3D Gaussian splatting have been proposed for high quality synthesis of human heads. However, existing methods stabilize training and enhance rendering quality from steep viewpoints by conditioning the random…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Florian Barthel , Wieland Morgenstern , Paul Hinzer , Anna Hilsmann , Peter Eisert