中文
相关论文

相关论文: Multimodal Crowd Counting with Pix2Pix GANs

200 篇论文

Visual crowd counting estimates the density of the crowd using deep learning models such as convolution neural networks (CNNs). The performance of the model heavily relies on the quality of the training data that constitutes crowd images.…

计算机视觉与模式识别 · 计算机科学 2023-10-12 Muhammad Asif Khan , Hamid Menouar , Ridha Hamila

Crowd counting aims to estimate the number of persons in a scene. Most state-of-the-art crowd counting methods based on color images can't work well in poor illumination conditions due to invisible objects. With the widespread use of…

计算机视觉与模式识别 · 计算机科学 2023-01-10 Zhengyi Liu , Wei Wu , Yacheng Tan , Guanghui Zhang

In the field of computer vision, multimodal image generation has become a research hotspot, especially the task of integrating text, image, and style. In this study, we propose a multimodal image generation method based on Generative…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Chaoyi Tan , Wenqing Zhang , Zhen Qi , Kowei Shih , Xinshi Li , Ao Xiang

Electrical tomography techniques have been widely employed for multiphase-flow monitoring owing to their non invasive nature, intrinsic safety, and low cost. Nevertheless, conventional reconstructions struggle to capture fine details, which…

图像与视频处理 · 电气工程与系统科学 2025-12-23 Wejian Yan

We give an overview of the different rendering methods and we demonstrate that the use of a Generative Adversarial Networks (GAN) for Global Illumination (GI) gives a superior quality rendered image to that of a rasterisations image. We…

计算机视觉与模式识别 · 计算机科学 2021-10-26 Jared Harris-Dewey , Richard Klein

This paper presents a generative adversarial network (GAN) based approach for radar image enhancement. Although radar sensors remain robust for operations under adverse weather conditions, their application in autonomous vehicles (AVs) is…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Thakshila Thilakanayake , Oscar De Silva , Thumeera R. Wanasinghe , George K. Mann , Awantha Jayasiri

Generative Adversarial Networks (GANs) have significantly advanced image processing, with Pix2Pix being a notable framework for image-to-image translation. This paper explores a novel application of Pix2Pix to transform abstract map images…

计算机视觉与模式识别 · 计算机科学 2024-05-02 Zhenglin Li , Bo Guan , Yuanzhou Wei , Yiming Zhou , Jingyu Zhang , Jinxin Xu

Crowd counting is a fundamental yet challenging task, which desires rich information to generate pixel-wise crowd density maps. However, most previous methods only used the limited information of RGB images and cannot well discover…

计算机视觉与模式识别 · 计算机科学 2021-04-07 Lingbo Liu , Jiaqi Chen , Hefeng Wu , Guanbin Li , Chenglong Li , Liang Lin

In medical imaging, a general problem is that it is costly and time consuming to collect high quality data from healthy and diseased subjects. Generative adversarial networks (GANs) is a deep learning method that has been developed for…

计算机视觉与模式识别 · 计算机科学 2018-06-21 Per Welander , Simon Karlsson , Anders Eklund

In this work, we tackle the problem of crowd counting in images. We present a Convolutional Neural Network (CNN) based density estimation approach to solve this problem. Predicting a high resolution density map in one go is a challenging…

计算机视觉与模式识别 · 计算机科学 2018-07-27 Viresh Ranjan , Hieu Le , Minh Hoai

Several visual tasks, such as pedestrian detection and image-to-image translation, are challenging to accomplish in low light using RGB images. Heat variation of objects in thermal images can be used to overcome this. In this work, an…

计算机视觉与模式识别 · 计算机科学 2023-11-10 Md Azim Khan

Generative Adversarial Networks (GANs) have recently introduced effective methods of performing Image-to-Image translations. These models can be applied and generalized to a variety of domains in Image-to-Image translation without changing…

计算机视觉与模式识别 · 计算机科学 2022-08-30 Sagar Saxena , Mohammad Nayeem Teli

Generative adversarial networks (GANs) are unsupervised Deep Learning approach in the computer vision community which has gained significant attention from the last few years in identifying the internal structure of multimodal medical…

图像与视频处理 · 电气工程与系统科学 2020-05-22 Nripendra Kumar Singh , Khalid Raza

Advanced Driver Assistance Systems (ADAS) in intelligent vehicles rely on accurate driver perception within the vehicle cabin, often leveraging a combination of sensing modalities. However, these modalities operate at varying rates, posing…

计算机视觉与模式识别 · 计算机科学 2024-03-04 Mathias Viborg Andersen , Ross Greer , Andreas Møgelmose , Mohan Trivedi

RGB-Thermal (RGB-T) crowd counting is a challenging task, which uses thermal images as complementary information to RGB images to deal with the decreased performance of unimodal RGB-based methods in scenes with low-illumination or similar…

计算机视觉与模式识别 · 计算机科学 2022-08-16 Pengyu Chen , Junyu Gao , Yuan Yuan , Qi Wang

In this paper, we propose a three-stream adaptive fusion network named TAFNet, which uses paired RGB and thermal images for crowd counting. Specifically, TAFNet is divided into one main stream and two auxiliary streams. We combine a pair of…

计算机视觉与模式识别 · 计算机科学 2022-02-18 Haihan Tang , Yi Wang , Lap-Pui Chau

Understanding and predicting the intention of pedestrians is essential to enable autonomous vehicles and mobile robots to navigate crowds. This problem becomes increasingly complex when we consider the uncertainty and multimodality of…

计算机视觉与模式识别 · 计算机科学 2020-07-14 Stuart Eiffert , Kunming Li , Mao Shan , Stewart Worrall , Salah Sukkarieh , Eduardo Nebot

While cloud/sky image segmentation has extensive real-world applications, a large amount of labelled data is needed to train a highly accurate models to perform the task. Scarcity of such volumes of cloud/sky images with corresponding…

计算机视觉与模式识别 · 计算机科学 2021-06-08 Mayank Jain , Conor Meegan , Soumyabrata Dev

The increasing prevalence of gigapixel resolutions has presented new challenges for crowd counting. Such resolutions are far beyond the memory and computation limits of current GPUs, and available deep neural network architectures and…

计算机视觉与模式识别 · 计算机科学 2023-05-17 Arian Bakhtiarnia , Qi Zhang , Alexandros Iosifidis

Magnetic Resonance Imaging (MRI) of the brain has been used to investigate a wide range of neurological disorders, but data acquisition can be expensive, time-consuming, and inconvenient. Multi-site studies present a valuable opportunity to…

计算机视觉与模式识别 · 计算机科学 2018-04-13 Harrison Nguyen , Richard W. Morris , Anthony W. Harris , Mayuresh S. Korgoankar , Fabio Ramos
‹ 上一页 1 2 3 10 下一页 ›