中文
相关论文

相关论文: Contrastive-SDXL: Annotation-Preserving Night-Time…

200 篇论文

We present SDXL, a latent diffusion model for text-to-image synthesis. Compared to previous versions of Stable Diffusion, SDXL leverages a three times larger UNet backbone: The increase of model parameters is mainly due to more attention…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Dustin Podell , Zion English , Kyle Lacey , Andreas Blattmann , Tim Dockhorn , Jonas Müller , Joe Penna , Robin Rombach

Nighttime surveillance suffers from degradation due to poor illumination and arduous human annotations. It is challengable and remains a security risk at night. Existing methods rely on multi-spectral images to perceive objects in the dark,…

计算机视觉与模式识别 · 计算机科学 2023-07-12 Guanzhou Lan , Bin Zhao , Xuelong Li

Existing deep learning-based object detection models perform well under daytime conditions but face significant challenges at night, primarily because they are predominantly trained on daytime images. Additionally, training with nighttime…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Yunxiang Yang , Hao Zhen , Yongcan Huang , Jidong J. Yang

The significance of background information is frequently overlooked in contemporary research concerning channel attention mechanisms. This study addresses the issue of suboptimal single-spectral nighttime pedestrian detection performance…

计算机视觉与模式识别 · 计算机科学 2024-08-09 He Yao , Yongjun Zhang , Huachun Jian , Li Zhang , Ruzhong Cheng

Efficiently capturing the complex spatiotemporal representations from large-scale unlabeled traffic data remains to be a challenging task. In considering of the dilemma, this work employs the advanced contrastive learning and proposes a…

机器学习 · 计算机科学 2023-12-19 Lincan Li , Kaixiang Yang , Fengji Luo , Jichao Bi

This paper investigates the impact of various data augmentation techniques on the performance of object detection models. Specifically, we explore classical augmentation methods, image compositing, and advanced generative models such as…

计算机视觉与模式识别 · 计算机科学 2025-02-20 Ang Jia Ning Shermaine , Michalis Lazarou , Tania Stathaki

In recent years, image and video surveillance have made considerable progresses to the Intelligent Transportation Systems (ITS) with the help of deep Convolutional Neural Networks (CNNs). As one of the state-of-the-art perception…

计算机视觉与模式识别 · 计算机科学 2021-05-12 Lan Fu , Hongkai Yu , Felix Juefei-Xu , Jinlong Li , Qing Guo , Song Wang

While the diffusion transformer (DiT) has become a focal point of interest in recent years, its application in low-light image enhancement remains a blank area for exploration. Current methods recover the details from low-light images while…

计算机视觉与模式识别 · 计算机科学 2026-01-14 Xiangchen Yin , Zhenda Yu , Longtao Jiang , Xin Gao , Xiao Sun , Zhi Liu , Xun Yang

Pedestrian detection in intelligent transportation systems has made significant progress but faces two critical challenges: (1) insufficient fusion of complementary information between visible and infrared spectra, particularly in complex…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Rui Zhao , Zeyu Zhang , Yi Xu , Yi Yao , Yan Huang , Wenxin Zhang , Zirui Song , Xiuying Chen , Yang Zhao

Diffusion large language models (D-LLMs) have emerged as a promising alternative to auto-regressive models due to their iterative refinement capabilities. However, hallucinations remain a critical issue that hinders their reliability. To…

计算与语言 · 计算机科学 2026-03-18 Yanyu Qian , Yue Tan , Yixin Liu , Wang Yu , Shirui Pan

Existing paradigms for inferring pedestrian crossing behavior, ranging from statistical models to supervised learning methods, demonstrate limited generalizability and perform inadequately on new sites. Recent advances in Large Language…

人工智能 · 计算机科学 2026-01-05 Qingwen Pu , Kun Xie , Hong Yang , Guocong Zhai

We propose a diffusion distillation method that achieves new state-of-the-art in one-step/few-step 1024px text-to-image generation based on SDXL. Our method combines progressive and adversarial distillation to achieve a balance between…

计算机视觉与模式识别 · 计算机科学 2024-03-05 Shanchuan Lin , Anran Wang , Xiao Yang

We propose a method that augments a simulated dataset using diffusion models to improve the performance of pedestrian detection in real-world data. The high cost of collecting and annotating data in the real-world has motivated the use of…

计算机视觉与模式识别 · 计算机科学 2023-05-17 Andrew Farley , Mohsen Zand , Michael Greenspan

Nighttime light (NTL) remote sensing observation serves as a unique proxy for quantitatively assessing progress toward meeting a series of Sustainable Development Goals (SDGs), such as poverty estimation, urban sustainable development, and…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Lixian Zhang , Runmin Dong , Shuai Yuan , Jinxiao Zhang , Mengxuan Chen , Juepeng Zheng , Haohuan Fu

Recently, skeleton-based human action has become a hot research topic because the compact representation of human skeletons brings new blood to this research domain. As a result, researchers began to notice the importance of using RGB or…

计算机视觉与模式识别 · 计算机科学 2023-07-26 Yifan Jiang , Han Chen , Hanseok Ko

In complex scenes and varied conditions, effectively integrating spatial-temporal context is crucial for accurately identifying changes. However, current RS-CD methods lack a balanced consideration of performance and efficiency. CNNs lack…

计算机视觉与模式识别 · 计算机科学 2025-04-18 Zhenkai Wu , Xiaowen Ma , Rongrong Lian , Kai Zheng , Wei Zhang

In this paper we propose a method for improving pedestrian detection in the thermal domain using two stages: first, a generative data augmentation approach is used, then a domain adaptation method using generated data adapts an RGB…

计算机视觉与模式识别 · 计算机科学 2021-02-04 My Kieu , Lorenzo Berlincioni , Leonardo Galteri , Marco Bertini , Andrew D. Bagdanov , Alberto Del Bimbo

The performance of nighttime semantic segmentation is restricted by the poor illumination and a lack of pixel-wise annotation, which severely limit its application in autonomous driving. Existing works, e.g., using the twilight as the…

计算机视觉与模式识别 · 计算机科学 2022-05-03 Huan Gao , Jichang Guo , Guoli Wang , Qian Zhang

Recent advances in text-to-image diffusion models, particularly Stable Diffusion, have enabled the generation of highly detailed and semantically rich images. However, personalizing these models to represent novel subjects based on a few…

计算机视觉与模式识别 · 计算机科学 2025-05-19 Amritanshu Tiwari , Cherish Puniani , Kaustubh Sharma , Ojasva Nema

The recent emergence of latent diffusion models such as SDXL and SD 1.5 has shown significant capability in generating highly detailed and realistic images. Despite their remarkable ability to produce images, generating accurate text within…

计算机视觉与模式识别 · 计算机科学 2024-10-24 Jun Young Koh , Sang Hyun Park , Joy Song
‹ 上一页 1 2 3 10 下一页 ›