中文
相关论文

相关论文: CHIMLE: Conditional Hierarchical IMLE for Multimod…

200 篇论文

Text-to-image synthesis has made encouraging progress and attracted lots of public attention recently. However, popular evaluation metrics in this area, like the Inception Score and Fr'echet Inception Distance, incur several issues. First…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Qi Chen , Chaorui Deng , Zixiong Huang , Bowen Zhang , Mingkui Tan , Qi Wu

High-resolution satellite imagery has proven useful for a broad range of tasks, including measurement of global human population, local economic livelihoods, and biodiversity, among many others. Unfortunately, high-resolution imagery is…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Yutong He , Dingjie Wang , Nicholas Lai , William Zhang , Chenlin Meng , Marshall Burke , David B. Lobell , Stefano Ermon

This paper presents a novel method to deal with the challenging task of generating photographic images conditioned on semantic image descriptions. Our method introduces accompanying hierarchical-nested adversarial objectives inside the…

计算机视觉与模式识别 · 计算机科学 2018-04-10 Zizhao Zhang , Yuanpu Xie , Lin Yang

Multiplex imaging is revolutionizing pathology by enabling the simultaneous visualization of multiple biomarkers within tissue samples, providing molecular-level insights that traditional hematoxylin and eosin (H&E) staining cannot provide.…

图像与视频处理 · 电气工程与系统科学 2026-01-06 Hyun-Jic Oh , Junsik Kim , Zhiyi Shi , Yichen Wu , Yu-An Chen , Peter K Sorger , Hanspeter Pfister , Won-Ki Jeong

Multi-modal image registration spatially aligns two images with different distributions. One of its major challenges is that images acquired from different imaging machines have different imaging distributions, making it difficult to focus…

计算机视觉与模式识别 · 计算机科学 2023-03-03 Lingke Kong , X. Sharon Qi , Qijin Shen , Jiacheng Wang , Jingyi Zhang , Yanle Hu , Qichao Zhou

Transfer learning from large-scale pre-trained models has become essential for many computer vision tasks. Recent studies have shown that datasets like ImageNet are weakly labeled since images with multiple object classes present are…

计算机视觉与模式识别 · 计算机科学 2021-11-25 Sai Rajeswar , Pau Rodriguez , Soumye Singhal , David Vazquez , Aaron Courville

Multi-subject image generation aims to synthesize images that faithfully preserve the identities of multiple reference subjects while following textual instructions. However, existing methods often suffer from identity inconsistency and…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Yijia Xu , Zihao Wang , Jinshi Cui

Recent diffusion model advancements have enabled high-fidelity images to be generated using text prompts. However, a domain gap exists between generated images and real-world images, which poses a challenge in generating high-quality…

计算机视觉与模式识别 · 计算机科学 2023-11-08 Yuechen Zhang , Jinbo Xing , Eric Lo , Jiaya Jia

Label-free cell classification is advantageous for supplying pristine cells for further use or examination, yet existing techniques frequently fall short in terms of specificity and speed. In this study, we address these limitations through…

图像与视频处理 · 电气工程与系统科学 2025-02-25 Khayrul Islam , Ratul Paul , Shen Wang , Yuwen Zhao , Partho Adhikary , Qiying Li , Xiaochen Qin , Yaling Liu

In medical imaging, access to data is commonly limited due to patient privacy restrictions and the issue that it can be difficult to acquire enough data in the case of rare diseases.[1] The purpose of this investigation was to develop a…

计算机视觉与模式识别 · 计算机科学 2024-03-29 John R. McNulty , Lee Kho , Alexandria L. Case , Charlie Fornaca , Drew Johnston , David Slater , Joshua M. Abzug , Sybil A. Russell

A good Text-to-Image model should not only generate high quality images, but also ensure the consistency between the text and the generated image. Previous models failed to simultaneously fix both sides well. This paper proposes a Gradual…

计算机视觉与模式识别 · 计算机科学 2022-06-22 Bo Yang , Fangxiang Feng , Xiaojie Wang

Multimode fibers (MMFs) provide a compact, high-throughput platform for minimally invasive imaging and information transmission. However, their utility is fundamentally constrained by mode mixing, which renders image transmission spatially…

Utility companies increasingly rely on drone imagery for post-event and routine inspection, but training accurate defect-type classifiers remains difficult because defect examples are rare and inspection datasets are often limited or…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Xuesong Wang , Caisheng Wang

Image and video generative models that are pre-trained on Internet-scale data can greatly increase the generalization capacity of robot learning systems. These models can function as high-level planners, generating intermediate subgoals for…

Conditional image generation has gained significant attention for its ability to personalize content. However, the field faces challenges in developing task-agnostic, reliable, and explainable evaluation metrics. This paper introduces…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Jifang Wang , Xue Yang , Longyue Wang , Zhenran Xu , Yiyu Wang , Yaowei Wang , Weihua Luo , Kaifu Zhang , Baotian Hu , Min Zhang

Conditional image synthesis for generating photorealistic images serves various applications for content editing to content generation. Previous conditional image synthesis algorithms mostly rely on semantic maps, and often fail in complex…

计算机视觉与模式识别 · 计算机科学 2020-04-23 Aysegul Dundar , Karan Sapra , Guilin Liu , Andrew Tao , Bryan Catanzaro

Low-light image enhancement, such as recovering color and texture details from low-light images, is a complex and vital task. For automated driving, low-light scenarios will have serious implications for vision-based applications. To…

图像与视频处理 · 电气工程与系统科学 2021-09-01 Yangyang Qu , Kai Chen , Chao Liu , Yongsheng Ou

Few-shot image generation aims to train generative models using a small number of training images. When there are few images available for training (e.g. 10 images), Learning From Scratch (LFS) methods often generate images that closely…

计算机视觉与模式识别 · 计算机科学 2023-11-15 Ziqiang Li , Chaoyue Wang , Xue Rui , Chao Xue , Jiaxu Leng , Bin Li

In this paper, we propose an instance similarity learning (ISL) method for unsupervised feature representation. Conventional methods assign close instance pairs in the feature space with high similarity, which usually leads to wrong…

计算机视觉与模式识别 · 计算机科学 2021-08-06 Ziwei Wang , Yunsong Wang , Ziyi Wu , Jiwen Lu , Jie Zhou

Large language models (LLMs) have emerged as a powerful foundation for intelligent reasoning and decision-making, demonstrating substantial impact across a wide range of domains and applications. However, their massive parameter scales and…

分布式、并行与集群计算 · 计算机科学 2025-12-29 Mingyu Sun , Xiao Zhang , Shen Qu , Yan Li , Mengbai Xiao , Yuan Yuan , Dongxiao Yu