中文
相关论文

相关论文: SynthVision -- Harnessing Minimal Input for Maxima…

200 篇论文

Deep learning models have emerged as a powerful tool for various medical applications. However, their success depends on large, high-quality datasets that are challenging to obtain due to privacy concerns and costly annotation. Generative…

计算机视觉与模式识别 · 计算机科学 2025-03-04 Milad Yazdani , Yasamin Medghalchi , Pooria Ashrafian , Ilker Hacihaliloglu , Dena Shahriari

Generating synthetic datasets for training face recognition models is challenging because dataset generation entails more than creating high fidelity images. It involves generating multiple images of same subjects under different factors…

计算机视觉与模式识别 · 计算机科学 2023-04-17 Minchul Kim , Feng Liu , Anil Jain , Xiaoming Liu

Current synthetic data pipelines for computer vision generate images without diagnosing what the downstream model actually needs. This open-loop paradigm treats synthetic data as cheap real data, randomly sampling the generator's output…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Krisanu Sarkar

The widespread diffusion of synthetically generated content is a serious threat that needs urgent countermeasures. As a matter of fact, the generation of synthetic content is not restricted to multimedia data like videos, photographs or…

Generating novel views of an object from a single image is a challenging task. It requires an understanding of the underlying 3D structure of the object from an image and rendering high-quality, spatially consistent new views. While recent…

计算机视觉与模式识别 · 计算机科学 2023-12-05 Jeong-gi Kwak , Erqun Dong , Yuhe Jin , Hanseok Ko , Shweta Mahajan , Kwang Moo Yi

In the field of deep learning applied to face recognition, securing large-scale, high-quality datasets is vital for attaining precise and reliable results. However, amassing significant volumes of high-quality real data faces hurdles such…

计算机视觉与模式识别 · 计算机科学 2023-05-18 Omer Granoviter , Alexey Gruzdev , Vladimir Loginov , Max Kogan , Orly Zvitia

Image-based Virtual Try-On (VITON) aims to transfer an in-shop garment image onto a target person. While existing methods focus on warping the garment to fit the body pose, they often overlook the synthesis quality around the garment-skin…

计算机视觉与模式识别 · 计算机科学 2023-12-07 xujie zhang , Xiu Li , Michael Kampffmeyer , Xin Dong , Zhenyu Xie , Feida Zhu , Haoye Dong , Xiaodan Liang

Synthetic generation of three-dimensional cell models from histopathological images aims to enhance understanding of cell mutation, and progression of cancer, necessary for clinical assessment and optimal treatment. Classical reconstruction…

图像与视频处理 · 电气工程与系统科学 2021-02-09 Yoav Alon , Xiang Yu , Huiyu Zhou

Advancements in deep image synthesis techniques, such as generative adversarial networks (GANs) and diffusion models (DMs), have ushered in an era of generating highly realistic images. While this technological progress has captured…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Mamadou Keita , Wassim Hamidouche , Hessen Bougueffa Eutamene , Abdenour Hadid , Abdelmalik Taleb-Ahmed

Melanoma is the most lethal form of skin cancer, and early detection is critical for improving patient outcomes. Although dermoscopy combined with deep learning has advanced automated skin-lesion analysis, progress is hindered by limited…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Pei-Yu Lin , Yidan Shen , Neville Mathew , Renjie Hu , Siyu Huang , Courtney M. Queen , Cameron E. West , Ana Ciurea , George Zouridakis

State-of-the-art face recognition models show impressive accuracy, achieving over 99.8% on Labeled Faces in the Wild (LFW) dataset. Such models are trained on large-scale datasets that contain millions of real human face images collected…

计算机视觉与模式识别 · 计算机科学 2022-10-07 Gwangbin Bae , Martin de La Gorce , Tadas Baltrusaitis , Charlie Hewitt , Dong Chen , Julien Valentin , Roberto Cipolla , Jingjing Shen

Artificial intelligence and machine learning techniques have the promise to revolutionize the field of digital pathology. However, these models demand considerable amounts of data, while the availability of unbiased training data is…

图像与视频处理 · 电气工程与系统科学 2023-02-14 Nati Daniel , Eliel Aknin , Ariel Larey , Yoni Peretz , Guy Sela , Yael Fisher , Yonatan Savir

Medical image segmentation models struggle with rare abnormalities due to scarce annotated pathological data. We propose DiffAug a novel framework that combines textguided diffusion-based generation with automatic segmentation validation to…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Maham Nazir , Muhammad Aqeel , Francesco Setti

The development of accurate medical image classification models is often constrained by privacy concerns and data scarcity for certain conditions, leading to small and imbalanced datasets. To address these limitations, this study explores…

计算机视觉与模式识别 · 计算机科学 2025-10-30 Iman Khazrak , Shakhnoza Takhirova , Mostafa M. Rezaee , Mehrdad Yadollahi , Robert C. Green , Shuteng Niu

The task of novel view synthesis aims to generate unseen perspectives of an object or scene from a limited set of input images. Nevertheless, synthesizing novel views from a single image still remains a significant challenge in the realm of…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Yifan Jiang , Hao Tang , Jen-Hao Rick Chang , Liangchen Song , Zhangyang Wang , Liangliang Cao

Predictive maintenance has been used to optimize system repairs in the industrial, medical, and financial domains. This technique relies on the consistent ability to detect and predict anomalies in critical systems. AI models have been…

Deep learning methods, and in particular convolutional neural networks (CNNs), have led to an enormous breakthrough in a wide range of computer vision tasks, primarily by using large-scale annotated datasets. However, obtaining such…

计算机视觉与模式识别 · 计算机科学 2019-01-23 Maayan Frid-Adar , Idit Diamant , Eyal Klang , Michal Amitai , Jacob Goldberger , Hayit Greenspan

Novel-view synthesis techniques achieve impressive results for static scenes but struggle when faced with the inconsistencies inherent to casual capture settings: varying illumination, scene motion, and other unintended effects that are…

We propose a virtual staining methodology based on Generative Adversarial Networks to map histopathology images of breast cancer tissue from H&E stain to PHH3 and vice versa. We use the resulting synthetic images to build Convolutional…

图像与视频处理 · 电气工程与系统科学 2020-03-18 Caner Mercan , Germonda Reijnen-Mooij , David Tellez Martin , Johannes Lotz , Nick Weiss , Marcel van Gerven , Francesco Ciompi

X-ray imaging is a rapid and cost-effective tool for visualizing internal human anatomy. While multi-view X-ray imaging provides complementary information that enhances diagnosis, intervention, and education, acquiring images from multiple…

图像与视频处理 · 电气工程与系统科学 2025-10-21 Chun Xie , Yuichi Yoshii , Itaru Kitahara