English
Related papers

Related papers: INCLG: Inpainting for Non-Cleft Lip Generation wit…

200 papers

Image inpainting techniques have shown significant improvements by using deep neural networks recently. However, most of them may either fail to reconstruct reasonable structures or restore fine-grained textures. In order to solve this…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Yurui Ren , Xiaoming Yu , Ruonan Zhang , Thomas H. Li , Shan Liu , Ge Li

Fine-tuning pretrained language models (LMs) without making any architectural changes has become a norm for learning various language downstream tasks. However, for non-language downstream tasks, a common practice is to employ task-specific…

Face recognition systems are increasingly vulnerable to morphing attacks, where a composite image is crafted to match multiple identities, enabling unauthorized access and identity fraud. Existing detection methods identify morphed images…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Nitish Shukla , Arun Ross

Image inpainting aims at restoring missing regions of corrupted images, which has many applications such as image restoration and object removal. However, current GAN-based generative inpainting models do not explicitly exploit the…

Computer Vision and Pattern Recognition · Computer Science 2020-09-15 Ang Li , Jianzhong Qi , Rui Zhang , Xingjun Ma , Kotagiri Ramamohanarao

Multimodal Large Language Models (MLLMs), such as GPT4o, have shown strong capabilities in visual reasoning and explanation generation. However, despite these strengths, they face significant challenges in the increasingly critical task of…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Fanrui Zhang , Jiawei Liu , Jiaying Zhu , Esther Sun , Dong Li , Qiang Zhang , Zheng-Jun Zha

Inverse graphics -- the task of inverting an image into physical variables that, when rendered, enable reproduction of the observed scene -- is a fundamental challenge in computer vision and graphics. Successfully disentangling an image…

Computer Vision and Pattern Recognition · Computer Science 2024-08-27 Peter Kulits , Haiwen Feng , Weiyang Liu , Victoria Abrevaya , Michael J. Black

Prior studies have made significant progress in image inpainting guided by either text description or subject image. However, the research on inpainting with flexible guidance or control, i.e., text-only, image-only, and their combination,…

Computer Vision and Pattern Recognition · Computer Science 2025-01-23 Yulin Pan , Chaojie Mao , Zeyinzi Jiang , Zhen Han , Jingfeng Zhang , Xiangteng He

We consider machine-learning-based malignancy prediction and lesion identification from clinical dermatological images, which can be indistinctly acquired via smartphone or dermoscopy capture. Additionally, we do not assume that images…

Computer Vision and Pattern Recognition · Computer Science 2021-04-07 Meng Xia , Meenal K. Kheterpal , Samantha C. Wong , Christine Park , William Ratliff , Lawrence Carin , Ricardo Henao

Constructing dataset for fashion style recognition is challenging due to the inherent subjectivity and ambiguity of style concepts. Recent advances in text-to-image models have facilitated generative data augmentation by synthesizing images…

Computer Vision and Pattern Recognition · Computer Science 2026-05-08 Yuki Hirakawa , Ryotaro Shimizu

Multilingual image captioning has recently been tackled by training with large-scale machine translated data, which is an expensive, noisy, and time-consuming process. Without requiring any multilingual caption data, we propose LMCap, an…

Computation and Language · Computer Science 2023-06-01 Rita Ramos , Bruno Martins , Desmond Elliott

In-context learning allows adapting a model to new tasks given a task description at test time. In this paper, we present IMProv - a generative model that is able to in-context learn visual tasks from multimodal prompts. Given a textual…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Jiarui Xu , Yossi Gandelsman , Amir Bar , Jianwei Yang , Jianfeng Gao , Trevor Darrell , Xiaolong Wang

Numerous image processing techniques (IPTs) have been employed to detect crack defects, offering an alternative to human-conducted onsite inspections. These IPTs manipulate images to extract defect features, particularly cracks in surfaces…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 Mohsen Asghari Ilani , Leila Amini , Hossein Karimi , Maryam Shavali Kuhshuri

Thanks to the powerful language comprehension capabilities of Large Language Models (LLMs), existing instruction-based image editing methods have introduced Multimodal Large Language Models (MLLMs) to promote information exchange between…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Yujie Hu , Zecheng Tang , Xu Jiang , Weiqi Li , Jian Zhang

This paper examines the limitations of advanced text-to-image models in accurately rendering unconventional concepts which are scarcely represented or absent in their training datasets. We identify how these limitations not only confine the…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Jiyoon Myung , Jihyeon Park

Facial images have extensive practical applications. Although the current large-scale text-image diffusion models exhibit strong generation capabilities, it is challenging to generate the desired facial images using only text prompt. Image…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Dawei Dai , Mingming Jia , Yinxiu Zhou , Hang Xing , Chenghang Li

Tumor segmentation plays a critical role in histopathology, but it requires costly, fine-grained image-mask pairs annotated by pathologists. Thus, synthesizing histopathology data to expand the dataset is highly desirable. Previous works…

Computer Vision and Pattern Recognition · Computer Science 2025-03-07 Hong Liu , Haosen Yang , Evi M. C. Huijben , Mark Schuiveling , Ruisheng Su , Josien P. W. Pluim , Mitko Veta

In the image inpainting task, the ability to repair both high-frequency and low-frequency information in the missing regions has a substantial influence on the quality of the restored image. However, existing inpainting methods usually fail…

Computer Vision and Pattern Recognition · Computer Science 2020-06-12 Huali Xu , Xiangdong Su , Meng Wang , Xiang Hao , Guanglai Gao

A myriad of algorithms for the automatic analysis of brain MR images is available to support clinicians in their decision-making. For brain tumor patients, the image acquisition time series typically starts with an already pathological…

Image and Video Processing · Electrical Eng. & Systems 2024-09-24 Florian Kofler , Felix Meissen , Felix Steinbauer , Robert Graf , Stefan K Ehrlich , Annika Reinke , Eva Oswald , Diana Waldmannstetter , Florian Hoelzl , Izabela Horvath , Oezguen Turgut , Suprosanna Shit , Christina Bukas , Kaiyuan Yang , Johannes C. Paetzold , Ezequiel de da Rosa , Isra Mekki , Shankeeth Vinayahalingam , Hasan Kassem , Juexin Zhang , Ke Chen , Ying Weng , Alicia Durrer , Philippe C. Cattin , Julia Wolleb , M. S. Sadique , M. M. Rahman , W. Farzana , A. Temtam , K. M. Iftekharuddin , Maruf Adewole , Syed Muhammad Anwar , Ujjwal Baid , Anastasia Janas , Anahita Fathi Kazerooni , Dominic LaBella , Hongwei Bran Li , Ahmed W Moawad , Gian-Marco Conte , Keyvan Farahani , James Eddy , Micah Sheller , Sarthak Pati , Alexandros Karagyris , Alejandro Aristizabal , Timothy Bergquist , Verena Chung , Russell Takeshi Shinohara , Farouk Dako , Walter Wiggins , Zachary Reitman , Chunhao Wang , Xinyang Liu , Zhifan Jiang , Elaine Johanson , Zeke Meier , Ariana Familiar , Christos Davatzikos , John Freymann , Justin Kirby , Michel Bilello , Hassan M Fathallah-Shaykh , Roland Wiest , Jan Kirschke , Rivka R Colen , Aikaterini Kotrotsou , Pamela Lamontagne , Daniel Marcus , Mikhail Milchenko , Arash Nazeri , Marc-André Weber , Abhishek Mahajan , Suyash Mohan , John Mongan , Christopher Hess , Soonmee Cha , Javier Villanueva-Meyer , Errol Colak , Priscila Crivellaro , Andras Jakab , Abiodun Fatade , Olubukola Omidiji , Rachel Akinola Lagos , O O Olatunji , Goldey Khanna , John Kirkpatrick , Michelle Alonso-Basanta , Arif Rashid , Miriam Bornhorst , Ali Nabavizadeh , Natasha Lepore , Joshua Palmer , Antonio Porras , Jake Albrecht , Udunna Anazodo , Mariam Aboian , Evan Calabrese , Jeffrey David Rudie , Marius George Linguraru , Juan Eugenio Iglesias , Koen Van Leemput , Spyridon Bakas , Benedikt Wiestler , Ivan Ezhov , Marie Piraud , Bjoern H Menze

Compared to the prosperity of pre-training models in natural image understanding, the research on large-scale pre-training models for facial knowledge learning is still limited. Current approaches mainly rely on manually assembled and…

Computer Vision and Pattern Recognition · Computer Science 2025-10-22 Yudong Li , Hao Li , Xianxu Hou , Linlin Shen

Deep learning (DL) has demonstrated its powerful capabilities in the field of image inpainting. The DL-based image inpainting approaches can produce visually plausible results, but often generate various unpleasant artifacts, especially in…

Computer Vision and Pattern Recognition · Computer Science 2020-09-03 Haiwei Wu , Jiantao Zhou , Yuanman Li