English
Related papers

Related papers: Disentangled Latent Energy-Based Style Translation…

200 papers

Scene Text Image Super-resolution (STISR) has recently achieved great success as a preprocessing method for scene text recognition. STISR aims to transform blurred and noisy low-resolution (LR) text images in real-world settings into clear…

Computer Vision and Pattern Recognition · Computer Science 2023-12-25 Chihiro Noguchi , Shun Fukuda , Masao Yamanaka

Auto-regressive speech-text models pre-trained on interleaved text tokens and discretized speech tokens demonstrate strong speech understanding and generation, yet remain substantially less compute-efficient than text LLMs, partly due to…

Computation and Language · Computer Science 2026-03-11 Yen-Ju Lu , Yashesh Gaur , Wei Zhou , Benjamin Muller , Jesus Villalba , Najim Dehak , Luke Zettlemoyer , Gargi Ghosh , Mike Lewis , Srinivasan Iyer , Duc Le

Machine learning analysis of longitudinal neuroimaging data is typically based on supervised learning, which requires a large number of ground-truth labels to be informative. As ground-truth labels are often missing or expensive to obtain…

Machine Learning · Computer Science 2021-06-29 Qingyu Zhao , Zixuan Liu , Ehsan Adeli , Kilian M. Pohl

Multisequence Magnetic Resonance Imaging (MRI) provides a more reliable diagnosis in clinical applications through complementary information across sequences. However, in practice, the absence of certain MR sequences is a common problem…

Image and Video Processing · Electrical Eng. & Systems 2025-10-21 Jihoon Cho , Jonghye Woo , Jinah Park

Multimodal 3D MRI brain tumor segmentation is a pivotal step in radiotherapy target delineation, surgical planning and post-treatment assessment. Existing methods often assume artifact-free MRI images. However, inevitable patient motion…

Image and Video Processing · Electrical Eng. & Systems 2026-05-18 Yuchun Wang , Xiaosong Li , Gefei Liang , Yang Liu

Generative models based on deep learning have shown significant potential in medical imaging, particularly for modality transformation and multimodal fusion in MRI-based brain imaging. This study introduces GM-LDM, a novel framework that…

Image and Video Processing · Electrical Eng. & Systems 2025-06-17 Hu Xu , Yang Jingling , Jia Sihan , Bi Yuda , Calhoun Vince

Deep learning methods, including Convolutional Neural Networks, Transformers and Mamba, have achieved remarkable success in hyperspectral image (HSI) classification. Nevertheless, existing methods exhibit inflexible integration of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 Jiawen Wen , Suixuan Qiu , Zihang Luo , Xiaofei Yang , Haotian Shi

We present HyperNST; a neural style transfer (NST) technique for the artistic stylization of images, based on Hyper-networks and the StyleGAN2 architecture. Our contribution is a novel method for inducing style transfer parameterized by a…

Computer Vision and Pattern Recognition · Computer Science 2022-08-10 Dan Ruta , Andrew Gilbert , Saeid Motiian , Baldo Faieta , Zhe Lin , John Collomosse

Dynamic magnetic resonance imaging (DMRI) is an effective imaging tool for diagnosis tasks that require motion tracking of a certain anatomy. To speed up DMRI acquisition, k-space measurements are commonly undersampled along spatial or…

Image and Video Processing · Electrical Eng. & Systems 2023-09-20 Di Xu , Hengjie Liu , Dan Ruan , Ke Sheng

Vision-and-language (V-L) tasks require the system to understand both vision content and natural language, thus learning fine-grained joint representations of vision and language (a.k.a. V-L representations) is of paramount importance.…

Computer Vision and Pattern Recognition · Computer Science 2022-11-01 Fenglin Liu , Xian Wu , Shen Ge , Xuancheng Ren , Wei Fan , Xu Sun , Yuexian Zou

We propose a novel approach for multi-modal Image-to-image (I2I) translation. To tackle the one-to-many relationship between input and output domains, previous works use complex training objectives to learn a latent embedding, jointly with…

Computer Vision and Pattern Recognition · Computer Science 2021-04-16 Moustafa Meshry , Yixuan Ren , Larry S Davis , Abhinav Shrivastava

Recent neural style transfer frameworks have obtained astonishing visual quality and flexibility in Single-style Transfer (SST), but little attention has been paid to Multi-style Transfer (MST) which refers to simultaneously transferring…

Computer Vision and Pattern Recognition · Computer Science 2019-10-30 Zixuan Huang , Jinghuai Zhang , Jing Liao

In MRI, images of the same contrast (e.g., T$_1$) from the same subject can exhibit noticeable differences when acquired using different hardware, sequences, or scan parameters. These differences in images create a domain gap that needs to…

Image and Video Processing · Electrical Eng. & Systems 2023-08-17 Hwihun Jeong , Heejoon Byun , Dong Un Kang , Jongho Lee

StyleGANs have shown impressive results on data generation and manipulation in recent years, thanks to its disentangled style latent space. A lot of efforts have been made in inverting a pretrained generator, where an encoder is trained ad…

Computer Vision and Pattern Recognition · Computer Science 2021-10-19 Ligong Han , Sri Harsha Musunuri , Martin Renqiang Min , Ruijiang Gao , Yu Tian , Dimitris Metaxas

Magnetic resonance imaging (MRI) is a non-invasive imaging modality and provides comprehensive anatomical and functional insights into the human body. However, its long acquisition times can lead to patient discomfort, motion artifacts, and…

Computer Vision and Pattern Recognition · Computer Science 2025-07-21 Mojtaba Safari , Zach Eidex , Chih-Wei Chang , Richard L. J. Qiu , Xiaofeng Yang

Understanding the underlying relationship between tongue and oropharyngeal muscle deformation seen in tagged-MRI and intelligible speech plays an important role in advancing speech motor control theories and treatment of speech…

Despite the suitability of graphs for capturing the relational structures inherent in architectural layout designs, there is a notable dearth of research on interpreting architectural design space using graph-based representation learning…

Machine Learning · Computer Science 2024-06-26 Jielin Chen , Rudi Stouffs

Designing generative models for 3D structural brain MRI that synthesize morphologically-plausible and attribute-specific (e.g., age, sex, disease state) samples is an active area of research. Existing approaches based on frameworks like…

Image and Video Processing · Electrical Eng. & Systems 2025-08-04 Alan Q. Wang , Fangrui Huang , Bailey Trang , Wei Peng , Mohammad Abbasi , Kilian Pohl , Mert Sabuncu , Ehsan Adeli

Self-supervised learning (SSL) and diffusion models have advanced representation learning and image synthesis, but in 3D medical imaging they are still largely used separately for analysis and synthesis, respectively. Unifying them is…

Image and Video Processing · Electrical Eng. & Systems 2026-04-07 Junkai Liu , Ling Shao , Le Zhang

22. Shortening acquisition time and reducing the motion-artifact are two of the most critical issues in MRI. As a promising solution, high-quality MRI image restoration provides a new approach to achieve higher resolution without costing…

Image and Video Processing · Electrical Eng. & Systems 2021-02-02 Hao Li , Jianan Liu
‹ Prev 1 4 5 6 7 8 10 Next ›