English
Related papers

Related papers: Archives, archival bond, and digital representatio…

200 papers

Multi-modality image fusion (MMIF) combines complementary information from different image modalities to provide a comprehensive and objective interpretation of scenes. However, existing fusion methods cannot resist different weather…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Xilai Li , Wuyang Liu , Xiaosong Li , Fuqiang Zhou , Huafeng Li , Feiping Nie

How to tackle non-iid data is a crucial topic in federated learning. This challenging problem not only affects training process, but also harms performance of clients not participating in training. Existing literature mainly focuses on…

Machine Learning · Computer Science 2023-02-17 Meirui Jiang , Xiaoxiao Li , Xiaofei Zhang , Michael Kamp , Qi Dou

Image-to-image translation architectures may have limited effectiveness in some circumstances. For example, while generating rainy scenarios, they may fail to model typical traits of rain as water drops, and this ultimately impacts the…

Computer Vision and Pattern Recognition · Computer Science 2020-03-17 Fabio Pizzati , Raoul de Charette , Michela Zaccaria , Pietro Cerri

Multimodal machine learning has gained significant attention in recent years due to its potential for integrating information from multiple modalities to enhance learning and decision-making processes. However, it is commonly observed that…

Machine Learning · Computer Science 2025-09-12 Sahiti Yerramilli , Jayant Sravan Tamarapalli , Jonathan Francis , Eric Nyberg

Many deep learning based automated medical image segmentation systems, in reality, face difficulties in deployment due to the cost of massive data annotation and high latency in model iteration. We propose a dynamic interactive learning…

Computer Vision and Pattern Recognition · Computer Science 2023-12-12 Mu Tian , Xiaohui Chen , Yi Gao

Recent years have witnessed a broader range of applications of image processing technologies in multiple industrial processes, such as smoke detection, security monitoring, and workpiece inspection. Different kinds of distortion types and…

Computer Vision and Pattern Recognition · Computer Science 2024-02-19 Xuanchao Ma , Yanlin Jiang , Hongyan Liu , Chengxu Zhou , Ke Gu

The digitisation of historical documents has provided historians with unprecedented research opportunities. Yet, the conventional approach to analysing historical documents involves converting them from images to text using OCR, a process…

Computation and Language · Computer Science 2023-11-07 Nadav Borenstein , Phillip Rust , Desmond Elliott , Isabelle Augenstein

We introduce Pixel-aligned Implicit Function (PIFu), a highly effective implicit representation that locally aligns pixels of 2D images with the global context of their corresponding 3D object. Using PIFu, we propose an end-to-end deep…

Computer Vision and Pattern Recognition · Computer Science 2019-12-05 Shunsuke Saito , Zeng Huang , Ryota Natsume , Shigeo Morishima , Angjoo Kanazawa , Hao Li

Multimodal medical image fusion plays an instrumental role in several areas of medical image processing, particularly in disease recognition and tumor detection. Traditional fusion methods tend to process each modality independently before…

Image and Video Processing · Electrical Eng. & Systems 2023-10-11 Lin Liu , Xinxin Fan , Chulong Zhang , Jingjing Dai , Yaoqin Xie , Xiaokun Liang

Articulated objects are prevalent in daily life. Interactable digital twins of such objects have numerous applications in embodied AI and robotics. Unfortunately, current methods to digitize articulated real-world objects require carefully…

Graphics · Computer Science 2025-11-18 Weikun Peng , Jun Lv , Cewu Lu , Manolis Savva

The appearance of histopathology images depends on tissue type, staining and digitization procedure. These vary from source to source and are the potential causes for domain-shift problems. Owing to this problem, despite the great success…

Computer Vision and Pattern Recognition · Computer Science 2022-08-24 Trinh Thi Le Vuong , Quoc Dang Vu , Mostafa Jahanifar , Simon Graham , Jin Tae Kwak , Nasir Rajpoot

While deep learning models like Vision Transformer (ViT) have achieved significant advances, they typically require large datasets. With data privacy regulations, access to many original datasets is restricted, especially medical images.…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Xinyuan Zhao , Yihang Wu , Ahmad Chaddad , Tareef Daqqaq , Reem Kateb

Current foundation models (FMs) rely on token representations that directly fragment continuous real-world multimodal data into discrete tokens. They limit FMs to learning real-world knowledge and relationships purely through statistical…

Machine Learning · Computer Science 2025-05-08 Yiqing Shen , Hao Ding , Lalithkumar Seenivasan , Tianmin Shu , Mathias Unberath

Multi-modality image fusion aims at fusing modality-specific (complementarity) and modality-shared (correlation) information from multiple source images. To tackle the problem of the neglect of inter-feature relationships, high-frequency…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Xiaoli Zhang , Liying Wang , Libo Zhao , Xiongfei Li , Siwei Ma

The method for image-to-point cloud registration typically determines the rigid transformation using a coarse-to-fine pipeline. However, directly and uniformly matching image patches with point cloud patches may lead to focusing on…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Zhixin Cheng , Jiacheng Deng , Xinjun Li , Baoqun Yin , Tianzhu Zhang

Recent advances in image-based 3D human shape estimation have been driven by the significant improvement in representation power afforded by deep neural networks. Although current approaches have demonstrated the potential in real world…

Computer Vision and Pattern Recognition · Computer Science 2020-04-02 Shunsuke Saito , Tomas Simon , Jason Saragih , Hanbyul Joo

In this paper we explore opportunities for the post-trade industry to standardize and simplify in order to significantly increase efficiency and reduce costs. We start by summarizing relevant industry problems (inconsistent processes,…

Computers and Society · Computer Science 2022-08-10 Aishwarya Nair , Lee Braine

As Joint Audio-Visual Generation Models see widespread commercial deployment, embedding watermarks has become essential for protecting vendor copyright and ensuring content provenance. However, existing techniques suffer from an…

Cryptography and Security · Computer Science 2026-03-10 Luyang Si , Leyi Pan , Lijie Wen

Typically, foundation models are hosted on cloud servers to meet the high demand for their services. However, this exposes them to security risks, as attackers can modify them after uploading to the cloud or transferring from a local…

Cryptography and Security · Computer Science 2023-05-18 Zhaoxia Yin , Heng Yin , Hang Su , Xinpeng Zhang , Zhenzhe Gao

Video signals are vulnerable in multimedia communication and storage systems, as even slight bitstream-domain corruption can lead to significant pixel-domain degradation. To recover faithful spatio-temporal content from corrupted inputs,…

Image and Video Processing · Electrical Eng. & Systems 2025-10-30 Tianyi Liu , Kejun Wu , Chen Cai , Yi Wang , Kim-Hui Yap , Lap-Pui Chau
‹ Prev 1 8 9 10 Next ›