English
Related papers

Related papers: ROI-based Deep Image Compression with Swin Transfo…

200 papers

Recently, Transformers have shown promising performance in various vision tasks. However, the high costs of global self-attention remain challenging for Transformers, especially for high-resolution vision tasks. Local self-attention runs…

Computer Vision and Pattern Recognition · Computer Science 2023-04-28 Zhemin Zhang , Xun Gong

Motivated by recent work on deep neural network (DNN)-based image compression methods showing potential improvements in image quality, savings in storage, and bandwidth reduction, we propose to perform image understanding tasks such as…

Computer Vision and Pattern Recognition · Computer Science 2018-03-19 Robert Torfason , Fabian Mentzer , Eirikur Agustsson , Michael Tschannen , Radu Timofte , Luc Van Gool

Deep networks can be trained to map images into a low-dimensional latent space. In many cases, different images in a collection are articulated versions of one another; for example, same object with different lighting, background, or pose.…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Rakib Hyder , M. Salman Asif

Transformers and their derivatives have achieved state-of-the-art performance across text, vision, and speech recognition tasks. However, minimal effort has been made to train transformers capable of evaluating the output quality of other…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Maxwell Meyer , Jack Spruyt

While MPEG-standardized video-based point cloud compression (VPCC) achieves high compression efficiency for human perception, it struggles with a poor trade-off between bitrate savings and detection accuracy when supporting 3D object…

Computer Vision and Pattern Recognition · Computer Science 2025-02-10 Mingxuan Yan , Ruijie Zhang , Xuedou Xiao , Wei Wang

In this paper, we propose a novel fragile block based medical image watermarking technique for embedding data of patient into medical image, verifying the integrity of ROI (Region of Interest), detecting the tampered blocks inside ROI and…

Multimedia · Computer Science 2014-12-22 Eswaraiah Rayachoti , Sreenivasa Reddy Edara

Magnetic resonance imaging (MRI) is a widely used non-radiative and non-invasive method for clinical interrogation of organ structures and metabolism, with an inherently long scanning time. Methods by k-space undersampling and deep learning…

Image and Video Processing · Electrical Eng. & Systems 2022-04-05 Jiahao Huang , Yinzhe Wu , Huanjun Wu , Guang Yang

Conventional image inpainting techniques typically process entire images, which often leads to computational inefficiency and susceptibility to information redundancy, particularly in occluded or cluttered scenes. Inspired by cortical…

Pattern Formation and Solitons · Physics 2025-10-21 Yonghao Wu , Chang Liu , Vladimir Filaretov , Dmitry Yukhimets

The Swin transformer has recently attracted attention in medical image analysis due to its computational efficiency and long-range modeling capability. Owing to these properties, the Swin Transformer is suitable for establishing more…

Computer Vision and Pattern Recognition · Computer Science 2024-08-21 Mingrui Ma , Tao Wang , Lei Song , Weijie Wang , Guixia Liu

While some studies have proven that Swin Transformer (Swin) with window self-attention (WSA) is suitable for single image super-resolution (SR), the plain WSA ignores the broad regions when reconstructing high-resolution images due to a…

Computer Vision and Pattern Recognition · Computer Science 2023-03-21 Haram Choi , Jeongmin Lee , Jihoon Yang

ROI (Region of Interest) video selective encryption based on H.265/HEVC is a technology that protects the sensitive regions of videos by perturbing the syntax elements associated with target areas. However, existing methods typically adopt…

Image and Video Processing · Electrical Eng. & Systems 2026-04-10 Xiang Zhang , Haoyan Lu , Ziqiang Li , Ziwen He , Zhenshan Tan , Fei Peng , Zhangjie Fu

Segmentation models in automated optical inspection of wire-bonded semiconductors are typically device-specific and must be re-trained when new devices or distribution shifts appear. We introduce AOI-SSL, a training-efficient framework for…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Joaquín Figueira , Rob Van Gastel , Giacomo D'Amicantonio , Zhuoran Liu , Ioan Gabriel Bucur , Faysal Boughorbel , Egor Bondarev

Driver distraction behavior recognition using in-vehicle cameras demands real-time inference on edge devices. However, lightweight models often fail to capture fine-grained behavioral cues, resulting in reduced performance on unseen drivers…

Computer Vision and Pattern Recognition · Computer Science 2025-12-11 Keito Inoshita

The emerging technology of snapshot compressive imaging (SCI) enables capturing high dimensional (HD) data in an efficient way. It is generally implemented by two components: an optical encoder that compresses HD signals into a 2D…

Image and Video Processing · Electrical Eng. & Systems 2022-02-03 Jiamian Wang , Yulun Zhang , Xin Yuan , Yun Fu , Zhiqiang Tao

Accurately segmenting roads is challenging due to substantial intra-class variations, indistinct inter-class distinctions, and occlusions caused by shadows, trees, and buildings. To address these challenges, attention to important texture…

Computer Vision and Pattern Recognition · Computer Science 2023-05-30 Tao Chen , Yiran Liu , Haoyu Jiang , Ruirui Li

Deep convolutional networks have attracted great attention in image restoration and enhancement. Generally, restoration quality has been improved by building more and more convolutional block. However, these methods mostly learn a specific…

Computer Vision and Pattern Recognition · Computer Science 2021-05-21 Yukai Shi , Jinghui Qin

Lossy image compression is generally formulated as a joint rate-distortion optimization to learn encoder, quantizer, and decoder. However, the quantizer is non-differentiable, and discrete entropy estimation usually is required for rate…

Computer Vision and Pattern Recognition · Computer Science 2017-09-20 Mu Li , Wangmeng Zuo , Shuhang Gu , Debin Zhao , David Zhang

In remote sensing images, complex backgrounds, weak object signals, and small object scales make accurate detection particularly challenging, especially under low-quality imaging conditions. A common strategy is to integrate single-image…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Ruo Qi , Linhui Dai , Yusong Qin , Chaolei Yang , Yanshan Li

Automated medical image captioning translates complex radiological images into diagnostic narratives that can support reporting workflows. We present a Swin-BART encoder-decoder system with a lightweight regional attention module that…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Zubia Naz , Farhan Asghar , Muhammad Ishfaq Hussain , Yahya Hadadi , Muhammad Aasim Rafique , Wookjin Choi , Moongu Jeon

\textbf{Purpose} This study aims to address the growing challenge of distinguishing computer-generated imagery (CGI) from authentic digital images in the RGB color space. Given the limitations of existing classification methods in handling…

Computer Vision and Pattern Recognition · Computer Science 2024-09-10 Preetu Mehta , Aman Sagar , Suchi Kumari