中文
相关论文

相关论文: ROI-Packing: Efficient Region-Based Compression fo…

200 篇论文

Thoracic aortic dissection and aneurysms are the most lethal diseases of the aorta. The major hindrance to treatment lies in the accurate analysis of the medical images. More particularly, aortic segmentation of the 3D image is often…

图像与视频处理 · 电气工程与系统科学 2026-01-14 Loris Giordano , Ine Dirks , Tom Lenaerts , Jef Vandemeulebroucke

Efficient 3D LiDAR point cloud compression (LPCC) and streaming are critical for edge server-assisted robotic systems, enabling real-time communication with compact data representations. A widely adopted approach represents LiDAR point…

图像与视频处理 · 电气工程与系统科学 2026-03-17 Shengqian Wang , Chang Tu , He Chen

With the increasing use of neural network (NN)-based computer vision applications that process image and video data as input, interest has emerged in video compression technology optimized for computer vision tasks. In fact, given the…

计算机视觉与模式识别 · 计算机科学 2025-09-26 Hyomin Choi , Heeji Han , Chris Rosewarne , Fabien Racapé

Multi-resolution methods such as Adaptive Mesh Refinement (AMR) can enhance storage efficiency for HPC applications generating vast volumes of data. However, their applicability is limited and cannot be universally deployed across all…

分布式、并行与集群计算 · 计算机科学 2024-10-02 Daoce Wang , Pascal Grosset , Jesus Pulido , Tushar M. Athawale , Jiannan Tian , Kai Zhao , Zarija Lukić , Axel Huebl , Zhe Wang , James Ahrens , Dingwen Tao

Efficient face detection is critical to provide natural human-robot interactions. However, computer vision tends to involve a large computational load due to the amount of data (i.e. pixels) that needs to be processed in a short amount of…

音频与语音处理 · 电气工程与系统科学 2024-03-19 William Aris , François Grondin

Conventional multi-view re-ranking methods usually perform asymmetrical matching between the region of interest (ROI) in the query image and the whole target image for similarity computation. Due to the inconsistency in the visual…

计算机视觉与模式识别 · 计算机科学 2018-08-15 Jun Li , Chang Xu , Wankou Yang , Changyin Sun , Dacheng Tao , Hong Zhang

With advances in image recognition technology based on deep learning, automatic video analysis by Artificial Intelligence is becoming more widespread. As the amount of video used for image recognition increases, efficient compression…

计算机视觉与模式识别 · 计算机科学 2023-04-04 Takahiro Shindo , Taiju Watanabe , Kein Yamada , Hiroshi Watanabe

Recently, pseudo analog transmission has gained increasing attentions due to its ability to alleviate the cliff effect in video multicast scenarios. The existing pseudo analog systems are sorely optimized under the minimum mean squared…

多媒体 · 计算机科学 2020-07-22 Xiao-Wei Tang , Xin-Lin Huang , Fei Hu , Qingjiang Shi

Attention-based vision models, such as Vision Transformer (ViT) and its variants, have shown promising performance in various computer vision tasks. However, these emerging architectures suffer from large model sizes and high computational…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Jinqi Xiao , Miao Yin , Yu Gong , Xiao Zang , Jian Ren , Bo Yuan

The pests captured with imaging devices may be relatively small in size compared to the entire images, and complex backgrounds have colors and textures similar to those of the pests, which hinders accurate feature extraction and makes pest…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Ga-Eun Kim , Chang-Hwan Son

While most existing neural image compression (NIC) and neural video compression (NVC) methodologies have achieved remarkable success, their optimization is primarily focused on human visual perception. However, with the rapid development of…

计算机视觉与模式识别 · 计算机科学 2025-01-09 Lei Liu , Zhenghao Chen , Zhihao Hu , Dong Xu

In photography, low depth of field (DOF) is an important technique to emphasize the object of interest (OOI) within an image. Thus, low DOF images are widely used in the application area of macro, portrait or sports photography. When…

计算机视觉与模式识别 · 计算机科学 2013-02-19 Franz Graf , Hans-Peter Kriegel , Michael Weiler

Recently, visual encoding based on functional magnetic resonance imaging (fMRI) have realized many achievements with the rapid development of deep network computation. Visual encoding model is aimed at predicting brain activity in response…

神经元与认知 · 定量生物学 2019-07-30 Kai Qiao , Chi Zhang , Jian Chen , Linyuan Wang , Li Tong , Bin Yan

Deep neural object detection or segmentation networks are commonly trained with pristine, uncompressed data. However, in practical applications the input images are usually deteriorated by compression that is applied to efficiently transmit…

图像与视频处理 · 电气工程与系统科学 2022-05-16 Kristian Fischer , Christian Blum , Christian Herglotz , André Kaup

Conventional cameras capture image irradiance on a sensor and convert it to RGB images using an image signal processor (ISP). The images can then be used for photography or visual computing tasks in a variety of applications, such as public…

计算机视觉与模式识别 · 计算机科学 2024-01-26 Zhihao Li , Ming Lu , Xu Zhang , Xin Feng , M. Salman Asif , Zhan Ma

Video cameras are pervasively deployed in city scale for public good or community safety (i.e. traffic monitoring or suspected person tracking). However, analyzing large scale video feeds in real time is data intensive and poses severe…

分布式、并行与集群计算 · 计算机科学 2021-05-17 Hongpeng Guo , Shuochao Yao , Zhe Yang , Qian Zhou , Klara Nahrstedt

We propose a camera-based assistive text reading framework to help blind persons read text labels and product packaging from hand-held objects in their daily life. To isolate the object from untidy backgrounds or other surrounding objects…

人机交互 · 计算机科学 2019-01-18 Rajkumar N , Anand M. G , Barathiraja N

Spectral photon-counting X-ray CT (sCT) opens up new possibilities for the quantitative measurement of materials in an object, compared to conventional energy-integrating CT or dual energy CT. However, achieving reliable and accurate…

图像与视频处理 · 电气工程与系统科学 2020-07-15 Bingqing Xie , Pei Niu , Ting Su , Valérie Kaftandjian , Loic Boussel , Philippe Douek Feng Yang , Philippe Duvauchelle , Yuemin Zhu

This paper presents a novel scheme to efficiently compress Light Detection and Ranging~(LiDAR) point clouds, enabling high-precision 3D scene archives, and such archives pave the way for a detailed understanding of the corresponding 3D…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Akihiro Kuwabara , Sorachi Kato , Toshiaki Koike-Akino , Takuya Fujihashi

Retrieving images from large and varied repositories using visual contents has been one of major research items, but a challenging task in the image management community. In this paper we present an efficient approach for region-based image…

计算机视觉与模式识别 · 计算机科学 2010-06-24 S. Sadek , A. Al-Hamadi , B. Michaelis , U. Sayed