English
Related papers

Related papers: Cube Padding for Weakly-Supervised Saliency Predic…

200 papers

Video stabilization technique is essential for most hand-held captured videos due to high-frequency shakes. Several 2D-, 2.5D- and 3D-based stabilization techniques are well studied, but to our knowledge, no solutions based on deep neural…

Graphics · Computer Science 2018-02-23 Miao Wang , Guo-Ye Yang , Jin-Kun Lin , Ariel Shamir , Song-Hai Zhang , Shao-Ping Lu , Shi-Min Hu

Augmented reality (AR) overlays digital content onto the reality. In AR system, correct and precise estimations of user's visual fixations and head movements can enhance the quality of experience by allocating more computation resources on…

Image and Video Processing · Electrical Eng. & Systems 2020-07-21 Yucheng Zhu , Xiongkuo Min , DanDan Zhu , Ke Gu , Jiantao Zhou , Guangtao Zhai , Xiaokang Yang , Wenjun Zhang

Thanks to the ability of providing an immersive and interactive experience, the uptake of 360 degree image content has been rapidly growing in consumer and industrial applications. Compared to planar 2D images, saliency prediction for 360…

Computer Vision and Pattern Recognition · Computer Science 2023-03-16 Pan Gao , Xinlang Chen , Rong Quan , Wei Xiang

Deep learning based salient object detection has recently achieved great success with its performance greatly outperforms any other unsupervised methods. However, annotating per-pixel saliency masks is a tedious and inefficient procedure.…

Computer Vision and Pattern Recognition · Computer Science 2018-03-20 Guanbin Li , Yuan Xie , Liang Lin

Viewport prediction is the crucial task for adaptive 360-degree video streaming, as the bitrate control algorithms usually require the knowledge of the user's viewing portions of the frames. Various methods are studied and adopted for…

Multimedia · Computer Science 2024-03-06 Lei Zhang , Tao Long , Weizhen Xu , Laizhong Cui , Jiangchuan Liu

Predicting attention is a popular topic at the intersection of human and computer vision. However, even though most of the available video saliency data sets and models claim to target human observers' fixations, they fail to differentiate…

Computer Vision and Pattern Recognition · Computer Science 2019-04-15 Mikhail Startsev , Michael Dorr

Recently, end-to-end trainable deep neural networks have significantly improved stereo depth estimation for perspective images. However, 360{\deg} images captured under equirectangular projection cannot benefit from directly adopting…

Computer Vision and Pattern Recognition · Computer Science 2020-03-27 Ning-Hsu Wang , Bolivar Solarte , Yi-Hsuan Tsai , Wei-Chen Chiu , Min Sun

Effective compression of 360$^\circ$ images, also referred to as omnidirectional images (ODIs), is of high interest for various virtual reality (VR) and related applications. 2D image compression methods ignore the equator-biased nature of…

Image and Video Processing · Electrical Eng. & Systems 2024-02-15 Oguzhan Gungordu , A. Murat Tekalp

Weakly supervised semantic segmentation (WSSS) aims to produce pixel-wise class predictions with only image-level labels for training. To this end, previous methods adopt the common pipeline: they generate pseudo masks from class activation…

Computer Vision and Pattern Recognition · Computer Science 2022-08-09 Sungpil Kho , Pilhyeon Lee , Wonyoung Lee , Minsong Ki , Hyeran Byun

The prediction of salient areas in images has been traditionally addressed with hand-crafted features based on neuroscience principles. This paper, however, addresses the problem with a completely data-driven approach by training a…

Computer Vision and Pattern Recognition · Computer Science 2016-03-03 Junting Pan , Kevin McGuinness , Elisa Sayrol , Noel O'Connor , Xavier Giro-i-Nieto

Inpainting has been continuously studied in the field of computer vision. As artificial intelligence technology developed, deep learning technology was introduced in inpainting research, helping to improve performance. Currently, the input…

Image and Video Processing · Electrical Eng. & Systems 2021-01-27 Seo Woo Han , Doug Young Suh

The task of human pose estimation (HPE) deals with the ill-posed problem of estimating the 3D position of human joints directly from images and videos. In recent literature, most of the works tackle the problem mostly by using convolutional…

Computer Vision and Pattern Recognition · Computer Science 2023-02-14 Nicola Garau , Nicola Conci

360-degree streaming videos can provide a rich immersive experiences to the users. However, it requires an extremely high bandwidth network. One of the common solutions for saving bandwidth consumption is to stream only a portion of video…

Multimedia · Computer Science 2022-06-07 Chuanzhe Jing , Tho Nguyen Duc , Phan Xuan Tan , Eiji Kamioka

Recently, video streams have occupied a large proportion of Internet traffic, most of which contain human faces. Hence, it is necessary to predict saliency on multiple-face videos, which can provide attention cues for many content based…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Yufan Liu , Minglang Qiao , Mai Xu , Bing Li , Weiming Hu , Ali Borji

In this paper we propose a highly scalable convolutional neural network, end-to-end trainable, for real-time 3D human pose regression from still RGB images. We call this approach the Scalable Sequential Pyramid Networks (SSP-Net) as it is…

Computer Vision and Pattern Recognition · Computer Science 2020-09-07 Diogo Luvizon , Hedi Tabia , David Picard

We aim to tackle sparse-view reconstruction of a 360 3D scene using priors from latent diffusion models (LDM). The sparse-view setting is ill-posed and underconstrained, especially for scenes where the camera rotates 360 degrees around a…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Soumava Paul , Christopher Wewer , Bernt Schiele , Jan Eric Lenssen

Image-based salient object detection (SOD) has been extensively studied in the past decades. However, video-based SOD is much less explored since there lack large-scale video datasets within which salient objects are unambiguously defined…

Computer Vision and Pattern Recognition · Computer Science 2017-05-10 Jia Li , Changqun Xia , Xiaowu Chen

Optical flow computation is essential in the early stages of the video processing pipeline. This paper focuses on a less explored problem in this area, the 360$^\circ$ optical flow estimation using deep neural networks to support…

Computer Vision and Pattern Recognition · Computer Science 2022-08-02 Yiheng Li , Connelly Barnes , Kun Huang , Fang-Lue Zhang

To watch 360{\deg} videos on normal 2D displays, we need to project the selected part of the 360{\deg} image onto the 2D display plane. In this paper, we propose a fully-automated framework for generating content-aware 2D normal-view…

Graphics · Computer Science 2017-09-12 Yeong Won Kim , Dae-Yong Jo , Chang-Ryeol Lee , Hyeok-Jae Choi , Yong Hoon Kwon , Kuk-Jin Yoon

The Dynamic Saliency Prediction (DSP) task simulates the human selective attention mechanism to perceive the dynamic scene, which is significant and imperative in many vision tasks. Most of existing methods only consider visual cues, while…

Computer Vision and Pattern Recognition · Computer Science 2022-05-03 Hailong Ning , Bin Zhao , Zhanxuan Hu , Lang He , Ercheng Pei