中文
相关论文

相关论文: ScreenSeg: On-Device Screenshot Layout Analysis

200 篇论文

The importance of hierarchical image organization has been witnessed by a wide spectrum of applications in computer vision and graphics. Different from image segmentation with the spatial whole-part consideration, this work designs a modern…

计算机视觉与模式识别 · 计算机科学 2021-04-06 Fu Yuanbin , Guoxiaojie , Hu Qiming , Lin Di , Ma Jiayi , Ling Haibin

Recently, automated medical image segmentation methods based on deep learning have achieved great success. However, they heavily rely on large annotated datasets, which are costly and time-consuming to acquire. Few-shot learning aims to…

人工智能 · 计算机科学 2024-08-20 Jiayu Huo , Ruiqiang Xiao , Haotian Zheng , Yang Liu , Sebastien Ourselin , Rachel Sparks

We introduce LighthouseGS, a practical novel view synthesis framework based on 3D Gaussian Splatting that utilizes simple panorama-style captures from a single mobile device. While convenient, this rotation-dominant motion and narrow…

图形学 · 计算机科学 2026-02-12 Seungoh Han , Jaehoon Jang , Hyunsu Kim , Jaeheung Surh , Junhyung Kwak , Hyowon Ha , Kyungdon Joo

A significant limitation of current smartphone-based eye-tracking algorithms is their low accuracy when applied to video-type visual stimuli, as they are typically trained on static images. Also, the increasing demand for real-time…

计算机视觉与模式识别 · 计算机科学 2025-01-15 Nishan Gunawardena , Gough Yumu Lui , Jeewani Anupama Ginige , Bahman Javadi

The ability to semantically interpret hand-drawn line sketches, although very challenging, can pave way for novel applications in multimedia. We propose SketchParse, the first deep-network architecture for fully automatic parsing of…

计算机视觉与模式识别 · 计算机科学 2017-09-06 Ravi Kiran Sarvadevabhatla , Isht Dwivedi , Abhijat Biswas , Sahil Manocha , R. Venkatesh Babu

We tackle the problem of semantic image layout manipulation, which aims to manipulate an input image by editing its semantic label map. A core problem of this task is how to transfer visual details from the input images to the new semantic…

计算机视觉与模式识别 · 计算机科学 2022-04-19 Haitian Zheng , Zhe Lin , Jingwan Lu , Scott Cohen , Jianming Zhang , Ning Xu , Jiebo Luo

High-resolution remote sensing images (HRRSIs) contain substantial ground object information, such as texture, shape, and spatial location. Semantic segmentation, which is an important task for element extraction, has been widely used in…

计算机视觉与模式识别 · 计算机科学 2020-05-08 Haifeng Li , Kaijian Qiu , Li Chen , Xiaoming Mei , Liang Hong , Chao Tao

Foundation segmentation models achieve reasonable leaf instance extraction from top-view crop images without training (i.e., zero-shot). However, segmenting entire plant individuals with each consisting of multiple overlapping leaves…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Junhao Xing , Ryohei Miyakawa , Yang Yang , Xinpeng Liu , Risa Shinoda , Hiroaki Santo , Yosuke Toda , Fumio Okura

Recent research on super-resolution (SR) has witnessed major developments with the advancements of deep convolutional neural networks. There is a need for information extraction from scenic text images or even document images on device,…

计算机视觉与模式识别 · 计算机科学 2022-01-03 Dhruval Jain , Arun D Prabhu , Gopi Ramena , Manoj Goyal , Debi Prasanna Mohanty , Sukumar Moharana , Naresh Purre

Document layout analysis is a known problem to the documents research community and has been vastly explored yielding a multitude of solutions ranging from text mining, and recognition to graph-based representation, visual feature…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Subhajit Maity , Sanket Biswas , Siladittya Manna , Ayan Banerjee , Josep Lladós , Saumik Bhattacharya , Umapada Pal

The challenge of estimating similarity between sets has been a significant concern in data science, finding diverse applications across various domains. However, previous approaches, such as MinHash, have predominantly centered around…

数据结构与算法 · 计算机科学 2024-05-31 Fenghao Dong , Yang He , Yutong Liang , Zirui Liu , Yuhan Wu , Peiqing Chen , Tong Yang

The fragmentation problem has extended from Android to different platforms, such as iOS, mobile web, and even mini-programs within some applications (app). In such a situation, recording and replaying test scripts is a popular automated…

软件工程 · 计算机科学 2021-02-23 Shengcheng Yu , Chunrong Fang , Yexiao Yun , Yang Feng

We propose a framework for the automatic one-shot segmentation of synthetic images generated by a StyleGAN. Our framework is based on the observation that the multi-scale hidden features in the GAN generator hold useful semantic information…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Ankit Manerikar , Avinash C. Kak

With the rapid development of AI hardware accelerators, applying deep learning-based algorithms to solve various low-level vision tasks on mobile devices has gradually become possible. However, two main problems still need to be solved:…

计算机视觉与模式识别 · 计算机科学 2023-08-17 Weiran Gou , Ziyao Yi , Yan Xiang , Shaoqing Li , Zibin Liu , Dehui Kong , Ke Xu

The performance of a camera network monitoring a set of targets depends crucially on the configuration of the cameras. In this paper, we investigate the reconfiguration strategy for the parameterized camera network model, with which the…

计算机视觉与模式识别 · 计算机科学 2023-03-01 Xuechao Zhang , Xuda Ding , Yi Ren , Yu Zheng , Chongrong Fang , Jianping He

This study introduces a novel approach to enhance the spatial-temporal resolution of time-event pixels based on luminance changes captured by event cameras. These cameras present unique challenges due to their low resolution and the sparse,…

图像与视频处理 · 电气工程与系统科学 2024-08-14 Waseem Shariff , Joe Lemley , Peter Corcoran

The outbreak of COVID-19 exposed the inadequacy of our technical tools for home health surveillance, and recent studies have shown the potential of smartphones as a universal optical microscopic imaging platform for such applications.…

图像与视频处理 · 电气工程与系统科学 2023-12-20 Haoran Zhang , Weiyi Zhang , Zirui Zuo , Jianlong Yang

With the recent surge in the use of touchscreen devices, free-hand sketching has emerged as a promising modality for human-computer interaction. While previous research has focused on tasks such as recognition, retrieval, and generation of…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Guangming Zhu , Siyuan Wang , Qing Cheng , Kelong Wu , Hao Li , Liang Zhang

Existing text-to-image (T2I) diffusion models face several limitations, including large model sizes, slow runtime, and low-quality generation on mobile devices. This paper aims to address all of these challenges by developing an extremely…

The main contributions of our work are two-fold. First, we present a Self-Attention MobileNet, called SA-MobileNet Network that can model long-range dependencies between the image features instead of processing the local region as done by…

计算机视觉与模式识别 · 计算机科学 2021-11-02 Siddhant Garg , Debi Prasanna Mohanty , Siva Prasad Thota , Sukumar Moharana