中文
相关论文

相关论文: Sketch It Out: Exploring Label-Free Structural Cue…

200 篇论文

Graphic layout generation is a growing research area focusing on generating aesthetically pleasing layouts ranging from poster designs to documents. While recent research has explored ways to incorporate user constraints to guide the layout…

The key challenge in designing a sketch representation lies with handling the abstract and iconic nature of sketches. Existing work predominantly utilizes either, (i) a pixelative format that treats sketches as natural images employing…

计算机视觉与模式识别 · 计算机科学 2021-08-27 Yonggang Qi , Guoyao Su , Pinaki Nath Chowdhury , Mingkang Li , Yi-Zhe Song

Sketching enables many exciting applications, notably, image retrieval. The fear-to-sketch problem (i.e., "I can't sketch") has however proven to be fatal for its widespread adoption. This paper tackles this "fear" head on, and for the…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Ayan Kumar Bhunia , Subhadeep Koley , Abdullah Faiz Ur Rahman Khilji , Aneeshan Sain , Pinaki Nath Chowdhury , Tao Xiang , Yi-Zhe Song

Semantic scene completion (SSC) aims to predict the semantic occupancy of each voxel in the entire 3D scene from limited observations, which is an emerging and critical task for autonomous driving. Recently, many studies have turned to…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Jianbiao Mei , Yu Yang , Mengmeng Wang , Junyu Zhu , Jongwon Ra , Yukai Ma , Laijian Li , Yong Liu

To capture individual gait patterns, excluding identity-irrelevant cues in walking videos, such as clothing texture and color, remains a persistent challenge for vision-based gait recognition. Traditional silhouette- and pose-based methods,…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Dongyang Jin , Chao Fan , Jingzhe Ma , Jingkai Zhou , Weihua Chen , Shiqi Yu

In this paper, we propose a novel deep framework for part-level semantic parsing of freehand sketches, which makes three main contributions that are experimentally shown to have substantial practical merit. First, we propose a homogeneous…

计算机视觉与模式识别 · 计算机科学 2020-12-04 Ying Zheng , Hongxun Yao , Xiaoshuai Sun

A fundamental challenge faced by existing Fine-Grained Sketch-Based Image Retrieval (FG-SBIR) models is the data scarcity -- model performances are largely bottlenecked by the lack of sketch-photo pairs. Whilst the number of photos can be…

计算机视觉与模式识别 · 计算机科学 2021-03-26 Ayan Kumar Bhunia , Pinaki Nath Chowdhury , Aneeshan Sain , Yongxin Yang , Tao Xiang , Yi-Zhe Song

Visual knowledge bases such as Visual Genome power numerous applications in computer vision, including visual question answering and captioning, but suffer from sparse, incomplete relationships. All scene graph models to date are limited to…

计算机视觉与模式识别 · 计算机科学 2019-12-03 Vincent S. Chen , Paroma Varma , Ranjay Krishna , Michael Bernstein , Christopher Re , Li Fei-Fei

Serial sectioning Optical Coherence Tomography (sOCT) is a high-throughput, label free microscopic imaging technique that is becoming increasingly popular to study post-mortem neurovasculature. Quantitative analysis of the vasculature…

图像与视频处理 · 电气工程与系统科学 2024-05-24 Etienne Chollet , Yael Balbastre , Caroline Magnain , Bruce Fischl , Hui Wang

Frailty is a condition in aging medicine characterized by diminished physiological reserve and increased vulnerability to stressors. However, frailty assessment remains subjective, heterogeneous, and difficult to scale in clinical practice.…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Laura McDaniel , Basudha Pal , Crystal Szczesny , Yuxiang Guo , Zhaoyang Wang , Ryan Roemmich , Peter Abadir , Rama Chellappa

Human visual system has the strong ability to quick assess the perceptual similarity between two facial sketches. However, existing two widely-used facial sketch metrics, e.g., FSIM and SSIM fail to address this perceptual similarity in…

计算机视觉与模式识别 · 计算机科学 2019-09-05 Deng-Ping Fan , ShengChuan Zhang , Yu-Huan Wu , Yun Liu , Ming-Ming Cheng , Bo Ren , Paul L. Rosin , Rongrong Ji

Gait analysis provides an objective characterization of locomotor function and is widely used to support diagnosis and rehabilitation monitoring across neurological and orthopedic disorders. Deep learning has been increasingly applied to…

人工智能 · 计算机科学 2026-04-03 Elisa Motta , Marta Lorenzini , Clara Mouawad , Alberto Ranavolo , Mariano Serrao , Arash Ajoudani

Noise and artifacts during computed tomography (CT) scans are a fundamental challenge affecting disease diagnosis. However, current methods either involve excessively long reconstruction times or rely on data-driven models for optimization,…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Guoquan Wei , Liu Shi , Shaoyu Wang , Mohan Li , Cunfeng Wei , Qiegen Liu

Gait recognition holds the promise to robustly identify subjects based on walking patterns instead of appearance information. In recent years, this field has been dominated by learning methods based on two principal input representations:…

计算机视觉与模式识别 · 计算机科学 2024-10-15 Yuxiang Guo , Anshul Shah , Jiang Liu , Ayush Gupta , Rama Chellappa , Cheng Peng

Semi-supervised learning has emerged as a powerful paradigm for leveraging large amounts of unlabeled data to improve the performance of machine learning models when labeled data are scarce. Among existing approaches, methods derived from…

机器学习 · 计算机科学 2026-04-29 Ali Aghababaei-Harandi , Aude Sportisse , Massih-Reza Amini

Recent years have witnessed remarkable progress in generative AI, with natural language emerging as the most common conditioning input. As underlying models grow more powerful, researchers are exploring increasingly diverse conditioning…

计算机视觉与模式识别 · 计算机科学 2026-02-17 Ahmed Bourouis , Mikhail Bessmeltsev , Yulia Gryaditskaya

Scene graph generation (SGG) is a sophisticated task that suffers from both complex visual features and dataset long-tail problem. Recently, various unbiased strategies have been proposed by designing novel loss functions and data balancing…

计算机视觉与模式识别 · 计算机科学 2023-03-03 Xiaoguang Chang , Teng Wang , Shaowei Cai , Changyin Sun

Gait recognition, which refers to the recognition or identification of a person based on their body shape and walking styles, derived from video data captured from a distance, is widely used in crime prevention, forensic identification, and…

计算机视觉与模式识别 · 计算机科学 2022-07-14 Hung-Min Hsu , Yizhou Wang , Cheng-Yen Yang , Jenq-Neng Hwang , Hoang Le Uyen Thuc , Kwang-Ju Kim

Large-scale vision-language pre-training has achieved significant performance in multi-modal understanding and generation tasks. However, existing methods often perform poorly on image-text matching tasks that require structured…

计算与语言 · 计算机科学 2023-12-14 Yufeng Huang , Jiji Tang , Zhuo Chen , Rongsheng Zhang , Xinfeng Zhang , Weijie Chen , Zeng Zhao , Zhou Zhao , Tangjie Lv , Zhipeng Hu , Wen Zhang

Multimodal UI design and development tools that interpret sketches or natural language descriptions of UIs inherently have notations: the inputs they can understand. In AI-based systems, notations are implicitly defined by the data used to…

人机交互 · 计算机科学 2025-08-14 Sam H. Ross , Yunseo Lee , Coco K. Lee , Jayne Everson , R. Benjamin Shapiro