中文
相关论文

相关论文: From Benedict Cumberbatch to Sherlock Holmes: Char…

200 篇论文

An effective approach to automated movie content analysis involves building a network (graph) of its characters. Existing work usually builds a static character graph to summarize the content using metadata, scripts or manual annotations.…

计算机视觉与模式识别 · 计算机科学 2020-07-30 Prakhar Kulshreshtha , Tanaya Guha

Face recognition from image or video is a popular topic in biometrics research. Many public places usually have surveillance cameras for video capture and these cameras have their significant value for security purpose. It is widely…

计算机视觉与模式识别 · 计算机科学 2013-02-27 Faizan Ahmad , Aaima Najam , Zeeshan Ahmed

Deep neural networks have achieved great successes on the image captioning task. However, most of the existing models depend heavily on paired image-sentence datasets, which are very expensive to acquire. In this paper, we make the first…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Yang Feng , Lin Ma , Wei Liu , Jiebo Luo

In some specific scenarios, face sketch was used to identify a person. However, drawing a complete face sketch often needs skills and takes time, which hinder its widespread applicability in the practice. In this study, we proposed a new…

计算机视觉与模式识别 · 计算机科学 2023-02-14 Dawei Dai , Yutang Li , Liang Wang , Shiyu Fu , Shuyin Xia , Guoyin Wang

Current face recognition systems typically operate via classification into known identities obtained from supervised identity annotations. There are two problems with this paradigm: (1) current systems are unable to benefit from often…

计算机视觉与模式识别 · 计算机科学 2018-11-20 Daniel C. Castro , Sebastian Nowozin

The technological advancement and sophistication in cameras and gadgets prompt researchers to have focus on image analysis and text understanding. The deep learning techniques demonstrated well to assess the potential for classifying text…

计算机视觉与模式识别 · 计算机科学 2017-04-25 Saad Bin Ahmed , Saeeda Naz , Muhammad Imran Razzak , Rubiyah Yousaf

The human face constantly conveys information, both consciously and subconsciously. However, as basic as it is for humans to visually interpret this information, it is quite a big challenge for machines. Conventional semantic facial feature…

机器学习 · 计算机科学 2016-10-21 Amogh Gudi

There is an abundant literature on face detection due to its important role in many vision applications. Since Viola and Jones proposed the first real-time AdaBoost based face detector, Haar-like features have been adopted as the method of…

计算机视觉与模式识别 · 计算机科学 2010-09-30 Sakrapee Paisitkriangkrai , Chunhua Shen , Jian Zhang

In this paper we introduce the problem of Visual Semantic Role Labeling: given an image we want to detect people doing actions and localize the objects of interaction. Classical approaches to action recognition either study the task of…

计算机视觉与模式识别 · 计算机科学 2015-05-19 Saurabh Gupta , Jitendra Malik

In this paper we present a new data-driven method for robust skin detection from a single human portrait image. Unlike previous methods, we incorporate human body as a weak semantic guidance into this task, considering acquiring large-scale…

计算机视觉与模式识别 · 计算机科学 2019-08-07 Yi He , Jiayuan Shi , Chuan Wang , Haibin Huang , Jiaming Liu , Guanbin Li , Risheng Liu , Jue Wang

We investigate the importance of parts for the tasks of action and attribute classification. We develop a part-based approach by leveraging convolutional network features inspired by recent advances in computer vision. Our part detectors…

计算机视觉与模式识别 · 计算机科学 2015-05-07 Georgia Gkioxari , Ross Girshick , Jitendra Malik

This paper presents a novel approach in a rarely studied area of computer vision: Human interaction recognition in still images. We explore whether the facial regions and their spatial configurations contribute to the recognition of…

计算机视觉与模式识别 · 计算机科学 2015-09-18 Gokhan Tanisik , Cemil Zalluhoglu , Nazli Ikizler-Cinbis

When reading a literary piece, readers often make inferences about various characters' roles, personalities, relationships, intents, actions, etc. While humans can readily draw upon their past experiences to build such a character-centric…

计算与语言 · 计算机科学 2021-09-14 Faeze Brahman , Meng Huang , Oyvind Tafjord , Chao Zhao , Mrinmaya Sachan , Snigdha Chaturvedi

The need for a large amount of labeled data in the supervised setting has led recent studies to utilize self-supervised learning to pre-train deep neural networks using unlabeled data. Many self-supervised training strategies have been…

计算机视觉与模式识别 · 计算机科学 2022-04-06 Mojtaba Bahrami , Mahsa Ghorbani , Nassir Navab

This paper addresses the problem of automatically detecting human skin in images without reliance on color information. A primary motivation of the work has been to achieve results that are consistent across the full range of skin tones,…

计算机视觉与模式识别 · 计算机科学 2022-04-26 Han Xu , Abhijit Sarkar , A. Lynn Abbott

This paper targets the problem of image set-based face verification and identification. Unlike traditional single media (an image or video) setting, we encounter a set of heterogeneous contents containing orderless images and videos. The…

计算机视觉与模式识别 · 计算机科学 2019-08-06 Xiaofeng Liu , B. V. K Vijaya Kumar , Chao Yang , Qingming Tang , Jane You

To help the visually impaired enjoy movies, automatic movie narrating systems are expected to narrate accurate, coherent, and role-aware plots when there are no speaking lines of actors. Existing works benchmark this challenge as a normal…

计算机视觉与模式识别 · 计算机科学 2023-06-28 Zihao Yue , Qi Zhang , Anwen Hu , Liang Zhang , Ziheng Wang , Qin Jin

While natural language understanding of long-form documents is still an open challenge, such documents often contain structural information that can inform the design of models for encoding them. Movie scripts are an example of such richly…

计算与语言 · 计算机科学 2020-05-01 Gayatri Bhat , Avneesh Saluja , Melody Dye , Jan Florjanczyk

High-fidelity human 3D models can now be learned directly from videos, typically by combining a template-based surface model with neural representations. However, obtaining a template surface requires expensive multi-view capture systems,…

计算机视觉与模式识别 · 计算机科学 2023-09-04 Shih-Yang Su , Timur Bagautdinov , Helge Rhodin

Building robust recognizers for Arabic has always been challenging. We demonstrate the effectiveness of an end-to-end trainable CNN-RNN hybrid architecture in recognizing Arabic text in videos and natural scenes. We outperform previous…

计算机视觉与模式识别 · 计算机科学 2017-11-08 Mohit Jain , Minesh Mathew , C. V. Jawahar