中文
相关论文

相关论文: AutoOEP -- A Multi-modal Framework for Online Exam…

200 篇论文

Previous research in $2D$ object detection focuses on various tasks, including detecting objects in generic and camouflaged images. These works are regarded as passive works for object detection as they take the input image as is. However,…

计算机视觉与模式识别 · 计算机科学 2023-10-31 Vishal Asnani , Abhinav Kumar , Suya You , Xiaoming Liu

Road safety is a critical global concern, with manual enforcement of helmet laws and vehicle safety standards (e.g., rear-view mirror presence) being resource-intensive and inconsistent. This paper presents an AI-powered system to automate…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Nishant Vasantkumar Hegde , Aditi Agarwal , Minal Moharir

This paper proposes an online visual multi-object tracking algorithm using a top-down Bayesian formulation that seamlessly integrates state estimation, track management, clutter rejection, occlusion and mis-detection handling into a single…

计算机视觉与模式识别 · 计算机科学 2017-08-07 Du Yong Kim , Ba-Ngu Vo , Ba-Tuong Vo

Open World Object Detection (OWOD) is a challenging computer vision task that extends standard object detection by (1) detecting and classifying unknown objects without supervision, and (2) incrementally learning new object classes without…

计算机视觉与模式识别 · 计算机科学 2025-07-18 Riku Inoue , Masamitsu Tsuchiya , Yuji Yasui

This paper proposes a novel approach to object detection on drone imagery, namely Multi-Proxy Detection Network with Unified Foreground Packing (UFPMP-Det). To deal with the numerous instances of very small scales, different from the common…

计算机视觉与模式识别 · 计算机科学 2022-01-04 Yecheng Huang , Jiaxin Chen , Di Huang

The growing need for video surveillance in public spaces has created a demand for systems that can track individuals across multiple cameras feeds in real-time. While existing tracking systems have achieved impressive performance using deep…

计算机视觉与模式识别 · 计算机科学 2023-09-26 Vipin Gautam , Shitala Prasad , Sharad Sinha

This paper introduces Test-time Correction (TTC), an online 3D detection system designed to rectify test-time errors using various auxiliary feedback, aiming to enhance the safety of deployed autonomous driving systems. Unlike conventional…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Hanxue Zhang , Zetong Yang , Yanan Sun , Li Chen , Fei Xia , Fatma Güney , Hongyang Li

Nowadays, the increasingly growing number of mobile and computing devices has led to a demand for safer user authentication systems. Face anti-spoofing is a measure towards this direction for bio-metric user authentication, and in…

计算机视觉与模式识别 · 计算机科学 2020-04-14 Suman Saha , Wenhao Xu , Menelaos Kanakis , Stamatios Georgoulis , Yuhua Chen , Danda Pani Paudel , Luc Van Gool

Reliable zero-shot detection of out-of-distribution (OOD) inputs is critical for deploying vision-language models in open-world settings. However, the lack of labeled negatives in zero-shot OOD detection necessitates proxy signals that…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Hao Tang , Yu Liu , Shuanglin Yan , Fei Shen , Shengfeng He , Jing Qin

Automatic deception detection is an important task that has gained momentum in computational linguistics due to its potential applications. In this paper, we propose a simple yet tough to beat multi-modal neural model for deception…

计算与语言 · 计算机科学 2018-03-21 Gangeshwar Krishnamurthy , Navonil Majumder , Soujanya Poria , Erik Cambria

To reduce the manpower consumption on box-level annotations, many weakly supervised object detection methods which only require image-level annotations, have been proposed recently. The training process in these methods is formulated into…

计算机视觉与模式识别 · 计算机科学 2021-01-21 Ruibing Jin , Guosheng Lin , Changyun Wen

Substantial advances in multi-modal Artificial Intelligence (AI) facilitate the combination of diverse medical modalities to achieve holistic health assessments. We present COMPRER , a novel multi-modal, multi-objective pretraining…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Guy Lutsker , Hagai Rossman , Nastya Godiva , Eran Segal

With the advancement of face manipulation technology, forgery images in multi-face scenarios are gradually becoming a more complex and realistic challenge. Despite this, detection and localization methods for such multi-face manipulations…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Changtao Miao , Qi Chu , Tao Gong , Zhentao Tan , Zhenchao Jin , Wanyi Zhuang , Man Luo , Honggang Hu , Nenghai Yu

Open-vocabulary object detectors (OVODs) unify vision and language to detect arbitrary object categories based on text prompts, enabling strong zero-shot generalization to novel concepts. As these models gain traction in high-stakes…

计算机视觉与模式识别 · 计算机科学 2026-01-14 Ankita Raj , Chetan Arora

Component-level audio Spoofing (Comp-Spoof) targets a new form of audio manipulation where only specific components of a signal, such as speech or environmental sound, are forged or substituted while other components remain genuine.…

声音 · 计算机科学 2026-02-02 Xueping Zhang , Yechen Wang , Linxi Li , Liwei Jin , Ming Li

Accurate trajectory prediction is crucial for safe and efficient autonomous driving, but handling partial observations presents significant challenges. To address this, we propose a novel trajectory prediction framework called Partial…

机器人学 · 计算机科学 2024-04-08 Sheng Wang , Yingbing Chen , Jie Cheng , Xiaodong Mei , Ren Xin , Yongkang Song , Ming Liu

In this work, we propose a novel deep online correction (DOC) framework for monocular visual odometry. The whole pipeline has two stages: First, depth maps and initial poses are obtained from convolutional neural networks (CNNs) trained in…

计算机视觉与模式识别 · 计算机科学 2021-12-17 Jiaxin Zhang , Wei Sui , Xinggang Wang , Wenming Meng , Hongmei Zhu , Qian Zhang

Human body orientation estimation (HBOE) is widely applied into various applications, including robotics, surveillance, pedestrian analysis and autonomous driving. Although many approaches have been addressing the HBOE problem from specific…

计算机视觉与模式识别 · 计算机科学 2023-03-17 Huayi Zhou , Fei Jiang , Jiaxin Si , Hongtao Lu

The field of object detection has made significant advances riding on the wave of region-based ConvNets, but their training procedure still includes many heuristics and hyperparameters that are costly to tune. We present a simple yet…

计算机视觉与模式识别 · 计算机科学 2016-04-13 Abhinav Shrivastava , Abhinav Gupta , Ross Girshick

Cross-topic automated essay scoring (AES) aims to develop a transferable model capable of effectively evaluating essays on a target topic. A significant challenge in this domain arises from the inherent discrepancies between topics. While…

计算与语言 · 计算机科学 2025-08-11 Chunyun Zhang , Hongyan Zhao , Chaoran Cui , Qilong Song , Zhiqing Lu , Shuai Gong , Kailin Liu