中文
相关论文

相关论文: BHDD: A Burmese Handwritten Digit Dataset

200 篇论文

We address the problem of registering synchronized color (RGB) and multi-spectral (MS) images featuring very different resolution by solving stereo matching correspondences. Purposely, we introduce a novel RGB-MS dataset framing 13…

计算机视觉与模式识别 · 计算机科学 2022-06-15 Fabio Tosi , Pierluigi Zama Ramirez , Matteo Poggi , Samuele Salti , Stefano Mattoccia , Luigi Di Stefano

We present the Multilingual Cloud Corpus, the first national-scale, parallel, multimodal linguistic dataset of Bangladesh's ethnic and indigenous languages. Despite being home to approximately 40 minority languages spanning four language…

计算与语言 · 计算机科学 2026-03-09 Mohammad Mamun Or Rashid

Handwritten Text Recognition (HTR) is a well-established research area. In contrast, Handwritten Text Generation (HTG) is an emerging field with significant potential. This task is challenging due to the variation in individual handwriting…

计算机视觉与模式识别 · 计算机科学 2025-12-29 Md. Rakibul Islam , Md. Kamrozzaman Bhuiyan , Safwan Muntasir , Arifur Rahman Jawad , Most. Sharmin Sultana Samu

Large DNNs with mixed-precision quantization can achieve ultra-high compression while retaining high classification performance. However, because of the challenges in finding an accurate metric that can guide the optimization process, these…

计算机视觉与模式识别 · 计算机科学 2021-12-30 Souvik Kundu , Shikai Wang , Qirui Sun , Peter A. Beerel , Massoud Pedram

Object grasping is critical for many applications, which is also a challenging computer vision problem. However, for the clustered scene, current researches suffer from the problems of insufficient training data and the lacking of…

计算机视觉与模式识别 · 计算机科学 2020-01-03 Hao-Shu Fang , Chenxi Wang , Minghao Gou , Cewu Lu

We present a new tool for colour-magnitude diagram (CMD) studies, $Powerful~CMD$. This tool is built on the basis of the advanced stellar population synthesis (ASPS) model, in which single stars, binary stars, rotating stars, and star…

星系天体物理 · 物理学 2017-07-19 Zhong-Mu Li , Cai-Yan Mao , Qi-Ping Luo , Zhou Fan , Wen-Chang Zhao , Li Chen , Ru-Xi Li , Jian-Po Guo

Vietnamese, a low-resource language, is typically categorized into three primary dialect groups that belong to Northern, Central, and Southern Vietnam. However, each province within these regions exhibits its own distinct pronunciation…

计算与语言 · 计算机科学 2024-10-07 Nguyen Van Dinh , Thanh Chi Dang , Luan Thanh Nguyen , Kiet Van Nguyen

We study the longitudinal stability of compact binary face templates and quantify ageing drift directly in bits per decade. Float embeddings from a modern face CNN are compressed with PCA-ITQ into 64- and 128-bit codes. For each identity in…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Abdelilah Ganmati , Karim Afdel , Lahcen Koutti

Machine Learning (ML) has garnered considerable attention from researchers and practitioners as a new and adaptable tool for disease diagnosis. With the advancement of ML and the proliferation of papers and research in this field, a…

机器学习 · 计算机科学 2022-01-11 Md Manjurul Ahsan , Zahed Siddique

We present a ghost handwritten digit recognition method for the unknown handwritten digits based on ghost imaging (GI) with deep neural network, where a few detection signals from the bucket detector, generated by the Cosine Transform…

图像与视频处理 · 电气工程与系统科学 2021-04-21 Xing He , Shengmei Zhao , Le Wang

The plethora of digitalised historical document datasets released in recent years has rekindled interest in advancing the field of handwriting pattern recognition. In the same vein, a recently published data set, known as ARDIS, presents…

计算机视觉与模式识别 · 计算机科学 2022-07-14 Mengqiao Zhao , Andre G. Hochuli , Abbas Cheddad

In the domain of Bangla Sign Language (BdSL) interpretation, prior approaches often imposed a burden on users, requiring them to spell words without hidden characters, which were subsequently corrected using Bangla grammar rules due to the…

计算机视觉与模式识别 · 计算机科学 2023-09-26 Naimul Haque , Meraj Serker , Tariq Bin Bashar

Computer-aided detection systems based on deep learning have shown good performance in breast cancer detection. However, high-density breasts show poorer detection performance since dense tissues can mask or even simulate masses. Therefore,…

Origami is becoming more and more relevant to research. However, there is no public dataset yet available and there hasn't been any research on this topic in machine learning. We constructed an origami dataset using images from the…

计算机视觉与模式识别 · 计算机科学 2021-01-15 Daniel Ma , Gerald Friedland , Mario Michael Krell

Consumer Health Queries (CHQs) in Bengali (Bangla), a low-resource language, often contain extraneous details, complicating efficient medical responses. This study investigates the zero-shot performance of nine advanced large language…

计算与语言 · 计算机科学 2025-09-16 Ajwad Abrar , Farzana Tabassum , Sabbir Ahmed

A new methodology to measure coded image/video quality using the just-noticeable-difference (JND) idea was proposed. Several small JND-based image/video quality datasets were released by the Media Communications Lab at the University of…

The Quick, Draw! Dataset is a Google dataset with a collection of 50 million drawings, divided in 345 categories, collected from the users of the game Quick, Draw!. In contrast with most of the existing image datasets, in the Quick, Draw!…

计算机视觉与模式识别 · 计算机科学 2019-10-24 Raul Fernandez-Fernandez , Juan G. Victores , David Estevez , Carlos Balaguer

The ability to grasp objects, signal with gestures, and share emotion through touch all stem from the unique capabilities of human hands. Yet creating high-quality personalized hand avatars from images remains challenging due to complex…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Zicong Fan , Edoardo Remelli , David Dimond , Fadime Sener , Liuhao Ge , Bugra Tekin , Cem Keskin , Shreyas Hampali

The Bengali language is the 5th most spoken native and 7th most spoken language in the world, and Bengali handwritten character recognition has attracted researchers for decades. However, other languages such as English, Arabic, Turkey, and…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Farhanul Haque , Md. Al-Hasan , Sumaiya Tabssum Mou , Abu Saleh Musa Miah , Jungpil Shin , Md Abdur Rahim

In this paper, we present TituLLMs, the first large pretrained Bangla LLMs, available in 1b and 3b parameter sizes. Due to computational constraints during both training and inference, we focused on smaller models. To train TituLLMs, we…