中文
相关论文

相关论文: The HASYv2 dataset

200 篇论文

We present Fashion-MNIST, a new dataset comprising of 28x28 grayscale images of 70,000 fashion products from 10 categories, with 7,000 images per category. The training set has 60,000 images and the test set has 10,000 images. Fashion-MNIST…

机器学习 · 计算机科学 2017-09-19 Han Xiao , Kashif Rasul , Roland Vollgraf

Reducing the atmospheric haze and enhancing image clarity is crucial for computer vision applications. The lack of real-life hazy ground truth images necessitates synthetic datasets, which often lack diverse haze types, impeding effective…

计算机视觉与模式识别 · 计算机科学 2024-09-27 Md Tanvir Islam , Nasir Rahim , Saeed Anwar , Muhammad Saqib , Sambit Bakshi , Khan Muhammad

Context: When software is released publicly, it is common to include with it either the full text of the license or licenses under which it is published, or a detailed reference to them. Therefore public licenses, including FOSS (free, open…

软件工程 · 计算机科学 2023-08-23 Jesús M. González-Barahona , Sergio Montes-Leon , Gregorio Robles , Stefano Zacchiroli

Origami is becoming more and more relevant to research. However, there is no public dataset yet available and there hasn't been any research on this topic in machine learning. We constructed an origami dataset using images from the…

计算机视觉与模式识别 · 计算机科学 2021-01-15 Daniel Ma , Gerald Friedland , Mario Michael Krell

The short note presents an image classification dataset consisting of 10 executable code varieties and approximately 50,000 virus examples. The malicious classes include 9 families of computer viruses and one benign set. The image…

密码学与安全 · 计算机科学 2021-03-02 David Noever , Samantha E. Miller Noever

The rat race between user-generated data and data-processing systems is currently won by data. The increased use of machine learning leads to further increase in processing requirements, while data volume keeps growing. To win the race,…

网络与互联网体系结构 · 计算机科学 2022-06-22 Changgang Zheng , Zhaoqi Xiong , Thanh T Bui , Siim Kaupmees , Riyad Bensoussane , Antoine Bernabeu , Shay Vargaftik , Yaniv Ben-Itzhak , Noa Zilberman

Machine learning classification problems are widespread in bioinformatics, but the technical knowledge required to perform model training, optimization, and inference can prevent researchers from utilizing this technology. This article…

机器学习 · 计算机科学 2023-10-06 Aaron D. Mullen , Samuel E. Armstrong , Jeff Talbert , V. K. Cody Bumgardner

The research presents an overhead view of 10 important objects and follows the general formatting requirements of the most popular machine learning task: digit recognition with MNIST. This dataset offers a public benchmark extracted from…

计算机视觉与模式识别 · 计算机科学 2021-02-09 David Noever , Samantha E. Miller Noever

We introduce MedMNIST v2, a large-scale MNIST-like dataset collection of standardized biomedical images, including 12 datasets for 2D and 6 datasets for 3D. All images are pre-processed into a small size of 28x28 (2D) or 28x28x28 (3D) with…

计算机视觉与模式识别 · 计算机科学 2023-02-20 Jiancheng Yang , Rui Shi , Donglai Wei , Zequan Liu , Lin Zhao , Bilian Ke , Hanspeter Pfister , Bingbing Ni

Recent advancements in Digital Pathology (DP), particularly through artificial intelligence and Foundation Models, have underscored the importance of large-scale, diverse, and richly annotated datasets. Despite their critical role, publicly…

图像与视频处理 · 电气工程与系统科学 2025-05-20 Dmitry Nechaev , Alexey Pchelnikov , Ekaterina Ivanova

The Microsoft Malware Classification Challenge was announced in 2015 along with a publication of a huge dataset of nearly 0.5 terabytes, consisting of disassembly and bytecode of more than 20K malware samples. Apart from serving in the…

密码学与安全 · 计算机科学 2018-03-01 Royi Ronen , Marian Radu , Corina Feuerstein , Elad Yom-Tov , Mansour Ahmadi

With the rise of handy smart phones in the recent years, the trend of capturing selfie images is observed. Hence efficient approaches are required to be developed for recognising faces in selfie images. Due to the short distance between the…

计算机视觉与模式识别 · 计算机科学 2023-02-15 Laxman Kumarapu , Shiv Ram Dubey , Snehasis Mukherjee , Parkhi Mohan , Sree Pragna Vinnakoti , Subhash Karthikeya

We release two artificial datasets, Simulated Flying Shapes and Simulated Planar Manipulator that allow to test the learning ability of video processing systems. In particular, the dataset is meant as a tool which allows to easily assess…

计算机视觉与模式识别 · 计算机科学 2018-07-03 Fabio Ferreira , Jonas Rothfuss , Eren Erdal Aksoy , You Zhou , Tamim Asfour

In this technical report, we present Zyda-2: a five trillion token dataset for language model pretraining. Zyda-2 was used to train our Zamba2 series of models which are state-of-the-art for their weight class. We build Zyda-2 by collating…

计算与语言 · 计算机科学 2024-11-12 Yury Tokpanov , Paolo Glorioso , Quentin Anthony , Beren Millidge

Since its beginning visual recognition research has tried to capture the huge variability of the visual world in several image collections. The number of available datasets is still progressively growing together with the amount of samples…

计算机视觉与模式识别 · 计算机科学 2014-02-25 Tatiana Tommasi , Tinne Tuytelaars , Barbara Caputo

This paper introduces Task 2 of the DCASE2019 Challenge, titled "Audio tagging with noisy labels and minimal supervision". This task was hosted on the Kaggle platform as "Freesound Audio Tagging 2019". The task evaluates systems for…

声音 · 计算机科学 2020-01-22 Eduardo Fonseca , Manoj Plakal , Frederic Font , Daniel P. W. Ellis , Xavier Serra

Image dehazing is an ill-posed problem that has been extensively studied in the recent years. The objective performance evaluation of the dehazing methods is one of the major obstacles due to the lacking of a reference dataset. While the…

计算机视觉与模式识别 · 计算机科学 2020-05-08 Codruta O. Ancuti , Cosmin Ancuti , Radu Timofte

Single image dehazing is an ill-posed problem that has recently drawn important attention. Despite the significant increase in interest shown for dehazing over the past few years, the validation of the dehazing methods remains largely…

计算机视觉与模式识别 · 计算机科学 2019-04-08 Codruta O. Ancuti , Cosmin Ancuti , Mateu Sbert , Radu Timofte

The Virus-MNIST data set is a collection of thumbnail images that is similar in style to the ubiquitous MNIST hand-written digits. These, however, are cast by reshaping possible malware code into an image array. Naturally, it is poised to…

机器学习 · 计算机科学 2021-11-04 Erik Larsen , Korey MacVittie , John Lilly

Understanding satire and humor is a challenging task for even current Vision-Language models. In this paper, we propose the challenging tasks of Satirical Image Detection (detecting whether an image is satirical), Understanding (generating…

计算机视觉与模式识别 · 计算机科学 2024-09-23 Abhilash Nandy , Yash Agarwal , Ashish Patwa , Millon Madhur Das , Aman Bansal , Ankit Raj , Pawan Goyal , Niloy Ganguly
‹ 上一页 1 2 3 10 下一页 ›