rl

From MaRDI portal
Dataset:6037083



OpenML44037MaRDI QIDQ6037083

OpenML dataset with id 44037

No author found.

Full work available at URL: https://api.openml.org/data/v1/download/22103125/rl.arff

Upload date: 18 June 2022


Dataset Characteristics

Number of classes: 2
Number of features: 13 (numeric: 5, symbolic: 7 and in total binary: 3 )
Number of instances: 4,970
Number of instances with missing values: 0
Number of missing values: 0

Dataset used in the tabular data benchmark https://github.com/LeoGrin/tabular-benchmark,

                         transformed in the same way. This dataset belongs to the "classification on categorical and
                         numerical features" benchmark. Original description: 

The goal of this challenge is to expose the research community to real world datasets of interest to 4Paradigm. All datasets are formatted in a uniform way, though the type of data might differ. The data are provided as preprocessed matrices, so that participants can focus on classification, although participants are welcome to use additional feature extraction procedures (as long as they do not violate any rule of the challenge). All problems are binary classification problems and are assessed with the normalized Area Under the ROC Curve (AUC) metric (i.e. 2*AUC-1).

                  The identity of the datasets and the type of data is concealed, though its structure is revealed. The final score in  phase 2 will be the average of rankings  on all testing datasets, a ranking will be generated from such results, and winners will be determined according to such ranking.
                  The tasks are constrained by a time budget. The Codalab platform provides computational resources shared by all participants. Each code submission will be exceuted in a compute worker with the following characteristics: 2Cores / 8G Memory / 40G SSD with Ubuntu OS. To ensure the fairness of the evaluation, when a code submission is evaluated, its execution time is limited in time.
                  http://automl.chalearn.org/data