Deprecated: $wgMWOAuthSharedUserIDs=false is deprecated, set $wgMWOAuthSharedUserIDs=true, $wgMWOAuthSharedUserSource='local' instead [Called from MediaWiki\HookContainer\HookContainer::run in /var/www/html/w/includes/HookContainer/HookContainer.php at line 135] in /var/www/html/w/includes/Debug/MWDebug.php on line 372
guillermo - MaRDI portal

guillermo

From MaRDI portal
Dataset:6035439



OpenML41159MaRDI QIDQ6035439

OpenML dataset with id 41159

No author found.

Full work available at URL: https://api.openml.org/data/v1/download/19335532/guillermo.arff

Upload date: 16 August 2018



Dataset Characteristics

Number of classes: 2
Number of features: 4,297 (numeric: 4,296, symbolic: 1 and in total binary: 1 )
Number of instances: 20,000
Number of instances with missing values: 0
Number of missing values: 0

The goal of this challenge is to expose the research community to real world datasets of interest to 4Paradigm. All datasets are formatted in a uniform way, though the type of data might differ. The data are provided as preprocessed matrices, so that participants can focus on classification, although participants are welcome to use additional feature extraction procedures (as long as they do not violate any rule of the challenge). All problems are binary classification problems and are assessed with the normalized Area Under the ROC Curve (AUC) metric (i.e. 2*AUC-1).

                  The identity of the datasets and the type of data is concealed, though its structure is revealed. The final score in  phase 2 will be the average of rankings  on all testing datasets, a ranking will be generated from such results, and winners will be determined according to such ranking.
                  The tasks are constrained by a time budget. The Codalab platform provides computational resources shared by all participants. Each code submission will be exceuted in a compute worker with the following characteristics: 2Cores / 8G Memory / 40G SSD with Ubuntu OS. To ensure the fairness of the evaluation, when a code submission is evaluated, its execution time is limited in time.
                  http://automl.chalearn.org/data





This page was built for dataset: guillermo