Data
Filter results by:
2018-05-29 19:06:15
Jan van Rijn uploaded data:
No data.
0 runs0 likes0 downloads0 reach0 impact
150 instances - 5 features - 4 classes - 31 missing values
2016-09-21 15:43:02
Joaquin Vanschoren uploaded data:
This is the famous Australian dataset, retrieved 2014-11-14 from the libSVM site. It was normalized. The original version is from…
0 runs0 likes0 downloads0 reach11 impact
690 instances - 15 features - 2 classes - 0 missing values
2016-07-29 21:03:14
Rafael Gomes Mantovani uploaded data:
####1. Summary This database was generated by the Laboratory of Image Processing and Pattern Recognition (INPG-LTIRF) in the development of the Esprit project ELENA No. 6891 and the Esprit working…
0 runs0 likes0 downloads0 reach11 impact
5500 instances - 41 features - 11 classes - 0 missing values
2016-07-29 20:36:10
Rafael Gomes Mantovani uploaded data:
1. Title of Database: LED display domain 2. Sources: (a) Breiman,L., Friedman,J.H., Olshen,R.A., & Stone,C.J. (1984). Classification and Regression Trees. Wadsworth International Group: Belmont,…
0 runs0 likes0 downloads0 reach11 impact
500 instances - 8 features - 10 classes - 0 missing values
2016-04-11 19:57:27
Rafael Gomes Mantovani uploaded data:
####1. Summary This dataset contain attributes of dresses and their recommendations according to their sales.Sales are monitor on the basis of alternate days. The attributes present analyzed are:…
0 runs0 likes0 downloads0 reach11 impact
500 instances - 13 features - 2 classes - 835 missing values
2016-04-11 18:32:46
Joaquin Vanschoren uploaded data:
Data on tree growth used in the Case Study published in the September, 1995 issue of the Canadian Journal of Statistics. This data set was been provided by Dr. Fernando Camacho, Ontario Hydro…
0 runs0 likes0 downloads0 reach11 impact
2796 instances - 34 features - 6 classes - 68100 missing values
2016-03-30 22:09:17
Joaquin Vanschoren uploaded data:
Process delays known as cylinder banding in rotogravure printing were substantially mitigated using control rules discovered by decision tree induction. Attribute Information: > 1. timestamp:…
0 runs0 likes0 downloads0 reach11 impact
540 instances - 38 features - 2 classes - 999 missing values
2016-02-21 21:09:16
Hilda Fabiola Bernard uploaded data:
Abstract: Expression levels of 77 proteins measured in the cerebral cortex of 8 classes of control and Down syndrome mice exposed to context fear conditioning, a task used to assess associative…
0 runs0 likes0 downloads0 reach11 impact
1080 instances - 82 features - 8 classes - 1396 missing values
2016-02-20 13:08:52
Hilda Fabiola Bernard uploaded data:
Additionally, the authors require a citation to one or more publications from those cited as relevant papers. Source: Creators: Renata Cristina Barros Madeo (Madeo, R. C. B.) Priscilla Koch Wagner…
0 runs0 likes0 downloads0 reach11 impact
9873 instances - 33 features - 5 classes - 0 missing values
2016-02-16 15:30:33
Hilda Fabiola Bernard uploaded data:
Source: Rami Mustafa A Mohammad ( University of Huddersfield, rami.mohammad '@' hud.ac.uk, rami.mustafa.a '@' gmail.com) Lee McCluskey (University of Huddersfield,t.l.mccluskey '@' hud.ac.uk ) Fadi…
0 runs0 likes0 downloads0 reach11 impact
11055 instances - 31 features - 2 classes - 0 missing values
2016-01-18 18:45:44
Rafael Gomes Mantovani uploaded data:
Source: Owner of database: Volker Lohweg (University of Applied Sciences, Ostwestfalen-Lippe, volker.lohweg '@' hs-owl.de) Donor of database: Helene Doerksen (University of Applied Sciences,…
0 runs0 likes0 downloads0 reach11 impact
1372 instances - 5 features - 2 classes - 0 missing values
2015-11-09 21:02:59
Rafael Gomes Mantovani uploaded data:
Title: Blood Transfusion Service Center Data Set Abstract: Data taken from the Blood Transfusion Service Center in Hsin-Chu City in Taiwan -- this is a classification problem.…
0 runs0 likes0 downloads0 reach11 impact
748 instances - 5 features - 2 classes - 0 missing values
2015-11-09 20:49:14
Rafael Gomes Mantovani uploaded data:
* Source: Marques de Sá, J.P., jpmdesa '@' gmail.com, Biomedical Engineering Institute, Porto, Portugal. Bernardes, J., joaobern '@' med.up.pt, Faculty of Medicine, University of Porto, Portugal.…
0 runs0 likes0 downloads0 reach11 impact
2126 instances - 36 features - 10 classes - 0 missing values
2015-11-09 20:49:00
Rafael Gomes Mantovani uploaded data:
Source: D. Lucas (ddlucas .at. alum.mit.edu), Lawrence Livermore National Laboratory; R. Klein (rklein .at. astron.berkeley.edu), Lawrence Livermore National Laboratory & U.C. Berkeley; J. Tannahill…
0 runs0 likes0 downloads0 reach11 impact
540 instances - 21 features - 2 classes - 0 missing values
2015-11-09 20:42:10
Rafael Gomes Mantovani uploaded data:
All data is from one continuous EEG measurement with the Emotiv EEG Neuroheadset. The duration of the measurement was 117 seconds. The eye state was detected via a camera during the EEG measurement…
0 runs0 likes0 downloads0 reach11 impact
14980 instances - 15 features - 2 classes - 0 missing values
2015-11-09 20:36:55
Rafael Gomes Mantovani uploaded data:
Source: James P Bridge, Sean B Holden and Lawrence C Paulson University of Cambridge Computer Laboratory William Gates Building 15 JJ Thomson Avenue Cambridge CB3 0FD UK +44 (0)1223 763500…
0 runs0 likes0 downloads0 reach11 impact
6118 instances - 52 features - 6 classes - 0 missing values
2015-11-09 20:32:35
Rafael Gomes Mantovani uploaded data:
Source: 1. Bendi Venkata Ramana, ramana.bendi '@' gmail.com Associate Professor, Department of Information Technology, Aditya Instutute of Technology and Management, Tekkali - 532201, Andhra Pradesh,…
0 runs0 likes0 downloads0 reach11 impact
583 instances - 11 features - 2 classes - 0 missing values
2015-11-09 20:25:43
Rafael Gomes Mantovani uploaded data:
1 . Abstract: Two ground ozone level data sets are included in this collection. One is the eight hour peak set (eighthr.data), the other is the one hour peak set (onehr.data). Those data were…
0 runs0 likes0 downloads0 reach11 impact
2534 instances - 73 features - 2 classes - 0 missing values
2015-11-09 20:25:20
Rafael Gomes Mantovani uploaded data:
* Title: Phoneme dataset * Abstract: The aim of this dataset is to distinguish between nasal (class 0) and oral sounds (class 1). The class distribution is 3,818 samples in class 0 and 1,586 samples…
0 runs0 likes0 downloads0 reach11 impact
5404 instances - 6 features - 2 classes - 0 missing values
2015-11-09 20:21:22
Rafael Gomes Mantovani uploaded data:
1. One-hundred plant species leaves data set (class = margin). 2. Sources: (a) Original owners of colour Leaves Samples: James Cope, Thibaut Beghin, Paolo Remagnino, Sarah Barman. The colour images…
0 runs0 likes0 downloads0 reach11 impact
1600 instances - 65 features - 100 classes - 0 missing values
2015-11-09 20:21:11
Rafael Gomes Mantovani uploaded data:
1. One-hundred plant species leaves data set (class = shape). 2. Sources: (a) Original owners of colour Leaves Samples: James Cope, Thibaut Beghin, Paolo Remagnino, Sarah Barman. The colour images are…
0 runs0 likes0 downloads0 reach11 impact
1600 instances - 65 features - 100 classes - 0 missing values
2015-11-09 20:20:59
Rafael Gomes Mantovani uploaded data:
The data directory contains the binary images (masks) of the leaf samples. The colour images are not included. There are three features: Shape, Margin and Texture. As discussed in the paper(s) above.…
0 runs0 likes0 downloads0 reach11 impact
1599 instances - 65 features - 100 classes - 0 missing values
2015-11-09 20:20:47
Rafael Gomes Mantovani uploaded data:
QSAR biodegradation Data Set * Abstract: Data set containing values for 41 attributes (molecular descriptors) used to classify 1055 chemicals into 2 classes (ready and not ready biodegradable). *…
0 runs0 likes0 downloads0 reach11 impact
1055 instances - 42 features - 2 classes - 0 missing values
2015-11-09 20:19:18
Rafael Gomes Mantovani uploaded data:
* Dataset Title: Wall-Following Robot Navigation Data Data Set * Abstract: The data were collected as the SCITOS G5 robot navigates through the room following the wall in a clockwise direction, for 4…
0 runs0 likes0 downloads0 reach11 impact
5456 instances - 25 features - 4 classes - 0 missing values
2015-11-09 20:18:33
Rafael Gomes Mantovani uploaded data:
Tattile Via Gaetano Donizetti, 1-3-5,25030 Mairano (Brescia), Italy. * Title: Semeion Handwritten Digit Data Set * Abstract: 1593 handwritten digits from around 80 persons were scanned, stretched in a…
0 runs0 likes0 downloads0 reach11 impact
1593 instances - 257 features - 10 classes - 0 missing values
2015-11-09 20:17:37
Rafael Gomes Mantovani uploaded data:
(www.semeion.it) * Title: Steel Plates Faults Data Set * Abstract: A dataset of steel plates' faults, classified into 7 different types. The goal was to train machine learning for automatic pattern…
0 runs0 likes0 downloads0 reach11 impact
1941 instances - 34 features - 2 classes - 0 missing values
2015-11-09 20:17:26
Rafael Gomes Mantovani uploaded data:
* Title: Tamilnadu Electricity Board Hourly Readings Data Set * Abstract: This data can be effectively produced the result to fewer parameter of the Load profile can be reduced in the Database *…
0 runs0 likes0 downloads0 reach11 impact
45781 instances - 4 features - 20 classes - 0 missing values
2015-11-09 20:15:56
Rafael Gomes Mantovani uploaded data:
* Title: Breast Cancer Wisconsin (Diagnostic) Data Set (WDBC) * Abstract: Diagnostic Wisconsin Breast Cancer Database * Source: Creators: 1. Dr. William H. Wolberg, General Surgery Dept. University of…
0 runs0 likes0 downloads0 reach11 impact
569 instances - 31 features - 2 classes - 0 missing values
2015-11-05 00:06:06
Joaquin Vanschoren uploaded data:
Data from the Kaggle Amazon Employee Access Challenge: https://www.kaggle.com/c/amazon-employee-access-challenge When an employee at any company starts work, they first need to obtain the computer…
0 runs0 likes0 downloads0 reach11 impact
32769 instances - 10 features - 2 classes - 0 missing values
2015-11-04 23:59:56
Joaquin Vanschoren uploaded data:
Data from the Kaggle Bioresponse challenge: https://www.kaggle.com/c/bioresponse The objective of the competition is to help us build as good a model as possible so that we can, as optimally as this…
0 runs0 likes0 downloads0 reach11 impact
3751 instances - 1777 features - 2 classes - 0 missing values
2015-10-08 14:43:09
Rafael Gomes Mantovani uploaded data:
* Dataset: Wilt Data Set * Abstract: High-resolution Remote Sensing data set (Quickbird). Small number of training samples of diseased trees, large number for other land cover. Testing data set from…
0 runs0 likes0 downloads0 reach11 impact
4839 instances - 6 features - 2 classes - 0 missing values
2015-09-02 00:49:04
Jan van Rijn updated data:
Restoring dataset file
1. Title of Database: Annealing Data 2. Source Information: donated by David Sterling and Wray Buntine. 3. Past Usage: unknown 4. Relevant Information: -- Explanation: I suspect this was left by Ross…
0 runs0 likes0 downloads0 reach0 impact
898 instances - 39 features - 5 classes - 22175 missing values
2015-06-09 16:56:26
Joaquin Vanschoren updated data:
added target attribute
Prediction task is to determine whether a person makes over 50K a year. Extraction was done by Barry Becker from the 1994 Census database. A set of reasonably clean records was extracted using the…
0 runs0 likes0 downloads0 reach11 impact
48842 instances - 15 features - 2 classes - 6465 missing values
2015-06-02 11:17:44
Farooq Zuberi uploaded data:
libSVM","AAD group #Dataset from the LIBSVM data repository. Preprocessing: transform to two-class
0 runs0 likes0 downloads0 reach10 impact
862 instances - 3 features - 0 classes - 0 missing values
2015-06-01 16:57:25
Rafael Gomes Mantovani uploaded data:
* Dataset Title: MicroMass - Pure (pure spectra version) * Abstract: A dataset to explore machine learning approaches for the identification of microorganisms from mass-spectrometry data. * Source:…
0 runs0 likes0 downloads0 reach11 impact
571 instances - 1301 features - 20 classes - 0 missing values
2015-05-25 19:09:04
Rafael Gomes Mantovani uploaded data:
Relevant Papers: Laurent Candillier and Vincent Lemaire. Design and Analysis of the Nomao Challenge - Active Learning in the Real-World. In: Proceedings of the ALRA : Active Learning in Real-world…
0 runs0 likes0 downloads0 reach11 impact
34465 instances - 119 features - 2 classes - 0 missing values
2015-05-22 23:46:18
Rafael Gomes Mantovani uploaded data:
Abstract: MADELON is an artificial dataset, which was part of the NIPS 2003 feature selection challenge. This is a two-class classification problem with continuous input variables. The difficulty is…
0 runs0 likes0 downloads0 reach11 impact
2600 instances - 501 features - 2 classes - 0 missing values
2015-05-22 21:11:58
Rafael Gomes Mantovani uploaded data:
1. Source: Lee Graham (lee '@' stellaralchemy.com) Franz Oppacher (oppacher '@' scs.carleton.ca) Carleton University, Department of Computer Science Intelligent Systems Research Unit 1125 Colonel By…
0 runs0 likes0 downloads0 reach11 impact
1212 instances - 101 features - 2 classes - 0 missing values
2015-05-22 20:38:11
Rafael Gomes Mantovani uploaded data:
Title: Human Activity Recognition Using Smartphones Abstract: Human Activity Recognition database built from the recordings of 30 subjects performing activities of daily living (ADL) while carrying a…
0 runs0 likes0 downloads0 reach11 impact
10299 instances - 562 features - 6 classes - 0 missing values
2015-05-22 20:11:33
Rafael Gomes Mantovani uploaded data:
Title: Gas Sensor Array Drift Dataset Data Set Source: Creators: Alexander Vergara (vergara '@' ucsd.edu) BioCircutis Institute University of California San Diego San Diego, California, USA Donors of…
0 runs0 likes0 downloads0 reach11 impact
13910 instances - 129 features - 6 classes - 0 missing values
2015-05-21 23:19:32
Rafael Gomes Mantovani uploaded data:
Source: Patrick Marques Ciarelli, pciarelli '@' lcad.inf.ufes.br, Department of Electrical Engineering, Federal University of Espirito Santo Elias Oliveira, elias '@' lcad.inf.ufes.br, Department of…
0 runs0 likes0 downloads0 reach11 impact
1080 instances - 857 features - 9 classes - 0 missing values
2015-05-21 22:16:49
Rafael Gomes Mantovani uploaded data:
Available at: [pdf] http://hdl.handle.net/1822/14838 [bib] http://www3.dsi.uminho.pt/pcortez/bib/2011-esm-1.txt 1. Title: Bank Marketing 2. Sources Created by: Paulo Cortez (Univ. Minho) and Sérgio…
0 runs0 likes0 downloads0 reach11 impact
45211 instances - 17 features - 2 classes - 0 missing values
2015-05-21 20:58:53
Rafael Gomes Mantovani uploaded data:
Dataset artificially generated by using first order theory which describes structure of ten capital letters of English alphabet
0 runs0 likes0 downloads0 reach11 impact
10218 instances - 8 features - 10 classes - 0 missing values
2015-04-15 22:22:58
Tobias Kuehn updated data:
removed ranges after the definition of a numerical attribute
Scene recognition dataset Source: Matthew R. Boutell, Jiebo Luo, Xipeng Shen, and Christopher M. Brown. Learning multi-label scene classification. Pattern Recognition, 37(9):1757-1771, 2004. 1:…
0 runs0 likes0 downloads0 reach11 impact
2407 instances - 300 features - 2 classes - 0 missing values
2015-04-15 17:41:20
Joaquin Vanschoren updated data:
ID is a row id
Dataset from the MLRR repository: http://axon.cs.byu.edu:5000/
0 runs0 likes0 downloads0 reach11 impact
19020 instances - 11 features - 2 classes - 0 missing values
2015-04-15 17:37:23
Joaquin Vanschoren updated data:
ID is a row id
Dataset from the MLRR repository: http://axon.cs.byu.edu:5000/
0 runs0 likes0 downloads0 reach11 impact
6598 instances - 168 features - 2 classes - 0 missing values
2015-04-15 17:08:50
Joaquin Vanschoren updated data:
attribute counter is a row id
The following are data used in an analysis of the Brown and Frown corpora for my doctoral dissertation titled ``Variations in Written English: Characterizing Authors' Rhetorical Language Choices…
0 runs0 likes0 downloads0 reach11 impact
500 instances - 22 features - 15 classes - 0 missing values
2014-11-27 01:26:36
Joaquin Vanschoren uploaded data:
This data is derived from the 2012 KDD Cup. The data is subsampled to 0.1% of the original number of instances, downsampling the majority class (click=0) so that the target feature is reasonably…
0 runs0 likes0 downloads0 reach11 impact
39948 instances - 10 features - 2 classes - 0 missing values
2014-10-30 11:15:44
Joaquin Vanschoren uploaded data:
This dataset represents a set of possible advertisements on Internet pages. The features encode the geometry of the image (if available) as well as phrases occurring in the URL, the image's URL and…
0 runs0 likes0 downloads0 reach10 impact
2014-10-07 00:08:27
Joaquin Vanschoren uploaded data:
Datasets from ACM KDD Cup (http://www.sigkdd.org/kddcup/index.php) KDD Cup 2009 http://www.kddcup-orange.com Converted to ARFF format by TunedIT Customer Relationship Management (CRM) is a key element…
0 runs0 likes0 downloads0 reach11 impact
2014-10-07 00:08:02
Joaquin Vanschoren uploaded data:
The KDD Cup 2009 offers the opportunity to work on large marketing databases from the French Telecom company Orange to predict the propensity of customers to switch provider (churn). Churn (wikipedia…
0 runs0 likes0 downloads0 reach10 impact
2014-10-06 23:57:45
Joaquin Vanschoren uploaded data:
%-*- text -*- %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE Software Engineering Repository data set made publicly available in order to encourage…
0 runs0 likes0 downloads0 reach11 impact
2014-10-06 23:57:43
Joaquin Vanschoren uploaded data:
%-*- text -*- %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE Software Engineering Repository data set made publicly available in order to encourage…
0 runs0 likes0 downloads0 reach11 impact
2109 instances - 22 features - 2 classes - 0 missing values
2014-10-06 23:57:36
Joaquin Vanschoren uploaded data:
%-*- text -*- %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE Software Engineering Repository data set made publicly available in order to encourage…
0 runs0 likes0 downloads0 reach11 impact
522 instances - 22 features - 2 classes - 0 missing values
2014-10-06 23:57:19
Joaquin Vanschoren uploaded data:
This is a PROMISE data set made publicly available in order to encourage repeatable, verifiable, refutable, and/or improvable predictive models of software engineering. If you publish material based…
0 runs0 likes0 downloads0 reach11 impact
10885 instances - 22 features - 2 classes - 25 missing values
2014-10-06 23:57:13
Joaquin Vanschoren uploaded data:
%-*- text -*- %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE data set made publicly available in order to encourage repeatable, verifiable, refutable,…
0 runs0 likes0 downloads0 reach11 impact
1563 instances - 38 features - 2 classes - 0 missing values
2014-10-06 23:57:12
Joaquin Vanschoren uploaded data:
%-*- text -*- %%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE data set made publicly available in order to encourage repeatable, verifiable, refutable,…
0 runs0 likes0 downloads0 reach11 impact
1458 instances - 38 features - 2 classes - 0 missing values
2014-10-06 23:57:07
Joaquin Vanschoren uploaded data:
%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%%% This is a PROMISE Software Engineering Repository data set made publicly available in order to encourage repeatable,…
0 runs0 likes0 downloads0 reach11 impact
15545 instances - 6 features - 2 classes - 0 missing values
2014-10-06 23:56:15
Joaquin Vanschoren uploaded data:
Datasets from the Agnostic Learning vs. Prior Knowledge Challenge (http://www.agnostic.inf.ethz.ch) Dataset from: http://www.agnostic.inf.ethz.ch/datasets.php Modified by TunedIT (converted to ARFF…
0 runs0 likes0 downloads0 reach11 impact
4562 instances - 49 features - 2 classes - 0 missing values
2014-10-06 23:56:01
Joaquin Vanschoren uploaded data:
Datasets from the Agnostic Learning vs. Prior Knowledge Challenge (http://www.agnostic.inf.ethz.ch) Dataset from: http://www.agnostic.inf.ethz.ch/datasets.php Modified by TunedIT (converted to ARFF…
0 runs0 likes0 downloads0 reach11 impact
3468 instances - 971 features - 2 classes - 0 missing values
2014-10-06 23:55:56
Joaquin Vanschoren uploaded data:
Datasets from the Agnostic Learning vs. Prior Knowledge Challenge (http://www.agnostic.inf.ethz.ch) Dataset from: http://www.agnostic.inf.ethz.ch/datasets.php Modified by TunedIT (converted to ARFF…
0 runs0 likes0 downloads0 reach11 impact
2014-10-06 22:07:54
Jan van Rijn updated data:
corrected target
Data from StatLib (ftp stat.cmu.edu/datasets) Data from which conclusions were drawn in the article "Sleep in Mammals: Ecological and Constitutional Correlates" by Allison, T. and Cicchetti, D.…
0 runs0 likes0 downloads0 reach0 impact
2014-10-05 00:03:24
Jan van Rijn updated data:
set ignore feature
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Case number deleted. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric prediction using instance-based learning…
0 runs0 likes0 downloads0 reach0 impact
195 instances - 11 features - 0 classes - 2 missing values
2014-09-28 23:51:27
Joaquin Vanschoren uploaded data:
PRO FOOTBALL SCORES (raw data appears after the description below) How well do the oddsmakers of Las Vegas predict the outcome of professional football games? Is there really a home field advantage -…
0 runs0 likes0 downloads0 reach11 impact
672 instances - 10 features - 2 classes - 1200 missing values
2014-09-28 23:51:25
Joaquin Vanschoren uploaded data:
analcatdata A collection of data sets used in the book "Analyzing Categorical Data," by Jeffrey S. Simonoff, Springer-Verlag, New York, 2003. The submission consists of a zip file containing two…
0 runs0 likes0 downloads0 reach11 impact
797 instances - 5 features - 6 classes - 0 missing values
2014-09-28 23:51:06
Joaquin Vanschoren uploaded data:
analcatdata A collection of data sets used in the book "Analyzing Categorical Data," by Jeffrey S. Simonoff, Springer-Verlag, New York, 2003. The submission consists of a zip file containing two…
0 runs0 likes0 downloads0 reach11 impact
841 instances - 71 features - 4 classes - 0 missing values
2014-09-28 23:50:51
Joaquin Vanschoren uploaded data:
Irish Educational Transitions Data Below are shown data on educational transitions for a sample of 500 Irish schoolchildren aged 11 in 1967. The data were collected by Greaney and Kelleghan (1984),…
0 runs0 likes0 downloads0 reach11 impact
2014-09-27 11:35:53
Joaquin Vanschoren updated data:
set index feature
This data consists of synthetically generated control charts. This dataset contains 600 examples of control charts synthetically generated by the process in Alcock and Manolopoulos (1999). There are…
0 runs0 likes0 downloads0 reach11 impact
2014-09-27 11:19:36
Joaquin Vanschoren updated data:
set target feature
This dataset records 640 time series of 12 LPC cepstrum coefficients taken from nine male speakers. The data was collected for examining our newly developed classifier for multidimensional curves…
0 runs0 likes0 downloads0 reach11 impact
9961 instances - 15 features - 9 classes - 0 missing values
2014-09-27 10:55:59
Joaquin Vanschoren uploaded data:
This file contains 9 sets of sanitized user data drawn from the command histories of 8 UNIX computer users at Purdue over the course of up to 2 years (USER0 and USER1 were generated by the same…
0 runs0 likes0 downloads0 reach10 impact
9100 instances - 3 features - 9 classes - 0 missing values
2014-09-25 23:26:09
Tobias Kuehn updated data:
all numeric variables declared as factors before
This data set was generated as follows. 150 subjects spoke the name of each letter of the alphabet twice. Hence, we have 52 training examples from each speaker. The speakers are grouped into sets of…
0 runs0 likes0 downloads0 reach11 impact
7797 instances - 618 features - 26 classes - 0 missing values
2014-09-22 16:13:44
Jan van Rijn updated data:
set target feature
This data set is also obtained from the task of controlling the ailerons of a F16 aircraft, although the target variable and attributes are different from the ailerons domain. The target variable here…
0 runs0 likes0 downloads0 reach0 impact
9517 instances - 7 features - 0 classes - 0 missing values
2014-09-21 23:04:47
Jan van Rijn updated data:
added special attributes
No data.
0 runs0 likes0 downloads0 reach0 impact
2014-09-19 17:06:29
Jan van Rijn updated data:
Instance_name is an identifier and should be ignored for modelling
Donor: G. Towell, M. Noordewier, and J. Shavlik Primate splice-junction gene sequences (DNA) with associated imperfect domain theory. All examples taken from Genbank 64.1. Categories "ei" and "ie"…
0 runs0 likes0 downloads0 reach0 impact
3190 instances - 61 features - 3 classes - 0 missing values
2014-08-26 17:41:07
Joaquin Vanschoren uploaded data:
The Monk's Problems: Problem 3 This is a merged version of the separate train and test set which are usually distributed. On OpenML this train-test split can be found as one of the possible tasks.…
0 runs0 likes0 downloads0 reach11 impact
554 instances - 7 features - 2 classes - 0 missing values
2014-08-26 17:29:02
Joaquin Vanschoren uploaded data:
The Monk's Problems: Problem 2 This is a merged version of the separate train and test set which are usually distributed. On OpenML this train-test split can be found as one of the possible tasks.…
0 runs0 likes0 downloads0 reach11 impact
601 instances - 7 features - 2 classes - 0 missing values
2014-08-26 17:11:18
Joaquin Vanschoren uploaded data:
The Monk's Problems: Problem 1 This is a merged version of the separate train and test set which are usually distributed. On OpenML this train-test split can be found as one of the possible tasks.…
0 runs0 likes0 downloads0 reach11 impact
556 instances - 7 features - 2 classes - 0 missing values
2014-08-22 16:57:30
Joaquin Vanschoren uploaded data:
In my work on context-sensitive learning, I used the "Deterding Vowel Recognition Data", but I found it necessary to reformulate the data. Implicit in the original data is contextual information on…
0 runs0 likes0 downloads0 reach11 impact
990 instances - 13 features - 11 classes - 0 missing values
2014-04-23 13:17:30
Jan van Rijn uploaded data:
This data set concerns the study of the factors affecting patterns of insulin-dependent diabetes mellitus in children. The objective is to investigate the dependence of the level of serum C-peptide on…
0 runs0 likes0 downloads0 reach0 impact
43 instances - 3 features - 0 classes - 0 missing values
2014-04-23 13:17:28
Jan van Rijn uploaded data:
Data from StatLib (ftp stat.cmu.edu/datasets) The infamous Longley data, "An appraisal of least-squares programs from the point of view of the user", JASA, 62(1967) p819-841. Variables are: Number of…
0 runs0 likes0 downloads0 reach0 impact
16 instances - 7 features - 0 classes - 0 missing values
2014-04-23 13:17:26
Jan van Rijn uploaded data:
Data from StatLib (ftp stat.cmu.edu/datasets) These data are those collected in a cloud-seeding experiment in Tasmania between mid-1964 and January 1971. Their analysis, using regression techniques…
0 runs0 likes0 downloads0 reach0 impact
108 instances - 6 features - 0 classes - 0 missing values
2014-04-23 13:17:24
Jan van Rijn uploaded data:
Dataset from Smoothing Methods in Statistics (ftp stat.cmu.edu/datasets) Simonoff, J.S. (1996). Smoothing Methods in Statistics. New York: Springer-Verlag.
0 runs0 likes0 downloads0 reach0 impact
2178 instances - 4 features - 0 classes - 0 missing values
2014-04-23 13:17:21
Jan van Rijn uploaded data:
Data from StatLib (ftp stat.cmu.edu/datasets) This is the data set called `DETROIT' in the book `Subset selection in regression' by Alan J. Miller published in the Chapman & Hall series of monographs…
0 runs0 likes0 downloads0 reach0 impact
13 instances - 14 features - 0 classes - 0 missing values
2014-04-23 13:17:18
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! All nominal attributes and instances with missing values are deleted. Price treated as the class attribute. As used by…
0 runs0 likes0 downloads0 reach0 impact
159 instances - 16 features - 0 classes - 0 missing values
2014-04-23 13:17:16
Jan van Rijn uploaded data:
The problem is to learn a regression equation/rule/tree to predict the activity from the descriptive structural attributes. The data and methodology is described in detail in: - King, Ross .D., Hurst,…
0 runs0 likes0 downloads0 reach0 impact
2014-04-23 13:17:10
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Cholesterol treated as the class attribute. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric prediction using…
0 runs0 likes0 downloads0 reach0 impact
303 instances - 14 features - 0 classes - 6 missing values
2014-04-23 13:17:07
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Identification code deleted. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric prediction using instance-based…
0 runs0 likes0 downloads0 reach0 impact
189 instances - 10 features - 0 classes - 0 missing values
2014-04-23 13:17:04
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Horsepower treated as the class attribute. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric prediction using…
0 runs0 likes0 downloads0 reach0 impact
2014-04-23 13:17:01
Jan van Rijn uploaded data:
This is a commercial application described in Weiss & Indurkhya (1995). The data describes a telecommunication problem. No further information is available. Characteristics: (10000+5000) cases, 49…
0 runs0 likes0 downloads0 reach0 impact
2014-04-23 13:16:46
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Case number deleted. X treated as the class attribute. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric…
0 runs0 likes0 downloads0 reach0 impact
418 instances - 19 features - 0 classes - 1239 missing values
2014-04-23 13:16:42
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Identifier attribute deleted. !!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! NAME: Sexual activity and the lifespan of male fruitflies TYPE: Designed (almost factorial)…
0 runs0 likes0 downloads0 reach0 impact
125 instances - 5 features - 0 classes - 0 missing values
2014-04-23 13:16:32
Jan van Rijn uploaded data:
The Computer Activity databases are a collection of computer systems activity measures. The data was collected from a Sun Sparcstation 20/712 with 128 Mbytes of memory running in a multi-user…
0 runs0 likes0 downloads0 reach0 impact
8192 instances - 22 features - 0 classes - 0 missing values
2014-04-23 13:16:22
Jan van Rijn uploaded data:
!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!!! Identifier attribute deleted. As used by Kilpatrick, D. & Cameron-Jones, M. (1998). Numeric prediction using instance-based…
0 runs0 likes0 downloads0 reach0 impact
398 instances - 8 features - 0 classes - 6 missing values
2014-04-23 13:16:20
Jan van Rijn uploaded data:
This data set consists of three types of entities: (a) the specification of an auto in terms of various characteristics; (b) its assigned insurance risk rating,; (c) its normalized losses in use as…
0 runs0 likes0 downloads0 reach0 impact
159 instances - 16 features - 0 classes - 0 missing values
2014-04-23 13:16:17
Jan van Rijn uploaded data:
Donor: David W. Aha (aha@ics.uci.edu) This database contains 76 attributes, but all published experiments refer to using a subset of 14 of them. In particular, the Cleveland database is the only one…
0 runs0 likes0 downloads0 reach0 impact
303 instances - 14 features - 0 classes - 6 missing values
2014-04-23 13:16:14
Jan van Rijn uploaded data:
Data from StatLib (ftp stat.cmu.edu/datasets) SUMMARY: Data from an experiment on the affects of machine adjustments on the time to count bolts. Data appear as the STATS (Issue 10) Challenge. DATA:…
0 runs0 likes0 downloads0 reach0 impact
40 instances - 7 features - 0 classes - 0 missing values
2014-04-23 13:16:12
Jan van Rijn uploaded data:
Dataset from Smoothing Methods in Statistics (ftp stat.cmu.edu/datasets) Simonoff, J.S. (1996). Smoothing Methods in Statistics. New York: Springer-Verlag.
0 runs0 likes0 downloads0 reach0 impact
52 instances - 3 features - 0 classes - 0 missing values
2014-04-23 13:16:10
Jan van Rijn uploaded data:
1. Title: Wisconsin Prognostic Breast Cancer (WPBC) 2. Source Information a) Creators: Dr. William H. Wolberg, General Surgery Dept., University of Wisconsin, Clinical Sciences Center, Madison, WI…
0 runs0 likes0 downloads0 reach0 impact
194 instances - 33 features - 0 classes - 0 missing values
2014-04-23 13:16:07
Jan van Rijn uploaded data:
Dataset from Smoothing Methods in Statistics (ftp stat.cmu.edu/datasets) Simonoff, J.S. (1996). Smoothing Methods in Statistics. New York: Springer-Verlag.
0 runs0 likes0 downloads0 reach0 impact
61 instances - 3 features - 0 classes - 0 missing values
2014-04-23 13:16:04
Jan van Rijn uploaded data:
This is data set is concerned with the forward kinematics of an 8 link robot arm. Among the existing variants of this data set we have used the variant 8nm, which is known to be highly non-linear and…
0 runs0 likes0 downloads0 reach0 impact
8192 instances - 9 features - 0 classes - 0 missing values