Data
chscase_vine1

chscase_vine1

active ARFF Publicly available Visibility: public Uploaded 04-10-2014 by Joaquin Vanschoren
0 likes downloaded by 0 people , 0 total downloads 0 issues 0 downvotes
Issue #Downvotes for this reason By


Loading wiki
Help us complete this description Edit
Author: Source: Unknown - Date unknown Please cite: File README ----------- chscase A collection of the data sets used in the book "A Casebook for a First Course in Statistics and Data Analysis," by Samprit Chatterjee, Mark S. Handcock and Jeffrey S. Simonoff, John Wiley and Sons, New York, 1995. Submitted by Samprit Chatterjee (schatterjee@stern.nyu.edu), Mark Handcock (mhandcock@stern.nyu.edu) and Jeff Simonoff (jsimonoff@stern.nyu.edu) This submission consists of 38 files, plus this README file. Each file represents a data set analyzed in the book. The names of the files correspond to the names used in the book. The data files are written in plain ASCII (character) text. Missing values are represented by "M" in all data files. More information about the data sets and the book can be obtained via gopher at the address swis.stern.nyu.edu The information is filed under ---> Academic Departments & Research Centers ---> Statistics and Operations Research ---> Publications ---> A Casebook for a First Course in Statistics and Data Analysis ---> Welcome! It can also be accessed from the World Wide Web (WWW) using a WWW browser (e.g., netscape) starting from the URL address http://www.stern.nyu.edu/SOR/Casebook NOTICE: These datasets may be used freely for scientific, educational and/or non-commercial purposes, provided suitable acknowledgment is given (by citing the Chatterjee, Handcock and Simonoff reference above). File: vine1.dat Note: attribute names were generated automatically since there was no information in the data itself. Information about the dataset CLASSTYPE: numeric CLASSINDEX: none specific

10 features

col_10 (target)numeric19 unique values
0 missing
col_1numeric52 unique values
0 missing
col_2numeric25 unique values
0 missing
col_3numeric18 unique values
0 missing
col_4numeric22 unique values
0 missing
col_5numeric15 unique values
0 missing
col_6numeric18 unique values
0 missing
col_7numeric14 unique values
0 missing
col_8numeric14 unique values
0 missing
col_9numeric16 unique values
0 missing

19 properties

52
Number of instances (rows) of the dataset.
10
Number of attributes (columns) of the dataset.
0
Number of distinct values of the target attribute (if it is nominal).
0
Number of missing values in the dataset.
0
Number of instances with at least one value missing.
10
Number of numeric attributes.
0
Number of nominal attributes.
0
Percentage of binary attributes.
0
Percentage of instances having missing values.
0
Percentage of missing values.
-1.18
Average class difference between consecutive instances.
100
Percentage of numeric attributes.
0.19
Number of attributes divided by the number of instances.
0
Percentage of nominal attributes.
Percentage of instances belonging to the most frequent class.
Number of instances belonging to the most frequent class.
Percentage of instances belonging to the least frequent class.
Number of instances belonging to the least frequent class.
0
Number of binary attributes.

13 tasks

0 runs - estimation_procedure: 10 times 10-fold Crossvalidation - evaluation_measure: mean_absolute_error - target_feature: col_10
0 runs - estimation_procedure: 10-fold Crossvalidation - evaluation_measure: mean_absolute_error - target_feature: col_10
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
Define a new task