Data
chscase_whale

chscase_whale

active ARFF Publicly available Visibility: public Uploaded 04-10-2014 by Joaquin Vanschoren
0 likes downloaded by 0 people , 0 total downloads 0 issues 0 downvotes
  • Chemistry Life Science StatLib
Issue #Downvotes for this reason By


Loading wiki
Help us complete this description Edit
Author: Source: Unknown - Date unknown Please cite: File README ----------- chscase A collection of the data sets used in the book "A Casebook for a First Course in Statistics and Data Analysis," by Samprit Chatterjee, Mark S. Handcock and Jeffrey S. Simonoff, John Wiley and Sons, New York, 1995. Submitted by Samprit Chatterjee (schatterjee@stern.nyu.edu), Mark Handcock (mhandcock@stern.nyu.edu) and Jeff Simonoff (jsimonoff@stern.nyu.edu) This submission consists of 38 files, plus this README file. Each file represents a data set analyzed in the book. The names of the files correspond to the names used in the book. The data files are written in plain ASCII (character) text. Missing values are represented by "M" in all data files. More information about the data sets and the book can be obtained via gopher at the address swis.stern.nyu.edu The information is filed under ---> Academic Departments & Research Centers ---> Statistics and Operations Research ---> Publications ---> A Casebook for a First Course in Statistics and Data Analysis ---> Welcome! It can also be accessed from the World Wide Web (WWW) using a WWW browser (e.g., netscape) starting from the URL address http://www.stern.nyu.edu/SOR/Casebook NOTICE: These datasets may be used freely for scientific, educational and/or non-commercial purposes, provided suitable acknowledgment is given (by citing the Chatterjee, Handcock and Simonoff reference above). File: whale.dat Note: attribute names were generated automatically since there was no information in the data itself. Information about the dataset CLASSTYPE: numeric CLASSINDEX: none specific

0 features

col_12 (target)numeric2 unique values
5 missing
col_1 (row identifier)numeric228 unique values
0 missing
col_2numeric19 unique values
0 missing
col_3numeric156 unique values
0 missing
col_4numeric34 unique values
0 missing
col_6numeric2 unique values
0 missing
col_7numeric223 unique values
5 missing
col_8numeric18 unique values
5 missing
col_9numeric183 unique values
5 missing
col_10numeric38 unique values
5 missing

0 properties

Data properties are not analyzed yet. Refresh the page in a few minutes.

13 tasks

0 runs - estimation_procedure: 10-fold Crossvalidation - evaluation_measure: mean_absolute_error - target_feature: col_12
0 runs - estimation_procedure: 10 times 10-fold Crossvalidation - evaluation_measure: mean_absolute_error - target_feature: col_12
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
0 runs - estimation_procedure: 50 times Clustering
Define a new task