The experiments files in the tsml and sktime toolboxes classification output results a specifically formatted csv. Below we provide an example of an results file in this format and describe each of the compontents.
- Line 1 of the results file provides descriptive meta information about the experiment. The values in this line are:
- The dataset name
- The classifier name
- Whether these results are on the test data or a performance estimate using the train data
- The train/test split fold id
- The time unit for any recorded timings
- Line 2 provides classifier specific training and parameter information. This could be the seed, parameters such as number of trees or accuracy found during an internal tuning process.
- Line 3 provides simple summary performance statictics. These are:
- Accuracy
- Train time
- Sum prediction time
- Time to perform a set of benchmark operations for normalising over different hardware
- Maximum runtime memory
- Number of classes
- Train set accuracy estimate method i.e. cross-validation or out of bag error
- Train set estimate time
- Train time plus estimate time
- All remaining lines are details about each prediction made. The first prediction line will correspond to the first series in the data file, the second line to the second series etc. The components of each prediction line are:
- The series true class value
- The series predicted class value
- The predictions probabilities for each class value
- Time to make the prediction