ARFF
Serialize splits DataFrames (see src.process_dataset.splitting) to
the OpenML splits ARFF format.
arff_head(text, n=15)
First n lines of an ARFF string — handy for quick inspection.
Source code in src/process_dataset/arff.py
48 49 50 | |
save_splits_arff(splits, path, relation='splits')
Write a splits DataFrame to path as OpenML-format ARFF.
Source code in src/process_dataset/arff.py
37 38 39 40 41 42 43 44 45 | |
splits_to_arff(splits, relation='splits')
Serialize a splits DataFrame to an OpenML-format ARFF string.
Java's ArffMapping emits a 4- or 5-column ARFF (type, rowid,
repeat, fold, optional sample). We reproduce that schema with
liac-arff so the output is byte-compatible with what the OpenML server
accepts. type is nominal {TRAIN, TEST}; every other column is
NUMERIC. Columns are emitted in declaration order regardless of
DataFrame column order.
Source code in src/process_dataset/arff.py
10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 | |