Parquet, compression and encoding
Measured on the Intel Lab readings:
101 MB CSV → 18 MB Parquet. About 5× smaller, which is also less data to read from disk.
Per column, Parquet applies dictionary encoding for repeated values and run-length encoding for runs, then compresses on top. A mote id repeated thousands of times nearly vanishes.
Apache Parquet, file format