Pandas has a core function to_parquet()
. Just write the dataframe to parquet format like this:
df.to_parquet('myfile.parquet')
You still need to install a parquet library such as fastparquet
. If you have more than one parquet library installed, you also need to specify which engine you want pandas to use, otherwise it will take the first one to be installed (as in the documentation). For example:
df.to_parquet('myfile.parquet', engine='fastparquet')
与恶龙缠斗过久,自身亦成为恶龙;凝视深渊过久,深渊将回以凝视…