mzabaluev commented on PR #10775: URL: https://github.com/apache/arrow-rs/pull/10775#issuecomment-5523766294
The initial motivation for #9699 was to bring the output closer to one produced by parquet-java, so that our native parquet writer does not produce drastically larger storage load than Spark, which would negatively affect customers migrating from it. Telling the customers to perform tuning to get compression that Spark achieves automatically (albeit with its own pitfalls) is not a satisfactory option. -- This is an automated message from the Apache Git Service. To respond to the message, please log on to GitHub and use the URL above to go to the specific comment. To unsubscribe, e-mail: [email protected] For queries about this service, please contact Infrastructure at: [email protected]
