snicholasbarton opened a new pull request, #18398: URL: https://github.com/apache/iceberg/pull/18398
When examining a JFR of our process that writes Parquet files for Iceberg, we saw that for every file we were creating a new Hadoop Configuration object and trying to parse the core-default.xml file every time as well. This isn't needed for non-Hadoop files, and in fact is already handled for Parquet reads [here](https://github.com/apache/iceberg/blob/f74aea1e68fc8161905a748f069c055f47ca64b5/parquet/src/main/java/org/apache/iceberg/parquet/Parquet.java#L1532-L1541). This change switches the Parquet.WriteBuilder to accept ParquetConfiguration instead, allowing us to skip all the xml parsing. -- This is an automated message from the Apache Git Service. To respond to the message, please log on to GitHub and use the URL above to go to the specific comment. To unsubscribe, e-mail: [email protected] For queries about this service, please contact Infrastructure at: [email protected] --------------------------------------------------------------------- To unsubscribe, e-mail: [email protected] For additional commands, e-mail: [email protected]
