ChiehFu commented on issue #10115:
URL: https://github.com/apache/hudi/issues/10115#issuecomment-1815044211

   @ad1happy2go got it, thanks for the clarification. 
   
   I tried repartitioning the input df by 100 and observed the parallelism of 
both tagging and indexing stages were set to 100 as well which significantly 
speeded up the indexing performance.
    
   I submitted another issue where slowness was observed in writing stage with 
S3 storage with Hudi 0.12.3, could you please take a look when you have a 
chance? https://github.com/apache/hudi/issues/10121


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]

Reply via email to