chrevanthreddy commented on PR #19309: URL: https://github.com/apache/hudi/pull/19309#issuecomment-5482115323
RFC review closeout update: - Replied to and resolved all 23 review threads against the corresponding landed sections (limitations, terminology, generation/shard invariants, bootstrap recovery, engine scope, stale-position recovery, OCC/NBCC matrix, LIRE scope, and Appendix A evidence). - Draft implementation PR: #19802. - Independent core Spark task-retry/RLI correctness fix: #19801. - Corrected BIGANN 1B evidence is recorded in Appendix A; the retired-key “approximate” artifact remains explicitly excluded. - Final implementation acceptance is still running: corrected 10M COW/MOR lifecycle + physical 512-file-group proof, followed by any required corrected 1B rebuild and bounded 1B mutation test. I have kept the RFC’s claims conservative: the current 1B query numbers describe the measured 16,384-physical-group layout, not the intended 512-group layout, and the independent 1B RLI audit remains open because both attempted Hudi-reader validators bottlenecked on ten giant RLI HFile input partitions. -- This is an automated message from the Apache Git Service. To respond to the message, please log on to GitHub and use the URL above to go to the specific comment. To unsubscribe, e-mail: [email protected] For queries about this service, please contact Infrastructure at: [email protected]
