jackylee-ch opened a new pull request, #934: URL: https://github.com/apache/paimon-rust/pull/934
Java exposes `BucketsTable` (`<table>$buckets`) for per-bucket file statistics — the standard way to diagnose bucket skew and decide whether a bucket needs compaction. The DataFusion integration had no equivalent, so that diagnosis was unreachable from DataFusion. This adds the `$buckets` provider: it aggregates the table's data files by `(partition, bucket)` into `record_count`, `file_size_in_bytes`, `file_count` and `last_update_time`, mirroring Java's schema and partition/bucket ordering. Rows come from the same scan `$files` already uses; this only groups them, and reads fail closed under query-auth like the other metadata tables. Tested by aggregating `$files` over the shared fixture and asserting `$buckets` reproduces it exactly. -- This is an automated message from the Apache Git Service. To respond to the message, please log on to GitHub and use the URL above to go to the specific comment. To unsubscribe, e-mail: [email protected] For queries about this service, please contact Infrastructure at: [email protected]
