TheR1sing3un opened a new pull request, #10043: URL: https://github.com/apache/paimon/pull/10043
### Purpose `read_blobs()` and `stream_blobs()` reject `ARRAY<BLOB>` columns even though the storage reader supports them. Reading another BLOB column also leaves the array's descriptors in the scalar output. Recognize ARRAY BLOB columns, fetch their elements in the same coalesced range read as scalar and MAP BLOBs, and restore row-aligned lists. Preserve element order, duplicates, null arrays, empty arrays, null elements, and empty or inline payloads. Projection excludes unrequested ARRAY BLOB columns from scalar output. ### Tests - Multimodal table and contiguous-window suites: 124 passed. - Focused BLOB reads on PyArrow 16: 7 passed. - Coverage includes mixed scalar/MAP/ARRAY columns, coalesced reads, filtering, projection, row IDs, limits, and empty results. - Flake8, changed-file license checks, Python 3.6 grammar checks, and `git diff --check` passed. -- This is an automated message from the Apache Git Service. To respond to the message, please log on to GitHub and use the URL above to go to the specific comment. To unsubscribe, e-mail: [email protected] For queries about this service, please contact Infrastructure at: [email protected]
