TheR1sing3un opened a new pull request, #10043:
URL: https://github.com/apache/paimon/pull/10043

   ### Purpose
   
   `read_blobs()` and `stream_blobs()` reject `ARRAY<BLOB>` columns even though 
the storage reader supports them. Reading another BLOB column also leaves the 
array's descriptors in the scalar output.
   
   Recognize ARRAY BLOB columns, fetch their elements in the same coalesced 
range read as scalar and MAP BLOBs, and restore row-aligned lists. Preserve 
element order, duplicates, null arrays, empty arrays, null elements, and empty 
or inline payloads. Projection excludes unrequested ARRAY BLOB columns from 
scalar output.
   
   ### Tests
   
   - Multimodal table and contiguous-window suites: 124 passed.
   - Focused BLOB reads on PyArrow 16: 7 passed.
   - Coverage includes mixed scalar/MAP/ARRAY columns, coalesced reads, 
filtering, projection, row IDs, limits, and empty results.
   - Flake8, changed-file license checks, Python 3.6 grammar checks, and `git 
diff --check` passed.
   


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]

Reply via email to