[jira] [Comment Edited] (HIVE-13873) Column pruning for nested fields

Xuefu Zhang (JIRA) Wed, 06 Jul 2016 08:25:38 -0700

    [ 
https://issues.apache.org/jira/browse/HIVE-13873?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=15364453#comment-15364453
 ]


Xuefu Zhang edited comment on HIVE-13873 at 7/6/16 3:24 PM:
------------------------------------------------------------

[~Ferd], thanks for working on this. I will be reviewing it as well. Also, 
could you please attach a patch here too?


was (Author: xuefuz):
[~Ferd], thanks for working on this. I will be reviewing it as well.

> Column pruning for nested fields
> --------------------------------
>
>                 Key: HIVE-13873
>                 URL: https://issues.apache.org/jira/browse/HIVE-13873
>             Project: Hive
>          Issue Type: New Feature
>          Components: Logical Optimizer
>            Reporter: Xuefu Zhang
>            Assignee: Ferdinand Xu
>
> Some columnar file formats such as Parquet store fields in struct type also 
> column by column using encoding described in Google Dramel pager. It's very 
> common in big data where data are stored in structs while queries only needs 
> a subset of the the fields in the structs. However, presently Hive still 
> needs to read the whole struct regardless whether all fields are selected. 
> Therefore, pruning unwanted sub-fields in struct or nested fields at file 
> reading time would be a big performance boost for such scenarios.



--
This message was sent by Atlassian JIRA
(v6.3.4#6332)

[jira] [Comment Edited] (HIVE-13873) Column pruning for nested fields

Reply via email to