[
https://issues.apache.org/jira/browse/HIVE-29724?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=18105813#comment-18105813
]
Stamatis Zampetakis commented on HIVE-29724:
--------------------------------------------
[~dengzh] mentioned that after HIVE-28972 the column descriptors are reused. Is
the feature introduced by HIVE-29694 redundant/useless?
The schemaTool has already many functionalities for managing schema and
metadata information. We could possibly introduce another option (e.g.,
optimize or cleanup) for performing this deduplication task and potentially
others if we find more improvements in the future.
> Provide a Metastore tool for de-duplicating the columns
> -------------------------------------------------------
>
> Key: HIVE-29724
> URL: https://issues.apache.org/jira/browse/HIVE-29724
> Project: Hive
> Issue Type: Improvement
> Reporter: Zhihua Deng
> Priority: Major
>
> Following the discussion on HIVE-29694, we can provide a tool to de-duplicate
> the columns to avoid column metadata bloat, which could happen during
> replication.
--
This message was sent by Atlassian Jira
(v8.20.10#820010)