[
https://issues.apache.org/jira/browse/SPARK-29170?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
]
Maxim Gekk updated SPARK-29170:
-------------------------------
Description: The updated benchmark results show performance regression *(2
times slower)* in CSV/JSON datasources on partitioned tables, see
https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
. Need to find out the root cause of this and fix it. (was: The updated
benchmark results show performance regression in CSV/JSON datasources on
partitioned tables, see
https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
. Need to find out the root cause of this and fix it.)
> Performance regression in CSV/JSON on partitioned tables
> --------------------------------------------------------
>
> Key: SPARK-29170
> URL: https://issues.apache.org/jira/browse/SPARK-29170
> Project: Spark
> Issue Type: Improvement
> Components: SQL
> Affects Versions: 3.0.0
> Reporter: Maxim Gekk
> Priority: Major
>
> The updated benchmark results show performance regression *(2 times slower)*
> in CSV/JSON datasources on partitioned tables, see
> https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
> . Need to find out the root cause of this and fix it.
--
This message was sent by Atlassian Jira
(v8.3.4#803005)
---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]