[ 
https://issues.apache.org/jira/browse/SPARK-29170?page=com.atlassian.jira.plugin.system.issuetabpanels:all-tabpanel
 ]

Maxim Gekk updated SPARK-29170:
-------------------------------
    Description: The updated benchmark results show performance regression *(2 
times slower)* in CSV/JSON datasources on partitioned tables, see 
https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
 . Need to find out the root cause of this and fix it.  (was: The updated 
benchmark results show performance regression in CSV/JSON datasources on 
partitioned tables, see 
https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
 . Need to find out the root cause of this and fix it.)

> Performance regression in CSV/JSON on partitioned tables
> --------------------------------------------------------
>
>                 Key: SPARK-29170
>                 URL: https://issues.apache.org/jira/browse/SPARK-29170
>             Project: Spark
>          Issue Type: Improvement
>          Components: SQL
>    Affects Versions: 3.0.0
>            Reporter: Maxim Gekk
>            Priority: Major
>
> The updated benchmark results show performance regression *(2 times slower)* 
> in CSV/JSON datasources on partitioned tables, see 
> https://github.com/apache/spark/pull/25828/files#diff-e37c3433287c7371ff9c2404db821095R160
>  . Need to find out the root cause of this and fix it.



--
This message was sent by Atlassian Jira
(v8.3.4#803005)

---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to