[ 
https://issues.apache.org/jira/browse/SPARK-54182?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=18119768#comment-18119768
 ] 

Szehon Ho commented on SPARK-54182:
-----------------------------------

Removing Fix Version/s 4.2.1 during release cleanup: the implementation from PR 
#52897 was subsequently reverted by PR #53661, and both commits are on 
branch-4.2. Leaving this issue reopened for future work.

> Avoid intermediate pandas dataframe creation in df.toPandas without Arrow 
> optimization
> --------------------------------------------------------------------------------------
>
>                 Key: SPARK-54182
>                 URL: https://issues.apache.org/jira/browse/SPARK-54182
>             Project: Spark
>          Issue Type: Sub-task
>          Components: PySpark
>    Affects Versions: 4.2.0
>            Reporter: Ruifeng Zheng
>            Priority: Major
>              Labels: pull-request-available
>
> optimize out the temp pandas dataframe in 
> https://github.com/apache/spark/blob/4966fe99d3e368b933519c4f75e876a40a2d51ea/python/pyspark/sql/pandas/conversion.py#L193-L223



--
This message was sent by Atlassian Jira
(v8.20.10#820010)

---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to