Protected: Training 2020-02-27
There is no excerpt because this is a protected post.
Categories
Default There is no excerpt because this is a protected post.
There is no excerpt because this is a protected post.
This post is accessible via garrens.com/DataSnowCat and references material covered at Spark + AI Summit (link) 2019.
Python is the de facto language of data science and engineering, which affords it an outsized community of users. However, when many data scientists and engineers come to Spark with a Python background, unexpected performance potholes can stand in the way of progress. These “Performance Potholes” include PySpark’s ease of integration with existing packages (e.g.… Continue reading