Hands-On Big Data Analytics with PySpark:
Analyze large datasets and discover techniques for testing, immunizing, and parallelizing Spark jobs
English | 2019 | ISBN: 9781838644130 | 184 Pages | PDF EPUB MOBI (True) | 22 MB
Analyze large datasets and discover techniques for testing, immunizing, and parallelizing Spark jobs
English | 2019 | ISBN: 9781838644130 | 184 Pages | PDF EPUB MOBI (True) | 22 MB
Apache Spark is an open source parallel-processing framework that has been around for quite some time now. One of the many uses of Apache Spark is for data analytics applications across clustered computers. In this book, you will not only learn how to use Spark and the Python API to create high-performance analytics with big data, but also discover techniques for testing, immunizing, and parallelizing Spark jobs.