Apache Spark 2 for Beginners

Nonfiction, Computers, Database Management, Data Processing, Application Software, Business Software, Programming
Cover of the book Apache Spark 2 for Beginners by Rajanarayanan Thottuvaikkatumana, Packt Publishing
View on Amazon View on AbeBooks View on Kobo View on B.Depository View on eBay View on Walmart
Author: Rajanarayanan Thottuvaikkatumana ISBN: 9781785886690
Publisher: Packt Publishing Publication: October 14, 2016
Imprint: Packt Publishing Language: English
Author: Rajanarayanan Thottuvaikkatumana
ISBN: 9781785886690
Publisher: Packt Publishing
Publication: October 14, 2016
Imprint: Packt Publishing
Language: English

Develop large-scale distributed data processing applications using Spark 2 in Scala and Python

About This Book

  • This book offers an easy introduction to the Spark framework published on the latest version of Apache Spark 2
  • Perform efficient data processing, machine learning and graph processing using various Spark components
  • A practical guide aimed at beginners to get them up and running with Spark

Who This Book Is For

If you are an application developer, data scientist, or big data solutions architect who is interested in combining the data processing power of Spark from R, and consolidating data processing, stream processing, machine learning, and graph processing into one unified and highly interoperable framework with a uniform API using Scala or Python, this book is for you.

What You Will Learn

  • Get to know the fundamentals of Spark 2 and the Spark programming model using Scala and Python
  • Know how to use Spark SQL and DataFrames using Scala and Python
  • Get an introduction to Spark programming using R
  • Perform Spark data processing, charting, and plotting using Python
  • Get acquainted with Spark stream processing using Scala and Python
  • Be introduced to machine learning using Spark MLlib
  • Get started with graph processing using the Spark GraphX
  • Bring together all that you've learned and develop a complete Spark application

In Detail

Spark is one of the most widely-used large-scale data processing engines and runs extremely fast. It is a framework that has tools that are equally useful for application developers as well as data scientists.

This book starts with the fundamentals of Spark 2 and covers the core data processing framework and API, installation, and application development setup. Then the Spark programming model is introduced through real-world examples followed by Spark SQL programming with DataFrames. An introduction to SparkR is covered next. Later, we cover the charting and plotting features of Python in conjunction with Spark data processing. After that, we take a look at Spark's stream processing, machine learning, and graph processing libraries. The last chapter combines all the skills you learned from the preceding chapters to develop a real-world Spark application.

By the end of this book, you will have all the knowledge you need to develop efficient large-scale applications using Apache Spark.

Style and approach

Learn about Spark's infrastructure with this practical tutorial. With the help of real-world use cases on the main features of Spark we offer an easy introduction to the framework.

View on Amazon View on AbeBooks View on Kobo View on B.Depository View on eBay View on Walmart

Develop large-scale distributed data processing applications using Spark 2 in Scala and Python

About This Book

Who This Book Is For

If you are an application developer, data scientist, or big data solutions architect who is interested in combining the data processing power of Spark from R, and consolidating data processing, stream processing, machine learning, and graph processing into one unified and highly interoperable framework with a uniform API using Scala or Python, this book is for you.

What You Will Learn

In Detail

Spark is one of the most widely-used large-scale data processing engines and runs extremely fast. It is a framework that has tools that are equally useful for application developers as well as data scientists.

This book starts with the fundamentals of Spark 2 and covers the core data processing framework and API, installation, and application development setup. Then the Spark programming model is introduced through real-world examples followed by Spark SQL programming with DataFrames. An introduction to SparkR is covered next. Later, we cover the charting and plotting features of Python in conjunction with Spark data processing. After that, we take a look at Spark's stream processing, machine learning, and graph processing libraries. The last chapter combines all the skills you learned from the preceding chapters to develop a real-world Spark application.

By the end of this book, you will have all the knowledge you need to develop efficient large-scale applications using Apache Spark.

Style and approach

Learn about Spark's infrastructure with this practical tutorial. With the help of real-world use cases on the main features of Spark we offer an easy introduction to the framework.

More books from Packt Publishing

Cover of the book Getting Started with Kubernetes by Rajanarayanan Thottuvaikkatumana
Cover of the book Getting Started with XenDesktop® 7.x by Rajanarayanan Thottuvaikkatumana
Cover of the book OpenCV 3 Computer Vision with Python Cookbook by Rajanarayanan Thottuvaikkatumana
Cover of the book Creating your MySQL Database: Practical Design Tips and Techniques by Rajanarayanan Thottuvaikkatumana
Cover of the book ASP.NET Web API Security Essentials by Rajanarayanan Thottuvaikkatumana
Cover of the book Heroku Cloud Application Development by Rajanarayanan Thottuvaikkatumana
Cover of the book gnuplot Cookbook by Rajanarayanan Thottuvaikkatumana
Cover of the book Oracle WebLogic Server 12c: First Look by Rajanarayanan Thottuvaikkatumana
Cover of the book Apache Axis2 Web Services, 2nd Edition by Rajanarayanan Thottuvaikkatumana
Cover of the book Docker for Serverless Applications by Rajanarayanan Thottuvaikkatumana
Cover of the book Tabular Modeling with SQL Server 2016 Analysis Services Cookbook by Rajanarayanan Thottuvaikkatumana
Cover of the book Learning Python Design Patterns - Second Edition by Rajanarayanan Thottuvaikkatumana
Cover of the book Machine Learning with R Cookbook by Rajanarayanan Thottuvaikkatumana
Cover of the book Mastering Selenium WebDriver 3.0 by Rajanarayanan Thottuvaikkatumana
Cover of the book Effective Amazon Machine Learning by Rajanarayanan Thottuvaikkatumana
We use our own "cookies" and third party cookies to improve services and to see statistical information. By using this website, you agree to our Privacy Policy