Big Data Analytics with Spark

Big Data Analytics with Spark

Big Data Analytics with Spark is a step-by-step guide for learning Spark, which is an open-source fast and general-purpose cluster computing framework for large-scale data analysis. You will learn how to use Spark for different types of big data analytics projects, including batch, interactive, graph, and stream data analysis as well as machine learning. In addition, this book will help you become a much sought-after Spark expert. Spark is one of the hottest Big Data technologies. The amount of data generated today by devices, applications and users is exploding. Therefore, there is a critical need for tools that can analyze large-scale data and unlock value from it. Spark is a powerful technology that meets that need. You can, for example, use Spark to perform low latency computations through the use of efficient caching and iterative algorithms; leverage the features of its shell for easy and interactive Data analysis; employ its fast batch processing and low latency features to process your real time data streams and so on. As a result, adoption of Spark is rapidly growing and is replacing Hadoop MapReduce as the technology of choice for big data analytics. This book provides an introduction to Spark and related big-data technologies. It covers Spark core and its add-on libraries, including Spark SQL, Spark Streaming, GraphX, and MLlib. Big Data Analytics with Spark is therefore written for busy professionals who prefer learning a new technology from a consolidated source instead of spending countless hours on the Internet trying to pick bits and pieces from different sources. The book also provides a chapter on Scala, the hottest functional programming language, and the program that underlies Spark. You’ll learn the basics of functional programming in Scala, so that you can write Spark applications in it. What's more, Big Data Analytics with Spark provides an introduction to other big data technologies that are commonly used along with Spark, like Hive, Avro, Kafka and so on. So the book is self-sufficient; all the technologies that you need to know to use Spark are covered. The only thing that you are expected to know is programming in any language. There is a critical shortage of people with big data expertise, so companies are willing to pay top dollar for people with skills in areas like Spark and Scala. So reading this book and absorbing its principles will provide a boost—possibly a big boost—to your career.


Author
Publisher Apress
Release Date
ISBN 1484209648
Pages 290 pages
Rating 4/5 (46 users)

More Books:

Big Data Analytics with Spark
Language: en
Pages: 290
Authors: Mohammed Guller
Categories: Computers
Type: BOOK - Published: 2015-12-29 - Publisher: Apress

Big Data Analytics with Spark is a step-by-step guide for learning Spark, which is an open-source fast and general-purpose cluster computing framework for large
Data Analytics with Spark Using Python
Language: en
Pages: 400
Authors: Jeffrey Aven
Categories: Computers
Type: BOOK - Published: 2018-05-28 - Publisher: Addison-Wesley Professional

Spark for Data Professionals introduces and solidifies the concepts behind Spark 2.x, teaching working developers, architects, and data professionals exactly ho
Scala Programming for Big Data Analytics
Language: en
Pages: 306
Authors: Irfan Elahi
Categories: Business & Economics
Type: BOOK - Published: 2019-07-05 - Publisher: Apress

Gain the key language concepts and programming techniques of Scala in the context of big data analytics and Apache Spark. The book begins by introducing you to
Hands-On Big Data Analytics with PySpark
Language: en
Pages: 182
Authors: Rudy Lai
Categories: Computers
Type: BOOK - Published: 2019-03-29 - Publisher: Packt Publishing Ltd

Use PySpark to easily crush messy data at-scale and discover proven techniques to create testable, immutable, and easily parallelizable Spark jobs Key FeaturesW
Advanced Analytics with Spark
Language: en
Pages: 276
Authors: Sandy Ryza
Categories: Computers
Type: BOOK - Published: 2015-04-02 - Publisher: "O'Reilly Media, Inc."

In this practical book, four Cloudera data scientists present a set of self-contained patterns for performing large-scale data analysis with Spark. The authors
Big Data Analytics
Language: en
Pages: 326
Authors: Venkat Ankam
Categories: Computers
Type: BOOK - Published: 2016-09-28 - Publisher: Packt Publishing Ltd

A handy reference guide for data analysts and data scientists to help to obtain value from big data analytics using Spark on Hadoop clusters About This Book Thi
Big Data Analytics
Language: en
Pages: 399
Authors: Arun K. Somani
Categories: Computers
Type: BOOK - Published: 2017-10-30 - Publisher: CRC Press

The proposed book will discuss various aspects of big data Analytics. It will deliberate upon the tools, technology, applications, use cases and research direct
Big Data Processing with Apache Spark
Language: en
Pages: 106
Authors: Srini Penchikala
Categories: Computers
Type: BOOK - Published: 2018-03-13 - Publisher: Lulu.com

Apache Spark is a popular open-source big-data processing framework thatÕs built around speed, ease of use, and unified distributed computing architecture. Not
Practical Big Data Analytics
Language: en
Pages: 412
Authors: Nataraj Dasgupta
Categories: Computers
Type: BOOK - Published: 2018-01-15 - Publisher: Packt Publishing Ltd

Get command of your organizational Big Data using the power of data science and analytics Key Features A perfect companion to boost your Big Data storing, proce
Big Data Analytics with Spark and Hadoop
Language: en
Pages: 309
Authors: Venkat Ankam
Categories:
Type: BOOK - Published: 2016-08-26 - Publisher:

A handy reference guide for data analysts and data scientists to fetch "Value" out of big data analytics using Spark on Hadoop ClustersAbout This Book* Practica