Please note: In order to keep Hive up to date and provide users with the best features, we are no longer able to fully support Internet Explorer. The site is still available to you, however some sections of the site may appear broken. We would encourage you to move to a more modern browser like Firefox, Edge or Chrome in order to experience the site fully.

Introducing .NET for Apache Spark : Distributed  Processing for Massive Datasets, Paperback / softback Book

Introducing .NET for Apache Spark : Distributed Processing for Massive Datasets Paperback / softback

Paperback / softback

Description

Get started using Apache Spark via C# or F# and the .NET for Apache Spark bindings.

This book is an introduction to both Apache Spark and the .NET bindings.

Readers new to Apache Spark will get up to speed quickly using Spark for data processing tasks performed against large and very large datasets.

You will learn how to combine your knowledge of .NET with Apache Spark to bring massive computing power to bear by distributed processing of extremely large datasets across multiple servers. This book covers how to get a local instance of Apache Spark running on your developer machine and shows you how to create your first .NET program that uses the Microsoft .NET bindings for Apache Spark.

Techniques shown in the book allow you to use Apache Spark to distribute your data processing tasks over multiple compute nodes.

You will learn to process data using both batch mode and streaming mode so you can make the right choice depending on whether you are processing an existing dataset or are working against new records in micro-batches as they arrive.

The goal of the book is leave you comfortable in bringing the power of Apache Spark to your favorite .NET language.

What You Will LearnInstall and configure Spark .NET on Windows, Linux, and macOS Write Apache Spark programs in C# and F# using the .NET bindingsAccess and invoke the Apache Spark APIs from .NET with the same high performance as Python, Scala, and REncapsulate functionality in user-defined functionsTransform and aggregate large datasets Execute SQL queries against files through Apache HiveDistribute processing of large datasets across multiple serversCreate your own batch, streaming, and machine learning programsWho This Book Is For.NETdevelopers who want to perform big data processing without having to migrate to Python, Scala, or R; and Apache Spark developers who want to run natively on .NET and take advantage of the C# and F# ecosystems

Information

  • Format:Paperback / softback
  • Pages:262 pages, 41 Illustrations, black and white; XV, 262 p. 41 illus.
  • Publisher:APress
  • Publication Date:
  • Category:
  • ISBN:9781484269916

£54.99

 
Free Home Delivery

on all orders

 
Pick up orders

from local bookshops

Information

  • Format:Paperback / softback
  • Pages:262 pages, 41 Illustrations, black and white; XV, 262 p. 41 illus.
  • Publisher:APress
  • Publication Date:
  • Category:
  • ISBN:9781484269916