Introducing .NET for Apache Spark: Distributed Processing for Massive Datasets

iebukes Apress 241 次浏览 没有评论
Introducing .NET for Apache Spark: Distributed Processing for Massive Datasets Front Cover

Introducing .NET for Apache Spark: Distributed Processing for Massive Datasets

by Ed Elliott
  • Length: 277 pages
  • Edition: 1
  • Publisher: Apress
  • Publication Date: 2021-04-28
  • ISBN-10: 1484269918
  • ISBN-13: 9781484269916
Description

Get started using Apache Spark via C# or F# and the .NET for Apache Spark bindings. This book is an introduction to both Apache Spark and the .NET bindings. Readers new to Apache Spark will get up to speed quickly using Spark for data processing tasks performed against large and very large datasets. You will learn how to combine your knowledge of .NET with Apache Spark to bring massive computing power to bear by distributed processing of extremely large datasets across multiple servers.

This book covers how to get a local instance of Apache Spark running on your developer machine and shows you how to create your first .NET program that uses the Microsoft .NET bindings for Apache Spark. Techniques shown in the book allow you to use Apache Spark to distribute your data processing tasks over multiple compute nodes. You will learn to process data using both batch mode and streaming mode so you can make the right choice depending on whether you are processing an existing dataset or are working against new records in micro-batches as they arrive. The goal of the book is leave you comfortable in bringing the power of Apache Spark to your favorite .NET language.

What You Will Learn

  • Install and configure Spark .NET on Windows, Linux, and macOS
  • Write Apache Spark programs in C# and F# using the .NET bindings
  • Access and invoke the Apache Spark APIs from .NET with the same high performance as Python, Scala, and R
  • Encapsulate functionality in user-defined functions
  • Transform and aggregate large datasets
  • Execute SQL queries against files through Apache Hive
  • Distribute processing of large datasets across multiple servers
  • Create your own batch, streaming, and machine learning programs

Who This Book Is For

.NET developers who want to perform big data processing without having to migrate to Python, Scala, or R; and Apache Spark developers who want to run natively on .NET and take advantage of the C# and F# ecosystems

Introducing .NET for Apache Spark: Distributed Processing for Massive Datasets

 

 亲,网盘文件已删,下载链接已失效


因为,我,失业了!于是我老家十八线小县城找了份掏下水道的工作。。。
 
为了生活
 
我决定将iebueks电子网站由免费改为赞助入群:
 
一年45元
 
从百度网盘群满之日算起。
 
这45元除了最新的英文IT电子书,还包括:

免费找书服务,中文英文皆可

国内出版社出版的中文电子书  
中文电子书
2022年公考资料
2022年公考资料
2023年考研学习资料
2023年考研学习资料
人人素材网各种视频素材模板以及中文字幕教程
人人素材网

入群指南


扫描下面二维码关注微信公众号获取资源

微信公众号二维码

发表评论

您的电子邮箱地址不会被公开。 必填项已用*标注

Go