楼主: Nicolle
1121 9

Introducing .NET for Apache Spark: Distributed Processing for Massive Datasets [推广有奖]

巨擘

0%

还不是VIP/贵宾

-

TA的文库  其他...

Python(Must-Read Books)

SAS Programming

Must-Read Books

威望
16
论坛币
12402323 个
通用积分
1620.8615
学术水平
3305 点
热心指数
3329 点
信用等级
3095 点
经验
477211 点
帖子
23879
精华
91
在线时间
9878 小时
注册时间
2005-4-23
最后登录
2022-3-6

楼主
Nicolle 学生认证  发表于 2021-5-15 20:29:15 |只看作者 |坛友微信交流群|倒序 |AI写论文
提示: 作者被禁止或删除 内容自动屏蔽

本帖被以下文库推荐

沙发
xxka917 发表于 2021-5-15 20:29:16 |只看作者 |坛友微信交流群

本帖隐藏的内容

Introducing .NET for Apache Spark Distributed Processing for Massive Datasets by.pdf (4.49 MB, 需要: 30 个论坛币)


Get started using Apache Spark via C# or F# and the .NET for Apache Spark bindings. This book is an introduction to both Apache Spark and the .NET bindings. Readers new to Apache Spark will get up to speed quickly using Spark for data processing tasks performed against large and very large datasets. You will learn how to combine your knowledge of .NET with Apache Spark to bring massive computing power to bear by distributed processing of extremely large datasets across multiple servers.

This book covers how to get a local instance of Apache Spark running on your developer machine and shows you how to create your first .NET program that uses the Microsoft .NET bindings for Apache Spark. Techniques shown in the book allow you to use Apache Spark to distribute your data processing tasks over multiple compute nodes. You will learn to process data using both batch mode and streaming mode so you can make the right choice depending on whether you are processing an existing dataset or are working against new records in micro-batches as they arrive. The goal of the book is leave you comfortable in bringing the power of Apache Spark to your favorite .NET language.

What You Will Learn
* Install and configure Spark .NET on Windows, Linux, and macOS
* Write Apache Spark programs in C# and F# using the .NET bindings
* Access and invoke the Apache Spark APIs from .NET with the same high performance as Python, Scala, and R
* Encapsulate functionality in user-defined functions
* Transform and aggregate large datasets
* Execute SQL queries against files through Apache Hive
* Distribute processing of large datasets across multiple servers
* Create your own batch, streaming, and machine learning programs

Who This Book Is For
.NET developers who want to perform big data processing without having to migrate to Python, Scala, or R; and Apache Spark developers who want to run natively on .NET and take advantage of the C# and F# ecosystems



使用道具

藤椅
auirzxp 学生认证  发表于 2021-5-15 22:51:06 |只看作者 |坛友微信交流群

使用道具

板凳
HappyAndy_Lo 发表于 2021-5-17 23:03:36 |只看作者 |坛友微信交流群

使用道具

报纸
Nicolle 学生认证  发表于 2021-5-22 04:13:11 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

地板
Nicolle 学生认证  发表于 2021-5-22 04:14:17 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

7
Nicolle 学生认证  发表于 2021-5-22 04:16:40 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

8
Nicolle 学生认证  发表于 2021-5-22 04:18:05 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

9
Nicolle 学生认证  发表于 2021-5-22 04:20:35 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

10
Nicolle 学生认证  发表于 2021-5-22 04:22:01 |只看作者 |坛友微信交流群
提示: 作者被禁止或删除 内容自动屏蔽

使用道具

您需要登录后才可以回帖 登录 | 我要注册

本版微信群
加好友,备注jltj
拉您入交流群

京ICP备16021002-2号 京B2-20170662号 京公网安备 11010802022788号 论坛法律顾问:王进律师 知识产权保护声明   免责及隐私声明

GMT+8, 2024-4-19 11:15