搜索
人大经济论坛 附件下载

附件下载

所在主题:
文件名:  Packt.Taming.Big.Data.with.Apache.Spark.and.Python.1787287947_Code.zip
资料下载链接地址: https://bbs.pinggu.org/a-2286159.html
附件大小:
950.09 KB   举报本内容
  • Title: Taming Big Data with Apache Spark and Python – Hands On!
  • Author: Frank Kane
  • Length: 81 pages
  • Edition: 1
  • Language: English
  • Publisher: Packt Publishing
  • Publication Date: 2017-07-06
  • ISBN-10: 1787287947
  • ISBN-13: 9781787287945







Key Features
  • Understand how Spark can be distributed across computing clusters
  • Develop and run Spark jobs efficiently using Python
  • A hands-on tutorial with over 15 real-world examples teaching you Big Data processing with Spark

Book Description
Apache Spark has emerged as the next big thing in the Big Data domain - quickly rising from an ascending technology to an established superstar in just a matter of years. Spark allows you to quickly extract actionable insights from large amounts of data, on a real-time basis. This book is your companion to learn Apache Spark in a hands-on manner. Start with understanding how to set up Spark on a single system or on a cluster. From analyzing large data sets using Spark RDD to developing and running effective Spark jobs quickly using Python, this course will teach you everything. Packed with over 15 interactive, fun-filled examples relevant to the real-world, the course will empower you to understand the Spark ecosystem and implement production-grade real-time Spark projects with ease.


What you will learn
  • Learn how you can identify the Big Data problems as Spark problems
  • Install and run Apache Spark on your computer or on a cluster
  • Analyze large data sets across many CPUs using Spark's Resilient Distributed Datasets
  • Implement machine learning on Spark using the MLlib library
  • Process continuos streams of data in real time using the Spark streaming module
  • Perform complex network analysis using Spark's GraphX library
  • Use Amazon's Elastic MapReduce service to run your Spark jobs on a cluster

Table of Contents
Chapter 1. Getting Started with Spark
Chapter 2. Spark Basics and Spark Examples
Chapter 3. Advanced Examples of Spark Programs
Chapter 4. Running Spark on a Cluster
Chapter 5. SparkSQL, DataFrames, and DataSets
Chapter 6. Other Spark Technologies and Libraries
Chapter 7. Where to Go From Here? – Learning More About Spark and Data Science



    熟悉论坛请点击新手指南
下载说明
1、论坛支持迅雷和网际快车等p2p多线程软件下载,请在上面选择下载通道单击右健下载即可。
2、论坛会定期自动批量更新下载地址,所以请不要浪费时间盗链论坛资源,盗链地址会很快失效。
3、本站为非盈利性质的学术交流网站,鼓励和保护原创作品,拒绝未经版权人许可的上传行为。本站如接到版权人发出的合格侵权通知,将积极的采取必要措施;同时,本站也将在技术手段和能力范围内,履行版权保护的注意义务。
(如有侵权,欢迎举报)
二维码

扫码加我 拉你入群

请注明:姓名-公司-职位

以便审核进群资格,未注明则拒绝

GMT+8, 2025-12-29 07:28