官术网_书友最值得收藏!

Summary

In this chapter, we covered how to set up Spark locally on our own computer as well as in the cloud as a cluster running on Amazon EC2. You learned how to run Spark on top of Amazon's Elastic Map Reduce (EMR). You also learned how to use Google Compute Engine's Spark Service to create a cluster and run a simple job. We discussed the basics of Spark's programming model and API using the interactive Scala console, and we wrote the same basic Spark program in Scala, Java, R, and Python. We also compared the performance metrics of Hadoop versus Spark for different machine learning algorithms as well as SORT benchmark tests.

In the next chapter, we will consider how to go about using Spark to create a machine learning system.

主站蜘蛛池模板: 河间市| 阿巴嘎旗| 昌图县| 黔南| 山丹县| 兴安县| 江津市| 慈溪市| 武汉市| 泰州市| 开江县| 宽甸| 仁寿县| 甘孜| 宝清县| 平果县| 玛纳斯县| 廊坊市| 南宫市| 莱阳市| 渝北区| 宁明县| 南澳县| 江永县| 通许县| 梓潼县| 延寿县| 容城县| 称多县| 越西县| 瑞丽市| 大安市| 岫岩| 齐河县| 伊吾县| 邓州市| 白河县| 湘潭市| 昌都县| 福州市| 泸水县|