官术网_书友最值得收藏!

Chapter 3. Deep Dive into Apache Spark

Apache Spark is growing at a fast pace in terms of technology, community, and user base. Two new APIs were introduced in 2015: the DataFrame API and DataSet API. These two APIs are built on top of the core API, which is based on RDDs. It is essential to understand the deeper concepts of RDDs including runtime architecture and behavior on various resource managers of Spark.

This chapter is divided into the following sub topics:

  • Starting Spark daemons
  • Spark core concepts
  • Pairing RDDs
  • The lifecycle of a Spark program
  • Spark applications
  • Persistence and caching
  • Spark resource managers—Standalone, Yarn, and Mesos
主站蜘蛛池模板: 深州市| 沙坪坝区| 德庆县| 甘肃省| 台东市| 容城县| 延庆县| 新竹县| 耒阳市| 秭归县| 丰城市| 宣城市| 科技| 富源县| 华宁县| 东台市| 龙胜| 阳江市| 都昌县| 乌苏市| 庆云县| 本溪| 吴旗县| 巴里| 桃园县| 广州市| 当雄县| 和林格尔县| 乌审旗| 夹江县| 屏东市| 大连市| 崇义县| 松滋市| 江川县| 盐津县| 来凤县| 宜都市| 文成县| 渑池县| 莲花县|