{"product_id":"9781617295522","title":"Spark in Action, Second Edition: Covers Apache Spark 3 with Examples in Java, Python, and Scala","description":"\u003ctable\u003e\u003ctbody\u003e\n\u003ctr\u003e\n\u003ctd style=\"\"\u003e\u003cstrong\u003eAuthor\/Contributor(s):\u003c\/strong\u003e\u003c\/td\u003e\n\u003ctd style=\"\"\u003ePerrin, Jean-Georges\u003cbr\u003e\n\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd style=\"\"\u003e\u003cstrong\u003ePublisher:\u003c\/strong\u003e\u003c\/td\u003e\n\u003ctd\u003eManning\u003cbr\u003e\n\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd style=\"\"\u003e\u003cstrong\u003eDate:\u003c\/strong\u003e\u003c\/td\u003e\n\u003ctd\u003e6\/2\/2020\u003cbr\u003e\n\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd style=\"\"\u003e\u003cstrong\u003eBinding:\u003c\/strong\u003e\u003c\/td\u003e\n\u003ctd style=\"\"\u003ePaperback\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd style=\"\"\u003e\u003cstrong\u003eCondition:\u003c\/strong\u003e\u003c\/td\u003e\n\u003ctd style=\"\"\u003eNEW\u003cbr\u003e\n\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003c\/tbody\u003e\u003c\/table\u003eSummary\u003cbr\u003eThe Spark distributed data processing platform provides an easy-to-implement tool for ingesting, streaming, and processing data from any source. In \u003ci\u003eSpark in Action, Second Edition\u003c\/i\u003e, you’ll learn to take advantage of Spark’s core features and incredible processing speed, with applications including real-time computation, delayed evaluation, and machine learning. Spark skills are a hot commodity in enterprises worldwide, and with Spark’s powerful and flexible Java APIs, you can reap all the benefits without first learning Scala or Hadoop.\u003cbr\u003e\u003cbr\u003eForeword by Rob Thomas.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the technology\u003c\/b\u003e\u003cbr\u003eAnalyzing enterprise data starts by reading, filtering, and merging files and streams from many sources. The Spark data processing engine handles this varied volume like a champ, delivering speeds 100 times faster than Hadoop systems. Thanks to SQL support, an intuitive interface, and a straightforward multilanguage API, you can use Spark without learning a complex new ecosystem.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the book\u003c\/b\u003e\u003cbr\u003e\u003ci\u003eSpark in Action, Second Edition\u003c\/i\u003e, teaches you to create end-to-end analytics applications. In this entirely new book, you’ll learn from interesting Java-based examples, including a complete data pipeline for processing NASA satellite data. And you’ll discover Java, Python, and Scala code samples hosted on GitHub that you can explore and adapt, plus appendixes that give you a cheat sheet for installing tools and understanding Spark-specific terms.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eWhat's inside\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003eWriting Spark applications in Java\u003cbr\u003eSpark application architecture\u003cbr\u003eIngestion through files, databases, streaming, and Elasticsearch\u003cbr\u003eQuerying distributed datasets with Spark SQL\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the reader\u003c\/b\u003e\u003cbr\u003eThis book does not assume previous experience with Spark, Scala, or Hadoop.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eAbout the author\u003c\/b\u003e\u003cbr\u003e\u003cb\u003eJean-Georges Perrin\u003c\/b\u003e is an experienced data and software architect. He is France’s first IBM Champion and has been honored for 12 consecutive years.\u003cbr\u003e\u003cbr\u003e\u003cb\u003eTable of Contents\u003c\/b\u003e\u003cbr\u003e\u003cbr\u003ePART 1 - THE THEORY CRIPPLED BY AWESOME EXAMPLES\u003cbr\u003e\u003cbr\u003e1 So, what is Spark, anyway?\u003cbr\u003e\u003cbr\u003e2 Architecture and flow\u003cbr\u003e\u003cbr\u003e3 The majestic role of the dataframe\u003cbr\u003e\u003cbr\u003e4 Fundamentally lazy\u003cbr\u003e\u003cbr\u003e5 Building a simple app for deployment\u003cbr\u003e\u003cbr\u003e6 Deploying your simple app\u003cbr\u003e\u003cbr\u003ePART 2 - INGESTION\u003cbr\u003e\u003cbr\u003e7 Ingestion from files\u003cbr\u003e\u003cbr\u003e8 Ingestion from databases\u003cbr\u003e\u003cbr\u003e9 Advanced ingestion: finding data sources and building\u003cbr\u003e\u003cbr\u003eyour own\u003cbr\u003e\u003cbr\u003e10 Ingestion through structured streaming\u003cbr\u003e\u003cbr\u003ePART 3 - TRANSFORMING YOUR DATA\u003cbr\u003e\u003cbr\u003e11 Working with SQL\u003cbr\u003e\u003cbr\u003e12 Transforming your data\u003cbr\u003e\u003cbr\u003e13 Transforming entire documents\u003cbr\u003e\u003cbr\u003e14 Extending transformations with user-defined functions\u003cbr\u003e\u003cbr\u003e15 Aggregating your data\u003cbr\u003e\u003cbr\u003ePART 4 - GOING FURTHER\u003cbr\u003e\u003cbr\u003e16 Cache and checkpoint: Enhancing Spark’s performances\u003cbr\u003e\u003cbr\u003e17 Exporting data and building full data pipelines\u003cbr\u003e\u003cbr\u003e18 Exploring deployment","brand":"Manning","offers":[{"title":"Default Title","offer_id":44591465398527,"sku":"9781617295522","price":59.99,"currency_code":"USD","in_stock":false}],"thumbnail_url":"\/\/cdn.shopify.com\/s\/files\/1\/0452\/0886\/2873\/files\/Jacket_0d059562-0da1-42cc-be41-3d28a5477cbe.jpg?v=1771351749","url":"https:\/\/massivebookshop.com\/products\/9781617295522","provider":"MASSIVE BOOKSHOP","version":"1.0","type":"link"}