Prisbevakning
Få notis vid prissänkningAv: Rudy Lai, Bartłomiej Potaczek
Lägsta pris
Bokus

357 kr
Amazon
Bokbörsen
Vi har hittat boken hos 2 butiker med verifierade priser — alla är partnerbutiker som vi får provision från när du klickar på ”Visa hos butik”. Vissa butiker visas som extern länk utan pris — priset ser du först hos butiken. Priset för dig är detsamma. Frakt kan tillkomma och varierar mellan butiker och leveranssätt — kontrollera alltid aktuellt pris och leveransvillkor hos butiken innan du slutför köpet.
Skriver du om boken på en blogg eller sajt? .
Priset har nyligen gått ner jämfört med butikens eget tidigare pris.
Det lägsta priset vi sett för boken sedan Booki började mäta.
Billigaste butiken ligger under de övriga butikernas medianpris just nu — en jämförelse mellan butiker, inte ett prisfall över tid.
Butiken med lägst pris i prislistan på boksidan just nu.
Use PySpark to easily crush messy data at-scale and discover proven techniques to create testable, immutable, and easily parallelizable Spark jobsKey FeaturesWork with large amounts of agile data using distributed datasets and in-memory cachingSource data from all popular data hosting platforms, such as HDFS, Hive, JSON, and S3Employ the easy-to-use PySpark API to deploy big data Analytics for productionBook DescriptionApache Spark is an open source parallel-processing framework that has been around for quite some time now. One of the many uses of Apache Spark is for data analytics applications across clustered computers. In this book, you will not only learn how to use Spark and the Python API to create high-performance analytics with big data, but also discover techniques for testing, immunizing, and parallelizing Spark jobs.You will learn how to source data from all popular data hosting platforms, including HDFS, Hive, JSON, and S3, and deal with large datasets with PySpark to gain practical big data experience. This book will help you work on prototypes on local machines and subsequently go on to handle messy data in production and at scale. This book covers installing and setting up PySpark, RDD operations, big data cleaning and wrangling, and aggregating and summarizing data into useful reports. You will also learn how to implement some practical and proven techniques to improve certain aspects of programming and administration in Apache Spark.By the end of the book, you will be able to build big data analytical solutions using the various PySpark offerings and also optimize them effectively.What you will learnGet practical big data experience while working on messy datasetsAnalyze patterns with Spark SQL to improve your business intelligenceUse PySpark's interactive shell to speed up development timeCreate highly concurrent Spark programs by leveraging immutabilityDiscover ways to avoid the most expensive operation in the Spark API: the shuffle operationRe-design your jobs to use reduceByKey instead of groupByCreate robust processing pipelines by testing Apache Spark jobsWho this book is forThis book is for developers, data scientists, business analysts, or anyone who needs to reliably analyze large amounts of large-scale, real-world data. Whether you're tasked with creating your company's business intelligence function or creating great data platforms for your machine learning models, or are looking to use code to magnify the impact of your business, this book is for you.
Bra läge att köpa
Bokus
46 kr billigare
Rör sig ofta
Författare
Rudy Lai, Bartłomiej Potaczek
Författare
Rudy Lai, Bartłomiej Potaczek
Förlag
Packt Publishing Limited
Utgivningsår
2019
Sidantal
182
Språk
Svenska
Fysiska detaljer
illustrations
Dewey
004.2
ISBN
9781838644130
Av: Rudy Lai, Bartłomiej Potaczek
Lägsta pris
Bokus

357 kr
Amazon
Bokbörsen
Vi har hittat boken hos 2 butiker med verifierade priser — alla är partnerbutiker som vi får provision från när du klickar på ”Visa hos butik”. Vissa butiker visas som extern länk utan pris — priset ser du först hos butiken. Priset för dig är detsamma. Frakt kan tillkomma och varierar mellan butiker och leveranssätt — kontrollera alltid aktuellt pris och leveransvillkor hos butiken innan du slutför köpet.
Skriver du om boken på en blogg eller sajt? .
Priset har nyligen gått ner jämfört med butikens eget tidigare pris.
Det lägsta priset vi sett för boken sedan Booki började mäta.
Billigaste butiken ligger under de övriga butikernas medianpris just nu — en jämförelse mellan butiker, inte ett prisfall över tid.
Butiken med lägst pris i prislistan på boksidan just nu.
Use PySpark to easily crush messy data at-scale and discover proven techniques to create testable, immutable, and easily parallelizable Spark jobsKey FeaturesWork with large amounts of agile data using distributed datasets and in-memory cachingSource data from all popular data hosting platforms, such as HDFS, Hive, JSON, and S3Employ the easy-to-use PySpark API to deploy big data Analytics for productionBook DescriptionApache Spark is an open source parallel-processing framework that has been around for quite some time now. One of the many uses of Apache Spark is for data analytics applications across clustered computers. In this book, you will not only learn how to use Spark and the Python API to create high-performance analytics with big data, but also discover techniques for testing, immunizing, and parallelizing Spark jobs.You will learn how to source data from all popular data hosting platforms, including HDFS, Hive, JSON, and S3, and deal with large datasets with PySpark to gain practical big data experience. This book will help you work on prototypes on local machines and subsequently go on to handle messy data in production and at scale. This book covers installing and setting up PySpark, RDD operations, big data cleaning and wrangling, and aggregating and summarizing data into useful reports. You will also learn how to implement some practical and proven techniques to improve certain aspects of programming and administration in Apache Spark.By the end of the book, you will be able to build big data analytical solutions using the various PySpark offerings and also optimize them effectively.What you will learnGet practical big data experience while working on messy datasetsAnalyze patterns with Spark SQL to improve your business intelligenceUse PySpark's interactive shell to speed up development timeCreate highly concurrent Spark programs by leveraging immutabilityDiscover ways to avoid the most expensive operation in the Spark API: the shuffle operationRe-design your jobs to use reduceByKey instead of groupByCreate robust processing pipelines by testing Apache Spark jobsWho this book is forThis book is for developers, data scientists, business analysts, or anyone who needs to reliably analyze large amounts of large-scale, real-world data. Whether you're tasked with creating your company's business intelligence function or creating great data platforms for your machine learning models, or are looking to use code to magnify the impact of your business, this book is for you.
Bra läge att köpa
Bokus
46 kr billigare
Rör sig ofta
Författare
Rudy Lai, Bartłomiej Potaczek
Författare
Rudy Lai, Bartłomiej Potaczek
Förlag
Packt Publishing Limited
Utgivningsår
2019
Sidantal
182
Språk
Svenska
Fysiska detaljer
illustrations
Dewey
004.2
ISBN
9781838644130
2019 · Svenska
analyze large datasets and discover techniques for testing, immunizing, and parallelizing Spark jobs
Rudy Lai, Bartłomiej Potaczek
”24% billigare” visar hur mycket lägre det billigaste priset är än medianpriset hos de övriga butikerna just nu — inte ett tidsbegränsat prisfall.
ISBN 9781838644130 jämförs hos alla butiker
Use PySpark to easily crush messy data at-scale and discover proven techniques to create testable, immutable, and easily parallelizable Spark jobsKey FeaturesWork with large amounts of agile data using distributed datasets and in-memory cachingSource data from all popular data hosting platforms, such as HDFS, Hive, JSON, and S3Employ the easy-to-use PySpark API to deploy big data Analytics for productionBook DescriptionApache Spark is an open source parallel-processing framework that has been around for quite some time now. One of the many uses of Apache Spark is for data analytics applications across clustered computers. In this book, you will not only learn how to use Spark and the Python API to create high-performance analytics with big data, but also discover techniques for testing, immunizing, and parallelizing Spark jobs.You will learn how to source data from all popular data hosting platforms, including HDFS, Hive, JSON, and S3, and deal with large datasets with PySpark to gain practical big data experience. This book will help you work on prototypes on local machines and subsequently go on to handle messy data in production and at scale. This book covers installing and setting up PySpark, RDD operations, big data cleaning and wrangling, and aggregating and summarizing data into useful reports. You will also learn how to implement some practical and proven techniques to improve certain aspects of programming and administration in Apache Spark.By the end of the book, you will be able to build big data analytical solutions using the various PySpark offerings and also optimize them effectively.What you will learnGet practical big data experience while working on messy datasetsAnalyze patterns with Spark SQL to improve your business intelligenceUse PySpark's interactive shell to speed up development timeCreate highly concurrent Spark programs by leveraging immutabilityDiscover ways to avoid the most expensive operation in the Spark API: the shuffle operationRe-design your jobs to use reduceByKey instead of groupByCreate robust processing pipelines by testing Apache Spark jobsWho this book is forThis book is for developers, data scientists, business analysts, or anyone who needs to reliably analyze large amounts of large-scale, real-world data. Whether you're tasked with creating your company's business intelligence function or creating great data platforms for your machine learning models, or are looking to use code to magnify the impact of your business, this book is for you.
Bra läge att köpa
Bokus
46 kr billigare
Rör sig ofta
Författare
Rudy Lai, Bartłomiej Potaczek
Förlag
Packt Publishing Limited
Utgivningsår
2019
Sidantal
182
Språk
Svenska
ISBN
9781838644130
Det lägsta priset just nu är 357 kr hos Bokus, av 2 butiker vi jämför. Priser ändras löpande – kontrollera alltid slutpris och frakt hos butiken innan köp.
Priserna uppdateras automatiskt, vanligtvis minst en gång per dygn. Senaste registrerade uppdatering: 8 juli 2026.
Varje butik sätter sitt eget pris och kör olika kampanjer, så samma bok kan kosta olika mycket. Sverige har fri prissättning på böcker – därför lönar det sig att jämföra, och här ser du priserna samlade på ett ställe.
Nej. Priset vi visar är butikens bokpris – fraktkostnad tillkommer och varierar mellan butiker (flera erbjuder fri frakt över en viss summa). Den slutliga fraktkostnaden ser du i butikens kassa innan du betalar.
Ja. Sätt en kostnadsfri prisbevakning så får du besked när priset faller. Du kan också följa prisutvecklingen i prishistoriken här på sidan.
Mer om butikerna
Läs om frakt, betalning, retur och omdömen för butikerna vi jämför priser hos.