71.5% OFF

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

Original price was: $50.00.Current price is: $14.26.

SKU: advanced-analytics-with-spark-patterns-for-learning-from-data-at-scale-2nd-edition-isbn-13-978-1491972953 Category: Tags: , , , , ,

Description

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

[PDF eBook eTextbook]

  • Publisher: ‎ O’Reilly Media; 2nd edition (July 18, 2017)
  • Language: ‎ English
  • 280 pages
  • ISBN-10: ‎ 9781491972953
  • ISBN-13: ‎ 978-1491972953

In the second edition of this practical book, four Cloudera data scientists present a set of self-contained patterns for performing large-scale data analysis with Spark. The authors bring Spark, statistical methods, and real-world data sets together to teach you how to approach analytics problems by example. Updated for Spark 2.1, this edition acts as an introduction to these techniques and other best practices in Spark programming.

You’ll start with an introduction to Spark and its ecosystem, and then dive into patterns that apply common techniques—including classification, clustering, collaborative filtering, and anomaly detection—to fields such as genomics, security, and finance.

If you have an entry-level understanding of machine learning and statistics, and you program in Java, Python, or Scala, you’ll find the book’s patterns useful for working on your own data applications.

With this book, you will:

  • Familiarize yourself with the Spark programming model
  • Become comfortable within the Spark ecosystem
  • Learn general approaches in data science
  • Examine complete implementations that analyze large public data sets
  • Discover which machine learning tools make sense for particular problems
  • Acquire code that can be adapted to many uses

The first chapter will place Spark within the wider context of data science and big data analytics. After that, each chapter will comprise a self-contained analysis using Spark. The second chapter will introduce the basics of data processing in Spark and Scala through a use case in data cleansing. The next few chapters will delve into the meat and potatoes of machine learning with Spark, applying some of the most common algorithms in canonical applications. The remaining chapters are a bit more of a grab bag and apply Spark in slightly more exotic applications—for example, querying Wikipedia through latent semantic relationships in the text or analyzing genomics data.

Since the first edition, Spark has experienced a major version upgrade that instated an entirely new core API and sweeping changes in subcomponents like MLlib and Spark SQL. In the second edition, we’ve made major renovations to the example code and brought the materials up to date with Spark’s new best practices.

Sandy Ryza develops algorithms for public transit at Remix. Prior, he was a senior data scientist at Cloudera and Clover Health. He is an Apache Spark committer, Apache Hadoop PMC member, and founder of the Time Series for Spark project. He holds the Brown University computer science department’s 2012 Twining award for “Most Chill”.

Uri Laserson is an Assistant Professor of Genetics at the Icahn School of Medicine at Mount Sinai, where he develops scalable technology for genomics and immunology using the Hadoop ecosystem.
Sean Owen is Director of Data Science at Cloudera. He is an ApacheSpark committer and PMC member, and was an Apache Mahout committer.

Josh Wills is the Head of Data Engineering at Slack, the founder of the Apache Crunch project, and wrote a tweet about data scientists once.

What makes us different?

• Instant Download

• Always Competitive Pricing

• 100% Privacy

• FREE Sample Available

• 24-7 LIVE Customer Support

Reviews

There are no reviews yet.

Be the first to review “Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953”
Cart
The Elements of Statistical Learning: Data Mining, Inference, and Prediction 2nd Edition, ISBN-13: 978-0387848570The Elements of Statistical Learning: Data Mining, Inference, and Prediction 2nd Edition, ISBN-13: 978-0387848570
$13.46
×
Microfluid Mechanics: Principles and Modeling by William W. Liou, ISBN-13: 978-0071443227Microfluid Mechanics: Principles and Modeling by William W. Liou, ISBN-13: 978-0071443227
$14.54
×
Vector Calculus, Linear Algebra, and Differential Forms 5th Edition, ISBN-13: 978-0971576681Vector Calculus, Linear Algebra, and Differential Forms 5th Edition, ISBN-13: 978-0971576681
$28.86
×
Statistics for Engineers and Scientists 5th Edition by William Navidi, ISBN-13: 978-1259717604Statistics for Engineers and Scientists 5th Edition by William Navidi, ISBN-13: 978-1259717604
$20.99
×
Topology 2nd Edition by James Munkres, ISBN-13: 978-0131816299Topology 2nd Edition by James Munkres, ISBN-13: 978-0131816299
$19.99
×
Understanding Analysis 2nd Edition by Stephen Abbott, ISBN-13: 978-1493927111Understanding Analysis 2nd Edition by Stephen Abbott, ISBN-13: 978-1493927111
$15.90
×
The Black Swan: The Impact of the Highly Improbable, ISBN-13: 978-1400063512The Black Swan: The Impact of the Highly Improbable, ISBN-13: 978-1400063512
$9.99
×
The Joy of Abstraction: An Exploration of Math, Category Theory, and Life by Eugenia Cheng, ISBN-13: 978-1108477222The Joy of Abstraction: An Exploration of Math, Category Theory, and Life by Eugenia Cheng, ISBN-13: 978-1108477222
$13.90
×
Starting Out with Python 4th Edition, ISBN-13: 978-0134444321Starting Out with Python 4th Edition, ISBN-13: 978-0134444321
$12.43
×
The Art of Memory Forensics: Detecting Malware and Threats in Windows, Linux, and Mac Memory – PDFThe Art of Memory Forensics: Detecting Malware and Threats in Windows, Linux, and Mac Memory – PDF
$13.35
×
The Singularity Is Near: When Humans Transcend Biology Ray Kurzweil, ISBN-13: 978-0670033843The Singularity Is Near: When Humans Transcend Biology Ray Kurzweil, ISBN-13: 978-0670033843
$8.74
×
Biological Psychology (13th Edition) – eBookBiological Psychology (13th Edition) – eBook
$8.99
×
The Human Services Internship: Getting the Most from Your Experience 4th Edition, ISBN-13: 978-1305087347The Human Services Internship: Getting the Most from Your Experience 4th Edition, ISBN-13: 978-1305087347
$12.43
×
AACN Essentials of Critical Care Nursing 3rd Edition, ISBN-13: 9780071822794AACN Essentials of Critical Care Nursing 3rd Edition, ISBN-13: 9780071822794
$19.46
×
Abnormal Psychology 9th Edition Thomas Oltmanns, ISBN-13: 978-0134899053Abnormal Psychology 9th Edition Thomas Oltmanns, ISBN-13: 978-0134899053
$17.50
×
Geringer’s International Business – eBook PDFGeringer’s International Business – eBook PDF
$11.00
×
The New Regulatory Framework for Consumer Dispute Resolution, ISBN-13: 978-0198766353The New Regulatory Framework for Consumer Dispute Resolution, ISBN-13: 978-0198766353
$14.49
×
Python Crash Course 2nd Edition by Eric Matthes, ISBN-13: 978-1593279288Python Crash Course 2nd Edition by Eric Matthes, ISBN-13: 978-1593279288
$17.99
×
Tort Law: Principles in Practice 3rd Edition by James Underwood, ISBN-13: 978-1543838817Tort Law: Principles in Practice 3rd Edition by James Underwood, ISBN-13: 978-1543838817
$18.85
×
Modern Physics 4th Edition by Kenneth S. Krane, ISBN-13: 978-1119495550Modern Physics 4th Edition by Kenneth S. Krane, ISBN-13: 978-1119495550
$14.88
×
Professional Development of Chemistry Teachers: Theory and Practice, ISBN-13: 978-1782627067Professional Development of Chemistry Teachers: Theory and Practice, ISBN-13: 978-1782627067
$55.89
×
The Organic Chemistry of Drug Synthesis Volume 7 Edition Daniel Lednicer, ISBN-13: 978-0470107508The Organic Chemistry of Drug Synthesis Volume 7 Edition Daniel Lednicer, ISBN-13: 978-0470107508
$12.34
×
The Mathematical Theory of Communication by Claude E Shannon, ISBN-13: 978-1843761846The Mathematical Theory of Communication by Claude E Shannon, ISBN-13: 978-1843761846
$14.99
×
Tourism: Principles and Practice 6th edition Alan Fyall, ISBN-13: 978-1292172354Tourism: Principles and Practice 6th edition Alan Fyall, ISBN-13: 978-1292172354
$15.33
×
Student Solutions Manual for Skoog’s Fundamentals of Analytical Chemistry 9th Edition, ISBN-13: 978-0495558347Student Solutions Manual for Skoog’s Fundamentals of Analytical Chemistry 9th Edition, ISBN-13: 978-0495558347
$11.23
×
Handbook of Inorganic Chemicals 1st Edition Pradyot Patnaik, ISBN-13: 978-0070494398Handbook of Inorganic Chemicals 1st Edition Pradyot Patnaik, ISBN-13: 978-0070494398
$35.64
×
Sample Preparation Techniques in Analytical Chemistry Somenath Mitra, ISBN-13: 978-0471328452Sample Preparation Techniques in Analytical Chemistry Somenath Mitra, ISBN-13: 978-0471328452
$12.23
×
Statistical Mechanics: Algorithms and Computations by Werner Krauth, ISBN-13: 978-0198515364Statistical Mechanics: Algorithms and Computations by Werner Krauth, ISBN-13: 978-0198515364
$19.70
×
Solid State Chemistry and its Applications 2nd Edition, ISBN-13: 978-1119942948Solid State Chemistry and its Applications 2nd Edition, ISBN-13: 978-1119942948
$17.63
×
Photons and Atoms: Introduction to Quantum Electrodynamics, ISBN-13: 978-0471184331Photons and Atoms: Introduction to Quantum Electrodynamics, ISBN-13: 978-0471184331
$19.15
×
Using Aspen Plus in Thermodynamics Instruction: A Step-by-Step Guide by Stanley I. Sandler, ISBN-13: 978-1118996911Using Aspen Plus in Thermodynamics Instruction: A Step-by-Step Guide by Stanley I. Sandler, ISBN-13: 978-1118996911
$14.90
×
Physics for Scientists and Engineers 6th Edition Volume 1 by Paul A. Tipler, ISBN-13: 978-0716789642Physics for Scientists and Engineers 6th Edition Volume 1 by Paul A. Tipler, ISBN-13: 978-0716789642
$14.72
×
Systems Analysis and Design 12th Edition Scott Tilley, ISBN-13: 978-0357117811Systems Analysis and Design 12th Edition Scott Tilley, ISBN-13: 978-0357117811
$16.35
×