71.5% OFF

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

Original price was: $50.00.Current price is: $14.26.

SKU: advanced-analytics-with-spark-patterns-for-learning-from-data-at-scale-2nd-edition-isbn-13-978-1491972953 Category: Tags: , , , , ,

Description

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

[PDF eBook eTextbook]

  • Publisher: ‎ O’Reilly Media; 2nd edition (July 18, 2017)
  • Language: ‎ English
  • 280 pages
  • ISBN-10: ‎ 9781491972953
  • ISBN-13: ‎ 978-1491972953

In the second edition of this practical book, four Cloudera data scientists present a set of self-contained patterns for performing large-scale data analysis with Spark. The authors bring Spark, statistical methods, and real-world data sets together to teach you how to approach analytics problems by example. Updated for Spark 2.1, this edition acts as an introduction to these techniques and other best practices in Spark programming.

You’ll start with an introduction to Spark and its ecosystem, and then dive into patterns that apply common techniques—including classification, clustering, collaborative filtering, and anomaly detection—to fields such as genomics, security, and finance.

If you have an entry-level understanding of machine learning and statistics, and you program in Java, Python, or Scala, you’ll find the book’s patterns useful for working on your own data applications.

With this book, you will:

  • Familiarize yourself with the Spark programming model
  • Become comfortable within the Spark ecosystem
  • Learn general approaches in data science
  • Examine complete implementations that analyze large public data sets
  • Discover which machine learning tools make sense for particular problems
  • Acquire code that can be adapted to many uses

The first chapter will place Spark within the wider context of data science and big data analytics. After that, each chapter will comprise a self-contained analysis using Spark. The second chapter will introduce the basics of data processing in Spark and Scala through a use case in data cleansing. The next few chapters will delve into the meat and potatoes of machine learning with Spark, applying some of the most common algorithms in canonical applications. The remaining chapters are a bit more of a grab bag and apply Spark in slightly more exotic applications—for example, querying Wikipedia through latent semantic relationships in the text or analyzing genomics data.

Since the first edition, Spark has experienced a major version upgrade that instated an entirely new core API and sweeping changes in subcomponents like MLlib and Spark SQL. In the second edition, we’ve made major renovations to the example code and brought the materials up to date with Spark’s new best practices.

Sandy Ryza develops algorithms for public transit at Remix. Prior, he was a senior data scientist at Cloudera and Clover Health. He is an Apache Spark committer, Apache Hadoop PMC member, and founder of the Time Series for Spark project. He holds the Brown University computer science department’s 2012 Twining award for “Most Chill”.

Uri Laserson is an Assistant Professor of Genetics at the Icahn School of Medicine at Mount Sinai, where he develops scalable technology for genomics and immunology using the Hadoop ecosystem.
Sean Owen is Director of Data Science at Cloudera. He is an ApacheSpark committer and PMC member, and was an Apache Mahout committer.

Josh Wills is the Head of Data Engineering at Slack, the founder of the Apache Crunch project, and wrote a tweet about data scientists once.

What makes us different?

• Instant Download

• Always Competitive Pricing

• 100% Privacy

• FREE Sample Available

• 24-7 LIVE Customer Support

Reviews

There are no reviews yet.

Be the first to review “Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953”
Cart
Principles of Human Physiology (5th Edition) – eBook PDFPrinciples of Human Physiology (5th Edition) – eBook PDF
$9.99
×
Essentials of Stem Cell Biology (3rd Edition) – eBook PDFEssentials of Stem Cell Biology (3rd Edition) – eBook PDF
$16.00
×
Anatomy and Physiology: An Integrative Approach (3rd Edition) – PDF – eBookAnatomy and Physiology: An Integrative Approach (3rd Edition) – PDF – eBook
$11.99
×
Current Developments in Biotechnology and Bioengineering – eBook PDFCurrent Developments in Biotechnology and Bioengineering – eBook PDF
$14.00
×
Psychiatric Mental Health Nursing (8th Edition) – eBookPsychiatric Mental Health Nursing (8th Edition) – eBook
$11.00
×
Biological Science (6th Global Edition) By Scott Freeman – eBookBiological Science (6th Global Edition) By Scott Freeman – eBook
$10.25
×
A Guide to SQL 10th Edition by Mark Shellman, ISBN-13: 978-0357361689A Guide to SQL 10th Edition by Mark Shellman, ISBN-13: 978-0357361689
$74.95
×
Oxford Handbook of Infectious Diseases and Microbiology (2nd Edition) – eBook PDFOxford Handbook of Infectious Diseases and Microbiology (2nd Edition) – eBook PDF
$6.99
×
Vander's human physiology: the mechanisms of body function (15th Edition) – eBookVander's human physiology: the mechanisms of body function (15th Edition) – eBook
$10.00
×
A+ Guide to IT Technical Support 9th Edition by Jean Andrews, ISBN-13: 978-1305266438A+ Guide to IT Technical Support 9th Edition by Jean Andrews, ISBN-13: 978-1305266438
$23.69
×
A Primer on Scientific Programming with Python 5th Edition, ISBN-13: 978-3662498866A Primer on Scientific Programming with Python 5th Edition, ISBN-13: 978-3662498866
$72.75
×
Adobe Illustrator CC Classroom in a Book (2018 release), ISBN-13: 978-0134852492Adobe Illustrator CC Classroom in a Book (2018 release), ISBN-13: 978-0134852492
$14.86
×
A Practical Guide to Linux Commands, Editors, and Shell Programming 4th Edition, ISBN-13: 978-0134774602A Practical Guide to Linux Commands, Editors, and Shell Programming 4th Edition, ISBN-13: 978-0134774602
$18.50
×
Management and Leadership for Nurse Administrators (8th Edition) – eBookManagement and Leadership for Nurse Administrators (8th Edition) – eBook
$10.25
×
Biological Inorganic Chemistry: An Introduction – eBook PDFBiological Inorganic Chemistry: An Introduction – eBook PDF
$8.00
×
A Practical Guide to SysML: The Systems Modeling Language 3rd Edition by Sanford Friedenthal, ISBN-13: 978-0128002025A Practical Guide to SysML: The Systems Modeling Language 3rd Edition by Sanford Friedenthal, ISBN-13: 978-0128002025
$14.75
×
The Personality Puzzle (8th Edition) By David C. Funder - eBookThe Personality Puzzle (8th Edition) By David C. Funder - eBook
$8.00
×
Fundamentals of Anatomy and Physiology (4th edition) – eBook PDFFundamentals of Anatomy and Physiology (4th edition) – eBook PDF
$5.99
×
Campbell Biology: Australian and New Zealand Edition (11th Edition) – eBook PDFCampbell Biology: Australian and New Zealand Edition (11th Edition) – eBook PDF
$9.00
×
Comprehensive Clinical Nephrology (6th Edition) – eBookComprehensive Clinical Nephrology (6th Edition) – eBook
$11.98
×
Feedback Control of Dynamic Systems (8th Edition) -eBookFeedback Control of Dynamic Systems (8th Edition) -eBook
$15.99
×
Human Rights and Personal Self-Defense in International Law – eBook PDFHuman Rights and Personal Self-Defense in International Law – eBook PDF
$5.99
×
Biological Science (6th Edition) By Scott Freeman – eBookBiological Science (6th Edition) By Scott Freeman – eBook
$10.25
×
A Common-Sense Guide to Data Structures and Algorithms 2nd Edition, ISBN-13: 978-1680507225A Common-Sense Guide to Data Structures and Algorithms 2nd Edition, ISBN-13: 978-1680507225
$14.44
×
AI Superpowers: China, Silicon Valley, And The New World Order, ISBN-13: 978-1328546395AI Superpowers: China, Silicon Valley, And The New World Order, ISBN-13: 978-1328546395
$9.99
×
Employment Law for Business (9th Edition) – eBook PDFEmployment Law for Business (9th Edition) – eBook PDF
$12.00
×