20% Discount with Use Code SAVEON20
  • Cart
  • Contact us
  • FAQ
logo01 univebook
Login / Register
Wishlist
0 Compare
5 items $54.39
Menu
logo01 univebook
5 items $54.39
  • Home
  • Shop
  • My account
  • Blog
  • About us
  • Contact us
  • Request an eBook
“The Singularity Is Near: When Humans Transcend Biology Ray Kurzweil, ISBN-13: 978-0670033843” has been added to your cart. View cart
-71%
Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953
Click to enlarge
Home Computing Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953
Fundamentals of Nuclear Science and Engineering 3rd Edition, ISBN-13: 978-1498769297
Fundamentals of Nuclear Science and Engineering 3rd Edition, ISBN-13: 978-1498769297 $50.00 Original price was: $50.00.$12.32Current price is: $12.32.
Back to products
Architecting the Cloud: Design Decisions for Cloud Computing Service Models by Michael J. Kavis, ISBN-13: 978-8126550333
Architecting the Cloud: Design Decisions for Cloud Computing Service Models by Michael J. Kavis, ISBN-13: 978-8126550333 $50.00 Original price was: $50.00.$14.60Current price is: $14.60.

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

$50.00 Original price was: $50.00.$14.26Current price is: $14.26.

Compare
Add to wishlist
SKU: advanced-analytics-with-spark-patterns-for-learning-from-data-at-scale-2nd-edition-isbn-13-978-1491972953 Category: Computing Tags: Advanced Analytics with Spark Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953, Josh Wills, Sandy Ryza, Sean Owen, Uri Laserson
Share:
  • Description
  • Reviews (0)
  • Shipping & Delivery
Description

Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953

[PDF eBook eTextbook]

  • Publisher: ‎ O’Reilly Media; 2nd edition (July 18, 2017)
  • Language: ‎ English
  • 280 pages
  • ISBN-10: ‎ 9781491972953
  • ISBN-13: ‎ 978-1491972953

In the second edition of this practical book, four Cloudera data scientists present a set of self-contained patterns for performing large-scale data analysis with Spark. The authors bring Spark, statistical methods, and real-world data sets together to teach you how to approach analytics problems by example. Updated for Spark 2.1, this edition acts as an introduction to these techniques and other best practices in Spark programming.

You’ll start with an introduction to Spark and its ecosystem, and then dive into patterns that apply common techniques—including classification, clustering, collaborative filtering, and anomaly detection—to fields such as genomics, security, and finance.

If you have an entry-level understanding of machine learning and statistics, and you program in Java, Python, or Scala, you’ll find the book’s patterns useful for working on your own data applications.

With this book, you will:

  • Familiarize yourself with the Spark programming model
  • Become comfortable within the Spark ecosystem
  • Learn general approaches in data science
  • Examine complete implementations that analyze large public data sets
  • Discover which machine learning tools make sense for particular problems
  • Acquire code that can be adapted to many uses

The first chapter will place Spark within the wider context of data science and big data analytics. After that, each chapter will comprise a self-contained analysis using Spark. The second chapter will introduce the basics of data processing in Spark and Scala through a use case in data cleansing. The next few chapters will delve into the meat and potatoes of machine learning with Spark, applying some of the most common algorithms in canonical applications. The remaining chapters are a bit more of a grab bag and apply Spark in slightly more exotic applications—for example, querying Wikipedia through latent semantic relationships in the text or analyzing genomics data.

Since the first edition, Spark has experienced a major version upgrade that instated an entirely new core API and sweeping changes in subcomponents like MLlib and Spark SQL. In the second edition, we’ve made major renovations to the example code and brought the materials up to date with Spark’s new best practices.

Sandy Ryza develops algorithms for public transit at Remix. Prior, he was a senior data scientist at Cloudera and Clover Health. He is an Apache Spark committer, Apache Hadoop PMC member, and founder of the Time Series for Spark project. He holds the Brown University computer science department’s 2012 Twining award for “Most Chill”.

Uri Laserson is an Assistant Professor of Genetics at the Icahn School of Medicine at Mount Sinai, where he develops scalable technology for genomics and immunology using the Hadoop ecosystem.
Sean Owen is Director of Data Science at Cloudera. He is an ApacheSpark committer and PMC member, and was an Apache Mahout committer.

Josh Wills is the Head of Data Engineering at Slack, the founder of the Apache Crunch project, and wrote a tweet about data scientists once.

What makes us different?

• Instant Download

• Always Competitive Pricing

• 100% Privacy

• FREE Sample Available

• 24-7 LIVE Customer Support

Reviews (0)

Reviews

There are no reviews yet.

Be the first to review “Advanced Analytics with Spark: Patterns for Learning from Data at Scale 2nd Edition, ISBN-13: 978-1491972953” Cancel reply

You must be logged in to post a review.

Shipping & Delivery

You will receive the link of your eBook 30 seconds after purchase on your email (check you email or junk mail), and you can login to your account at anytime using your username to read or download your eBook.

If you have any problem or any other questions, you can email us or try the chat widget.

Visit contact us.

Related products

-71%
Programming with Microsoft Visual Basic 2017 8th Edition, ISBN-13: 978-1337102124
Compare

Programming with Microsoft Visual Basic 2017 8th Edition, ISBN-13: 978-1337102124

Computing
$50.00 Original price was: $50.00.$14.50Current price is: $14.50.
Programming with Microsoft Visual Basic 2017 8th Edition by Diane Zak, ISBN-13: 978-1337102124 [PDF eBook eTextbook] 912 pages Publisher: Cengage
Add to wishlist
Add to cart
Quick view
-70%
Programming Logic and Design, Comprehensive by Joyce Farrell, ISBN-13: 978-1337102070
Compare

Programming Logic and Design, Comprehensive by Joyce Farrell, ISBN-13: 978-1337102070

Computing
$50.00 Original price was: $50.00.$14.99Current price is: $14.99.
Programming Logic and Design, Comprehensive by Joyce Farrell, ISBN-13: 978-1337102070 [PDF eBook eTextbook] Publisher: ‎ Cengage Learning; 9th edition (January
Add to wishlist
Add to cart
Quick view
-60%
SPSS Demystified 3rd Edition by Ronald Yockey, ISBN-13: 978-1138286283
Compare

SPSS Demystified 3rd Edition by Ronald Yockey, ISBN-13: 978-1138286283

Computing
$50.00 Original price was: $50.00.$19.99Current price is: $19.99.
SPSS Demystified 3rd Edition by Ronald Yockey, ISBN-13: 978-1138286283 [PDF eBook eTextbook] 276 pages Publisher: Routledge; 3 edition (August 22,
Add to wishlist
Add to cart
Quick view
-75%
Virtual Reality Designs 1st Edition Adriana Peña Pérez Negrón, ISBN-13: 978-0367894979
Compare

Virtual Reality Designs 1st Edition Adriana Peña Pérez Negrón, ISBN-13: 978-0367894979

Computing
$50.00 Original price was: $50.00.$12.34Current price is: $12.34.
Virtual Reality Designs 1st Edition by Adriana Peña Pérez Negrón, ISBN-13: 978-0367894979 [PDF eBook eTextbook] Publisher: ‎ CRC Press; 1st
Add to wishlist
Add to cart
Quick view
-64%
Python Crash Course 2nd Edition by Eric Matthes, ISBN-13: 978-1593279288
Compare

Python Crash Course 2nd Edition by Eric Matthes, ISBN-13: 978-1593279288

Computing
$50.00 Original price was: $50.00.$17.99Current price is: $17.99.
Python Crash Course: A Hands-On, Project-Based Introduction to Programming by Eric Matthes, ISBN-13: 978-1593279288 [PDF eBook eTextbook] Publisher: ‎ NO
Add to wishlist
Add to cart
Quick view
-61%
Programming Multicore and Many-core Computing Systems, ISBN-13: 978-0470936900
Compare

Programming Multicore and Many-core Computing Systems, ISBN-13: 978-0470936900

Computing
$50.00 Original price was: $50.00.$19.50Current price is: $19.50.
Programming Multicore and Many-core Computing Systems, ISBN-13: 978-0470936900 [PDF eBook eTextbook] Series: Wiley Series on Parallel and Distributed Computing (Book
Add to wishlist
Add to cart
Quick view
-80%
Python 3 for Machine Learning by Oswald Campesato, ISBN-13: 978-1683924951
Compare

Python 3 for Machine Learning by Oswald Campesato, ISBN-13: 978-1683924951

Computing
$50.00 Original price was: $50.00.$9.99Current price is: $9.99.
Python 3 for Machine Learning by Oswald Campesato, ISBN-13: 978-1683924951  [PDF eBook eTextbook] Publisher: ‎ Mercury Learning and Information (March
Add to wishlist
Add to cart
Quick view
-70%
Problem Solving with C++ 10th Edition by Walter Savitch, ISBN-13: 978-0134448282
Compare

Problem Solving with C++ 10th Edition by Walter Savitch, ISBN-13: 978-0134448282

Computing
$50.00 Original price was: $50.00.$14.99Current price is: $14.99.
Problem Solving with C++ 10th Edition by Walter Savitch, ISBN-13: 978-0134448282 [PDF eBook eTextbook] Publisher: ‎ Pearson; 10th edition (February
Add to wishlist
Add to cart
Quick view

Free Shipping.

Via Email.

24/7 Support.

Contact Or Chat With Us.

Online Payment.

One Time Payement.

Fast Delivery.

30 Seconds After Purchase.

  • OUR COMPANY
    • UniveBook
    • Email: contact@univebook.com
    • Website: univebook.com
  • USEFUL LINKS
    • Home
    • Shop
    • Wishlist
    • Blog
  • OUR POLICY
    • Privacy Policy
    • Refund Policy
    • Terms & Conditions
    • DMCA
  • INFORMATIONS
    • About Us
    • FAQ
    • Contact Us
    • Request an eBook

Payment System:

UNIVEBOOK 2020-2025 CREATED BY UniveBook . PREMIUM E-COMMERCE SOLUTIONS.
  • Home
  • Shop
  • Blog
  • About us
  • Contact us
  • Request an eBook
  • Wishlist
  • Compare
  • Login / Register
Shopping cart
Close
Sign in
Close

Lost your password?

No account yet?

Create an Account
Shop
Wishlist
5 items Cart
My account