100 handpicked books like Advanced Analytics with Spark (picked by fans)

Designing Data-Intensive Applications

From my list on practical, hands-on books on DevOps and software delivery.

By Yevgeniy Brikman Author

Why did Yevgeniy love this book?

This is the best overview of data storage and distributed systems—two key concepts for building almost any piece of software today—that I've seen anywhere. Martin does a wonderful job of taking a massive body of research and distilling complicated concepts and difficult trade-offs down to a level anyone can understand.

I learned a lot about replication, partitioning, linearizability, locking, write skew, phantoms, transactions, event logs, and more. I'm also a big fan of the final chapter, The Future of Data Systems, which covers ideas such as "unbundling the database", end-to-end event streams, and an important discussion on ethics in programming and data systems.

Designing Data-Intensive Applications

By Martin Kleppmann,

Why should I read it?

2 authors picked Designing Data-Intensive Applications as one of their favorite books, and they share why you should read it.

What is this book about?

Data is at the center of many challenges in system design today. Difficult issues need to be figured out, such as scalability, consistency, reliability, efficiency, and maintainability. In addition, we have an overwhelming variety of tools, including NoSQL datastores, stream or batch processors, and message brokers. What are the right choices for your application? How do you make sense of all these buzzwords? In this practical and comprehensive guide, author Martin Kleppmann helps you navigate this diverse landscape by examining the pros and cons of various technologies for processing and storing data. Software keeps changing, but the fundamental principles remain…

Explore

Topics

Big data

Genres

Design

Hands-On Machine Learning with Scikit-Learn, Keras, and TensorFlow 3e

By Géron Aurélien,

From my list on big data processing ecosystem.

By Tomasz Lelek Author

Why did Tomasz love this book?

The Hands-on Machine Learning book presents an end-to-end approach to many problems that can be solved with machine learning.

Every concept and topic is backed up with a running code that you can experiment with and adapt to your real-world problems.

Thanks to this book, you will be able to understand the state of the art of today's machine learning and feel comfortable using the most up-to-date ML methods.

Hands-On Machine Learning with Scikit-Learn, Keras, and TensorFlow 3e

By Géron Aurélien,

Why should I read it?

1 author picked Hands-On Machine Learning with Scikit-Learn, Keras, and TensorFlow 3e as one of their favorite books, and they share why you should read it.

What is this book about?

Through a recent series of breakthroughs, deep learning has boosted the entire field of machine learning. Now, even programmers who know close to nothing about this technology can use simple, efficient tools to implement programs capable of learning from data. This best-selling book uses concrete examples, minimal theory, and production-ready Python frameworks--scikit-learn, Keras, and TensorFlow--to help you gain an intuitive understanding of the concepts and tools for building intelligent systems.

With this updated third edition, author Aurelien Geron explores a range of techniques, starting with simple linear regression and progressing to deep neural networks. Numerous code examples and exercises throughout…

Explore

Topics

Genres

Coming soon!

Kafka

By Neha Narkhede, Gwen Shapira, Todd Palino

From my list on big data processing ecosystem.

By Tomasz Lelek Author

Why did Tomasz love this book?

Apache Kafka is the backbone of almost every streaming-based system today.

The solutions created and implemented in Kafka are the key concepts in every streaming system that you will work with.

This book will allow you to fully understand the Kafka architecture, its internals, and APIs and allow you to become an expert in this technology.

Kafka

By Neha Narkhede, Gwen Shapira, Todd Palino

Why should I read it?

1 author picked Kafka as one of their favorite books, and they share why you should read it.

What is this book about?

Every enterprise application creates data, whether it's log messages, metrics, user activity, outgoing messages, or something else. And how to move all of this data becomes nearly as important as the data itself. If you're an application architect, developer, or production engineer new to Apache Kafka, this practical guide shows you how to use this open source streaming platform to handle real-time data feeds.

Engineers from Confluent and LinkedIn who are responsible for developing Kafka explain how to deploy production Kafka clusters, write reliable event-driven microservices, and build scalable stream-processing applications with this platform. Through detailed examples, you'll learn Kafka's…

Explore

Topics

Genres

Design

Database Internals

By Alex Petrov,

From my list on big data processing ecosystem.

By Tomasz Lelek Author

Why did Tomasz love this book?

The Database Internals will allow you to go one step further in your understanding of how distributed databases work.

The author has a lot of experience with one of the most successful distributed databases - Apache Cassandra and shares his knowledge about low-level details and internals of distributed databases.

Database Internals

By Alex Petrov,

Why should I read it?

1 author picked Database Internals as one of their favorite books, and they share why you should read it.

What is this book about?

When it comes to choosing, using, and maintaining a database, understanding its internals is essential. But with so many distributed databases and tools available today, it's often difficult to understand what each one offers and how they differ. With this practical guide, Alex Petrov guides developers through the concepts behind modern database and storage engine internals.

Throughout the book, you'll explore relevant material gleaned from numerous books, papers, blog posts, and the source code of several open source databases. These resources are listed at the end of parts one and two. You'll discover that the most significant distinctions among many…

Explore

Topics

Big data

Genres

Coming soon!

Be Data Literate

By Jordan Morrow,

From my list on for data science and analytics leaders.

By Jeremy Adamson Author

Why did Jeremy love this book?

Not everybody needs to be a data scientist, but everybody does need to be data literate. Without an intentional focus on evangelism and building a strong data culture in your organization it will be an uphill battle to make meaningful change. This book helps individuals and leaders to understand what data literacy is, and how we can build it like any other skill.

Be Data Literate

By Jordan Morrow,

Why should I read it?

1 author picked Be Data Literate as one of their favorite books, and they share why you should read it.

What is this book about?

In the fast moving world of the fourth industrial revolution not everyone needs to be a data scientist but everyone should be data literate, with the ability to read, analyze and communicate with data. It is not enough for a business to have the best data if those using it don't understand the right questions to ask or how to use the information generated to make decisions. Be Data Literate is the essential guide to developing the curiosity, creativity and critical thinking necessary to make anyone data literate, without retraining as a data scientist or statistician. With learnings to show…

Explore

Topics

Genres

Business

R for Data Science

By Hadley Wickham, Garrett Grolemund,

From my list on intro to programming and data science with R.

By Tilman M. Davies Author

Why did Tilman love this book?

For those intending to use R with an eye on the popular 'Tidyverse' suite of packages – which facilitate the handling, manipulation, and visualisation of data sets – it's hard to go past this book. From the founding contributors of the RStudio/Tidyverse worlds, this is a great way to learn about this dialect of R against the overarching backdrop of statistical data analysis and data science.

R for Data Science

By Hadley Wickham, Garrett Grolemund,

Why should I read it?

1 author picked R for Data Science as one of their favorite books, and they share why you should read it.

What is this book about?

Learn how to use R to turn raw data into insight, knowledge, and understanding. This book introduces you to R, RStudio, and the tidyverse, a collection of R packages designed to work together to make data science fast, fluent, and fun. Suitable for readers with no previous programming experience, R for Data Science is designed to get you doing data science as quickly as possible. Authors Hadley Wickham and Garrett Grolemund guide you through the steps of importing, wrangling, exploring, and modeling your data and communicating the results. You'll get a complete, big-picture understanding of the data science cycle, along…

Explore

Topics

Genres

Coming soon!

Introduction to Machine Learning with Python

By Andreas C. Müller, Sarah Guido,

From my list on machine learning for beginners.

By Yuxi (Hayden) Liu Author

Why did Yuxi love this book?

This book is more advanced than the first book I recommended. It presents ML theoretical and practical aspects step-by-step from the bottom up. Each chapter elaborates at length on a core building block in the ML life cycle. For example, feature engineering, supervised learning, and model evaluation have their own separate chapters, with intuitive discussions of how they work. Most of the concept is taught through the simple yet powerful Python Module Scikit-Learn so it won’t overburden you with heavy programming. This book will be perfect for practitioners with some understanding of statistics and linear algebra.

Introduction to Machine Learning with Python

By Andreas C. Müller, Sarah Guido,

Why should I read it?

1 author picked Introduction to Machine Learning with Python as one of their favorite books, and they share why you should read it.

What is this book about?

Machine learning has become an integral part of many commercial applications and research projects, but this field is not exclusive to large companies with extensive research teams. If you use Python, even as a beginner, this book will teach you practical ways to build your own machine learning solutions. With all the data available today, machine learning applications are limited only by your imagination. You'll learn the steps necessary to create a successful machine-learning application with Python and the scikit-learn library. Authors Andreas Muller and Sarah Guido focus on the practical aspects of using machine learning algorithms, rather than the…

Explore

Topics

Genres

Coming soon!

Machine Learning For Absolute Beginners

By Oliver Theobald,

From my list on machine learning for beginners.

By Yuxi (Hayden) Liu Author

Why did Yuxi love this book?

This could be the first stop of your brand new machine learning journey. I personally like how the technical concept is translated into plain English – each chapter starts with a high-level overview of a ML algorithm or methodology, concise and clear, followed by lots of visual examples and real world scenarios. I can guarantee you won’t get lost halfway. The book focuses on getting you introduced to ML with minimal math. But if you want to grasp some more of math, the next book I recommend is waiting for you.

Machine Learning For Absolute Beginners

By Oliver Theobald,

Why should I read it?

1 author picked Machine Learning For Absolute Beginners as one of their favorite books, and they share why you should read it.

What is this book about?

NOTICE: To buy the newest edition of this book (2021), please search "Machine Learning Absolute Beginners Third Edition" on Amazon. The product page you are currently viewing is for the 2nd Edition (2017) of this book.

Featured by Tableau as the first of "7 Books About Machine Learning for Beginners."

Ready to spin up a virtual GPU instance and smash through petabytes of data? Want to add 'Machine Learning' to your LinkedIn profile?

Well, hold on there...

Before you embark on your epic journey, there are some high-level theory and statistical principles to weave through first.
But rather than spend…

Explore

Topics

Genres

Coming soon!

Rage Inside the Machine

By Robert Elliott Smith,

From my list on no hype and no nonsense artificial intelligence.

By Peter J. Bentley Author

Why did Peter love this book?

OK, I’m biased here because Rob is an old friend of mine. We first met at academic conferences and had several heated debates (arguments). But after spending a little time together at a workshop we realised each probably knew what they were talking about after all. Robert Elliott Smith, I should make clear it's not the Rob Smith who writes about “Artificial Superintelligence”. Those books definitely do not make this list.

Our Rob is a coherent, grounded scientist with bags of real-world experience, and he brings his knowledge to this title with gusto, telling us about how AI is affecting our lives in ways you never thought possible – and often not in a good way. If you want to understand what can go wrong with AI and what we should be doing to stop it, don’t read about singularities or other such nonsense, read this.

Rage Inside the Machine

By Robert Elliott Smith,

Why should I read it?

1 author picked Rage Inside the Machine as one of their favorite books, and they share why you should read it.

What is this book about?

Shortlisted for the 2020 Business Book Awards

We live in a world increasingly ruled by technology; we seem as governed by technology as we do by laws and regulations. Frighteningly often, the influence of technology in and on our lives goes completely unchallenged by citizens and governments. We comfort ourselves with the soothing refrain that technology has no morals and can display no prejudice, and it's only the users of technology who distort certain aspects of it.

But is this statement actually true? Dr Robert Smith thinks it is dangerously untrue in the modern era.

Having worked in the field…

Explore

Topics

Genres

Business

Fundamentals of Machine Learning for Predictive Data Analytics, Second Edition

By John D. Kelleher, Brian Mac Namee, Aoife D'Arcy

From my list on machine learning for beginners.

By Yuxi (Hayden) Liu Author

Why did Yuxi love this book?

Another practical book that I highly recommend. Its intuitive structure is the first thing I like about it. It gives you a comprehensive walkthrough of the ML workflow, from data exploration to learning. It covers abundant practical guides that get you prepared for real world challenges, such as how to handle outliers and to impute missing data. As a ML practitioner, I appreciate the dedicated case studies throughout the entire book. They really excite learners for future real world applications.

Fundamentals of Machine Learning for Predictive Data Analytics, Second Edition

By John D. Kelleher, Brian Mac Namee, Aoife D'Arcy

Why should I read it?

1 author picked Fundamentals of Machine Learning for Predictive Data Analytics, Second Edition as one of their favorite books, and they share why you should read it.

What is this book about?

The second edition of a comprehensive introduction to machine learning approaches used in predictive data analytics, covering both theory and practice.

Machine learning is often used to build predictive models by extracting patterns from large datasets. These models are used in predictive data analytics applications including price prediction, risk assessment, predicting customer behavior, and document classification. This introductory textbook offers a detailed and focused treatment of the most important machine learning approaches used in predictive data analytics, covering both theoretical concepts and practical applications. Technical and mathematical material is augmented with explanatory worked examples, and case studies illustrate the application…

Explore

Topics

Genres

Coming soon!

56 books like Advanced Analytics with Spark

Designing Data-Intensive Applications

Why did Yevgeniy love this book?

Hands-On Machine Learning with Scikit-Learn, Keras, and TensorFlow 3e

Why did Tomasz love this book?

Kafka

Why did Tomasz love this book?

Database Internals

Why did Tomasz love this book?

Be Data Literate

Why did Jeremy love this book?

R for Data Science

Why did Tilman love this book?

Introduction to Machine Learning with Python

Why did Yuxi love this book?

Machine Learning For Absolute Beginners

Why did Yuxi love this book?

Rage Inside the Machine

Why did Peter love this book?

Fundamentals of Machine Learning for Predictive Data Analytics, Second Edition

Why did Yuxi love this book?

5 book lists we think you will like!

Interested in data mining, big data, and machine learning?

56 books like Advanced Analytics with Spark

Why am I passionate about this?

Yevgeniy's book list on practical, hands-on books on DevOps and software delivery

Why did Yevgeniy love this book?

Why am I passionate about this?

Tomasz's book list on big data processing ecosystem

Why did Tomasz love this book?

Why am I passionate about this?

Tomasz's book list on big data processing ecosystem

Why did Tomasz love this book?

Why am I passionate about this?

Tomasz's book list on big data processing ecosystem

Why did Tomasz love this book?

Why am I passionate about this?

Jeremy's book list on for data science and analytics leaders

Why did Jeremy love this book?

Why am I passionate about this?

Tilman's book list on intro to programming and data science with R

Why did Tilman love this book?

Why am I passionate about this?

Yuxi's book list on machine learning for beginners

Why did Yuxi love this book?

Why am I passionate about this?

Yuxi's book list on machine learning for beginners

Why did Yuxi love this book?

Why am I passionate about this?

Peter's book list on no hype and no nonsense artificial intelligence

Why did Peter love this book?

Why am I passionate about this?

Yuxi's book list on machine learning for beginners

Why did Yuxi love this book?

Share your top 3 reads of 2024!

1,176

5 book lists we think you will like!

Interested in data mining, big data, and machine learning?