Agile Data Science 2.0: Building Full-Stack Data Analytics Applications with Spark
Book information
Description
Data science teams looking to turn research into useful analytics applications require not only the right tools, but also the right approach if they’re to succeed. With the revised second edition of this hands-on guide, up-and-coming data scientists will learn how to use the Agile Data Science development methodology to build data applications with Python, Apache Spark, Kafka, and other tools. Author Russell Jurney demonstrates how to compose a data platform for building, deploying, and refining analytics applications with Apache Kafka, MongoDB, ElasticSearch, d3.js, scikit-learn, and Apache Airflow. You’ll learn an iterative approach that lets you quickly change the kind of analysis you’re doing, depending on what the data is telling you. Publish data science work as a web application, and affect meaningful change in your organization. Build value from your data in a series of agile sprints, using the data-value pyramidExtract features for statistical models from a single datasetVisualize data with charts, and expose different aspects through interactive reportsUse historical data to predict the future via classification and regressionTranslate predictions into actionsGet feedback from users after each sprint to keep your project on track
Similar books
Statistics for Machine Learning: Techniques for exploring supervised, unsupervised, and reinforcement learning models with Python and R
2017 · PDF
The Evolution of Data Products
2011 · PDF
Craft GraphQL APIs in Elixir with Absinthe: Flexible, Robust Services for Queries, Mutations, and Subscriptions
2018 · PDF
Geometric Transformations for 3D Modeling
2007 · RAR
Seven Databases in Seven Weeks: A Guide to Modern Databases and the NoSQL Movement
2018 · PDF
Supercharge Excel: When you learn to Write DAX for Power Pivot
2018 · PDF
Pro Entity Framework Core 2 for ASP.NET Core MVC
2018 · EPUB
Reservoir Modelling: A Practical Guide
2018 · PDF