Kukreja, Manoj

Data Engineering with - Apache Spark, Delta Lake, - and Lakehouse: Create - scalable pipelines that - ingest, curate, and - aggregate complex data in - a timely and secure way

(No reviews yet) Write a Review
ISBN 13:
9781801077743
author:
Kukreja, Manoj
format:
Paperback
publisher:
Packt Publishing Limited
language:
English
Publication Year:
2021
Pages:
480
Dimensions:
23.5 x 19.1 x 2.5 centimetres (0
Genre:
Computers, Computer Science, Data Modeling,
Condition:
New
Availability:
Item usually sent within 7 working days
£44.74

Description

Build Scalable Data Pipelines and Networks with Apache Spark, Delta Lake, and Lakehouse Manoj Kukreja's comprehensive guide to data engineering helps you create reliable data platforms for managers, data scientists, and analysts alike. Learn how to build data pipelines that can adapt to changing schemas and data volumes. This book covers the core concepts of Apache Spark and Delta Lake, as well as ingesting, processing, and analyzing complex data for machine learning model training. You'll also discover how to operationalize data models in production using curated data and deploy data lakes with fast performance and governance in mind. With practical examples and code snippets, Manoj Kukreja shares real-world scenarios from his 10 years of experience working with big data, providing a solid foundation for building scalable data pipelines and networks.

View AllClose