What do you get with Print?

Instant access to your digital copy whilst your Print order is Shipped

Paperback book shipped to your preferred address

Redeem a companion digital copy on all Print orders

Access this title in our online reader with advanced features

DRM FREE - Read whenever, wherever and however you want

AI Assistant (beta) to help accelerate your learning

Key benefits

Understand data engineering concepts, the role of a data engineer, and the benefits of using GCP for building your solution

Learn how to use the various GCP products to ingest, consume, and transform data and orchestrate pipelines

Discover tips to prepare for and pass the Professional Data Engineer exam

Description

With this book, you'll understand how the highly scalable Google Cloud Platform (GCP) enables data engineers to create end-to-end data pipelines right from storing and processing data and workflow orchestration to presenting data through visualization dashboards. Starting with a quick overview of the fundamental concepts of data engineering, you'll learn the various responsibilities of a data engineer and how GCP plays a vital role in fulfilling those responsibilities. As you progress through the chapters, you'll be able to leverage GCP products to build a sample data warehouse using Cloud Storage and BigQuery and a data lake using Dataproc. The book gradually takes you through operations such as data ingestion, data cleansing, transformation, and integrating data with other sources. You'll learn how to design IAM for data governance, deploy ML pipelines with the Vertex AI, leverage pre-built GCP models as a service, and visualize data with Google Data Studio to build compelling reports. Finally, you'll find tips on how to boost your career as a data engineer, take the Professional Data Engineer certification exam, and get ready to become an expert in data engineering with GCP. By the end of this data engineering book, you'll have developed the skills to perform core data engineering tasks and build efficient ETL data pipelines with GCP.

Who is this book for?

This book is for data engineers, data analysts, and anyone looking to design and manage data processing pipelines using GCP. You'll find this book useful if you are preparing to take Google's Professional Data Engineer exam. Beginner-level understanding of data science, the Python programming language, and Linux commands is necessary. A basic understanding of data processing and cloud computing, in general, will help you make the most out of this book.

What you will learn

Load data into BigQuery and materialize its output for downstream consumption

Build data pipeline orchestration using Cloud Composer

Develop Airflow jobs to orchestrate and automate a data warehouse

Build a Hadoop data lake, create ephemeral clusters, and run jobs on the Dataproc cluster

Leverage Pub/Sub for messaging and ingestion for event-driven systems

Use Dataflow to perform ETL on streaming data

Unlock the power of your data with Data Studio

Calculate the GCP cost estimation for your end-to-end data solutions

What do you get with Print?

Instant access to your digital copy whilst your Print order is Shipped

Paperback book shipped to your preferred address

Redeem a companion digital copy on all Print orders

Access this title in our online reader with advanced features

DRM FREE - Read whenever, wherever and however you want

AI Assistant (beta) to help accelerate your learning

Frequently bought together

Google Cloud Certified Professional Cloud Network Engineer Guide

NZ$71.99

NZ$94.99

Data Engineering with Google Cloud Platform

NZ$102.99

Total NZ$ 269.97

FAQs

What is the digital copy I get with my Print order?

When you buy any Print edition of our Books, you can redeem (for free) the eBook edition of the Print Book you’ve purchased. This gives you instant access to your book when you make an order via PDF, EPUB or our online Reader experience.

What is the delivery time and cost of print book?

Shipping Details

USA:

Economy: Delivery to most addresses in the US within 10-15 business days

Premium: Trackable Delivery to most addresses in the US within 3-8 business days

UK:

Economy: Delivery to most addresses in the U.K. within 7-9 business days.
Shipments are not trackable

Premium: Trackable delivery to most addresses in the U.K. within 3-4 business days!
Add one extra business day for deliveries to Northern Ireland and Scottish Highlands and islands

EU:

Premium: Trackable delivery to most EU destinations within 4-9 business days.

Australia:

Economy: Can deliver to P. O. Boxes and private residences.
Trackable service with delivery to addresses in Australia only.
Delivery time ranges from 7-9 business days for VIC and 8-10 business days for Interstate metro
Delivery time is up to 15 business days for remote areas of WA, NT & QLD.

Premium: Delivery to addresses in Australia only
Trackable delivery to most P. O. Boxes and private residences in Australia within 4-5 days based on the distance to a destination following dispatch.

India:

Premium: Delivery to most Indian addresses within 5-6 business days

Rest of the World:

Premium: Countries in the American continent: Trackable delivery to most countries within 4-7 business days

Asia:

Premium: Delivery to most Asian addresses within 5-9 business days

Disclaimer:
All orders received before 5 PM U.K time would start printing from the next business day. So the estimated delivery times start from the next day as well. Orders received after 5 PM U.K time (in our internal systems) on a business day or anytime on the weekend will begin printing the second to next business day. For example, an order placed at 11 AM today will begin printing tomorrow, whereas an order placed at 9 PM tonight will begin printing the day after tomorrow.

Unfortunately, due to several restrictions, we are unable to ship to the following countries:

Afghanistan
American Samoa
Belarus
Brunei Darussalam
Central African Republic
The Democratic Republic of Congo
Eritrea
Guinea-bissau
Iran
Lebanon
Libiya Arab Jamahriya
Somalia
Sudan
Russian Federation
Syrian Arab Republic
Ukraine
Venezuela

What is custom duty/charge?

Customs duty are charges levied on goods when they cross international borders. It is a tax that is imposed on imported goods. These duties are charged by special authorities and bodies created by local governments and are meant to protect local industries, economies, and businesses.

Do I have to pay customs charges for the print book order?

The orders shipped to the countries that are listed under EU27 will not bear custom charges. They are paid by Packt as part of the order.

List of EU27 countries: www.gov.uk/eu-eea:

A custom duty or localized taxes may be applicable on the shipment and would be charged by the recipient country outside of the EU27 which should be paid by the customer and these duties are not included in the shipping charges been charged on the order.

How do I know my custom duty charges?

The amount of duty payable varies greatly depending on the imported goods, the country of origin and several other factors like the total invoice amount or dimensions like weight, and other such criteria applicable in your country.

For example:

If you live in Mexico, and the declared value of your ordered items is over $ 50, for you to receive a package, you will have to pay additional import tax of 19% which will be $ 9.50 to the courier service.
Whereas if you live in Turkey, and the declared value of your ordered items is over € 22, for you to receive a package, you will have to pay additional import tax of 18% which will be € 3.96 to the courier service.

How can I cancel my order?

Cancellation Policy for Published Printed Books:

You can cancel any order within 1 hour of placing the order. Simply contact customercare@packt.com with your order details or payment transaction id. If your order has already started the shipment process, we will do our best to stop it. However, if it is already on the way to you then when you receive it, you can contact us at customercare@packt.com using the returns and refund process.

Please understand that Packt Publishing cannot provide refunds or cancel any order except for the cases described in our Return Policy (i.e. Packt Publishing agrees to replace your printed book because it arrives damaged or material defect in book), Packt Publishing will not accept returns.

What is your returns and refunds policy?

Return Policy:

We want you to be happy with your purchase from Packtpub.com. We will not hassle you with returning print books to us. If the print book you receive from us is incorrect, damaged, doesn't work or is unacceptably late, please contact Customer Relations Team on customercare@packt.com with the order number and issue details as explained below:

If you ordered (eBook, Video or Print Book) incorrectly or accidentally, please contact Customer Relations Team on customercare@packt.com within one hour of placing the order and we will replace/refund you the item cost.
Sadly, if your eBook or Video file is faulty or a fault occurs during the eBook or Video being made available to you, i.e. during download then you should contact Customer Relations Team within 14 days of purchase on customercare@packt.com who will be able to resolve this issue for you.
You will have a choice of replacement or refund of the problem items.(damaged, defective or incorrect)
Once Customer Care Team confirms that you will be refunded, you should receive the refund within 10 to 12 working days.
If you are only requesting a refund of one book from a multiple order, then we will refund you the appropriate single item.
Where the items were shipped under a free shipping offer, there will be no shipping costs to refund.

On the off chance your printed book arrives damaged, with book material defect, contact our Customer Relation Team on customercare@packt.com within 14 days of receipt of the book with appropriate evidence of damage and we will work with you to secure a replacement copy, if necessary. Please note that each printed book you order from us is individually made by Packt's professional book-printing partner which is on a print-on-demand basis.

What tax is charged?

Currently, no tax is charged on the purchase of any print book (subject to change based on the laws and regulations). A localized VAT fee is charged only to our European and UK customers on eBooks, Video and subscriptions that they buy. GST is charged to Indian customers for eBooks and video purchases.

What payment methods can I use?

You can pay with the following card types:

Visa Debit
Visa Credit
MasterCard
PayPal

What is the delivery time and cost of print books?

Shipping Details

USA:

Economy: Delivery to most addresses in the US within 10-15 business days

Premium: Trackable Delivery to most addresses in the US within 3-8 business days

UK:

Economy: Delivery to most addresses in the U.K. within 7-9 business days.
Shipments are not trackable

Premium: Trackable delivery to most addresses in the U.K. within 3-4 business days!
Add one extra business day for deliveries to Northern Ireland and Scottish Highlands and islands

EU:

Premium: Trackable delivery to most EU destinations within 4-9 business days.

Australia:

India:

Premium: Delivery to most Indian addresses within 5-6 business days

Rest of the World:

Premium: Countries in the American continent: Trackable delivery to most countries within 4-7 business days

Asia:

Premium: Delivery to most Asian addresses within 5-9 business days

Unfortunately, due to several restrictions, we are unable to ship to the following countries:

Afghanistan
American Samoa
Belarus
Brunei Darussalam
Central African Republic
The Democratic Republic of Congo
Eritrea
Guinea-bissau
Iran
Lebanon
Libiya Arab Jamahriya
Somalia
Sudan
Russian Federation
Syrian Arab Republic
Ukraine
Venezuela

Filter reviews by

All

Packt verified reviews

Amazon verified reviews

Kyle Malone May 29, 2022

As a data analyst looking to expand my data engineering skills, I have been looking for a book on this topic for a while now. It's exactly what I needed to better understand how to write ETL jobs using GCP.The code in Github has a few errors, but if you know basic Python you should be able to spot and correct the mistakes fairly easily. Overall, this books has helped me a ton. I would definitely purchase it again.

Amazon Verified review

Behnam Jun 19, 2022

This book helped me to get started working with GCP in an organized way, now I have the confidence to do some real projects on GCP.

cloud-learner Jun 28, 2024

well written and easy to understand!

Om S May 31, 2022

Google Cloud Platform had always been pioneered when it comes to big data and AI. This platform is amazingly simple for developers and the author has done a great job putting everything together step by step to facilitate readers' and users' experience.The book is divided into 12 solid chapters; the last chapter is related to those who are interested in getting certification and want to know more about it.[Chapter 1 ] - Explains fundamentals of data engineering, ETL concept in data engineering and data life cycle, the difference between ETL and ELT, etc. It also explains ETL, big data, and distributed systems in short.[Chapter 2] - This is more related to big data products and their capability, GCP console, cloud shell, and cloud editor are needed and explained in detail with other associated services. The author has explained serverless services or fully managed services. Most of the cloud companies want clients/users to be on and opt serverless side of the story![Chapter 3] - Cover Google Cloud Storage and BigQuery console, Different kind of scenarios the author has presented give a different aspect of the data modeling and understanding of the BigQuery operations. You will learn design data modeling for BigQuery. Diagram and code to cover your practice that is your key.[Chapter 4] - In this chapter automating tasks, jobs, and how to handle their dependencies has been explained. Cloud Composer, the introduction of the open-source tool “Airflow” and data pipeline for BigQuery data warehouse. [Chapter 5] - If you know and have an idea about Spark and PySpark you will be enjoying this chapter! Developing Spark ETL from GCS to BigQuery is a fun part. Brief Intro to Dataproc which is Data Lake, building a data lake on a Dataproc cluster, creating and running jobs on a Dataproc cluster, the concept of the ephemeral cluster, using Dataproc and Cloud Composer is explained. My suggestion is to learn HDFS and Spark a bit in detail, this is an important chapter.[Chapter 6] – You will learn about streaming data and how to handle incoming data as soon as data is created using the Pub/Sub publisher client.[Chapter 7] - Data Studio for visualizing data with connectivity to BigQuery. Data studio explorer gives a lot of options to make charts/aggregations and visualize your data.[Chapter 8] - No one understands better than Google Cloud about Machine Learning and Artificial Intelligence. In this chapter, author has a provided bird eye view of ML, MLOps, some pre-built GCP models as a service, and deploying ML pipelines with Vertex AI are highlights. If you like ML you will be enjoying this chapter. Different Exercises make this a more fun chapter. [Chapter 9 ] IAM, project structure, and BigQuery ACLs, controlling user access/ infrastructure as code. You will also learn the power of Terraform but my intake is practicing “Terraform” a bit outside of this book is more fun and you will see new challenges. [Chapter 10] - This chapter cover cost strategy/saving money and also highlights how to estimate the overall data solution using GCP. End-to-end data solution cost with tools and tips for optimizing services. Happy clients bring good business![Chapter 11] - Continuous integration and continuous deployment (CI/CD) on Google Cloud Platform for Data Engineers, explains the concept of CI/CD and its relevance to data engineers. Without CI/CD cloud is like a handicap! If you have used that before you will understand its importance.[Chapter 12 ] – As I explained in the beginning it's all about Boosting Your Confidence as a Data Engineer, and preparing you for the GCP certification. Finally, I would conclude it as a great book. Nice references/URL and summary for revision.For a new version of the book, I would be expecting more questions and case studies for readers.

Yashar Mansouri Jun 26, 2022

Data Engineering with GCP by Adi Wijaya is a hands-on-book that covers processes such as Data Warehousing , ETL/ELT, and workflow management through the use of Google Cloud Platform's tech stack.Services such as IAM & Admin, BigQuery, CloudSQL, Cloud Storage, Composer, Dataproc, Pub/Sub, and Dataflow are covered in multiple chapters of the book with hands-on or coding examples through Cloud Console, Cloud Shell scripting, and even Python code. Most examples are presented by trying to answer business requirements for real world scenarios.What I specifically liked is that the author also covers the concepts around data engineering processes such as comparing the Inmon and Kimball Data Warehousing methods and when to use each as well as user and project management and the optimal cost strategy when using the tools provided by GCP.

Data Engineering with Google Cloud Platform: A practical guide to operationalizing scalable data analytics systems on GCP

What do you get with Print?