Role Overview
As a GCP Data Engineer, you will design, develop, and optimize ETL/ELT pipelines using GCP-native tools. You will be responsible for building and managing data warehousing solutions, developing ingestion frameworks, and ensuring data quality and security across all pipelines.
Responsibilities
- Design, develop, and optimize ETL/ELT pipelines using GCP-native tools (Dataflow, Dataproc, Cloud Composer/Airflow)
- Build and manage data warehousing solutions on BigQuery, including schema design, partitioning, and query optimization
- Develop data ingestion frameworks from structured and unstructured sources (batch and streaming)
- Implement data quality checks, monitoring, and alerting across pipelines
- Collaborate with data architects to translate business requirements into technical designs
- Work with Pub/Sub, Cloud Storage, and Cloud Functions to build event-driven data workflows
- Support CI/CD practices for data pipeline deployments (Cloud Build, Terraform)
- Ensure data security, access controls, and compliance with governance standards
- Participate in code reviews, sprint planning, and technical documentation
- Coordinate with onshore teams across time zones for delivery alignment
Requirements
- Proven experience in designing and implementing data pipelines on GCP
- Strong expertise in BigQuery and SQL
- Experience with orchestration tools like Airflow
- Familiarity with Infrastructure as Code (Terraform) and CI/CD
Skills
- GCP
- BigQuery
- Apache Airflow
- Terraform
- Python