Role Overview
This role focuses on cloud infrastructure, CI/CD automation, runtime observability, and security reliability within a platform engineering context.
Responsibilities
- Manage and maintain cloud infrastructure, demonstrating strong hands-on experience with GCP or another hyperscaler.
- Utilize Terraform for Infrastructure as Code.
- Manage cloud networking, IAM, and service accounts.
- Build and maintain CI/CD pipelines (e.g., GitHub Actions, Cloud Build, Jenkins).
- Automate operational tasks using scripting languages like Bash or Python.
- Implement release management and deployment strategies (e.g., blue/green, canary, rolling).
- Manage containerised platforms (Docker, Cloud Run, Kubernetes).
- Implement monitoring, logging, and alerting solutions.
- Apply cloud security fundamentals, including IAM and secrets management.
- Implement reliability and availability best practices.
Requirements
- 5–8+ years of experience in DevOps / Platform / Site Reliability Engineering roles.
- Strong hands-on experience with GCP or another hyperscaler.
- Proven experience with Infrastructure as Code (Terraform).
- Experience managing networking, IAM, and service accounts in cloud environments.
- Strong experience building and maintaining CI/CD pipelines (e.g., GitHub Actions, Cloud Build, Jenkins).
- Experience automating operational tasks using scripting languages (e.g., Bash, Python).
- Familiarity with release management and deployment strategies (blue/green, canary, rolling).
- Experience running containerised platforms (Docker, Cloud Run, Kubernetes).
- Hands-on experience with monitoring, logging, and alerting tools.
- Solid understanding of cloud security fundamentals (IAM, secrets management, least privilege).
- Experience working in distributed or offshore teams.