Location: Vancouver, BC
Contract: 12 months
Schedule: 40 hours/week, 5 days onsite
Start Date: ASAP
We are seeking a Cloud DevOps Engineer to join the Health Team. This role will focus on provisioning and maintaining Azure infrastructure, supporting Kubernetes-based applications, and building reliable CI/CD and GitOps deployment processes.
What You’ll Do
- Provision and maintain Azure infrastructure using Terraform, including compute, networking, storage, and services supporting application and data workloads.
- Manage and fine-tune Azure infrastructure to improve reliability, performance, resource utilization, and cost efficiency.
- Deploy, configure, and maintain applications on Kubernetes using Helm and Argo CD.
- Create and maintain Helm charts, Kubernetes manifests, and GitOps deployment configurations across environments.
- Troubleshoot deployment failures and application issues involving pods, services, ingress, configuration, resource limits, and connectivity.
- Work closely with developers and data engineers to support releases and troubleshoot application and data engineering workloads, including ETL pipelines.
- Build and maintain container images and manage their publication to container registries.
- Create and maintain CI/CD pipelines to make application releases reliable and repeatable.
- Monitor application and infrastructure health, investigate alerts, and resolve operational issues.
- Follow established security practices for access, secrets, and infrastructure and deployment changes.
- Maintain clear runbooks and operational documentation.
What You’ll Bring
- Experience in DevOps or cloud operations.
- Hands-on experience with Azure Kubernetes Service (AKS) or another managed Kubernetes platform such as EKS/GKE, or self-hosted Kubernetes.
- Hands-on experience deploying and troubleshooting applications on Kubernetes.
- Experience using Helm to configure and deploy applications, including maintaining charts and environment-specific values.
- Hands-on experience using Terraform to provision and maintain cloud infrastructure, including reusable configuration, state management, and reviewing and applying infrastructure changes.
- Hands-on experience managing Azure compute, networking, storage, access, and monitoring.
- Hands-on Linux administration and network troubleshooting skills.
- Experience with Git and at least one CI/CD platform, such as GitHub Actions.
- Ability to correlate logs, metrics, and events to diagnose and resolve issues.
- Clear communication, ownership of assigned work, and effective collaboration with other engineers.
Nice to Have
- Experience supporting data platforms, data pipelines, or Azure services used for data processing and storage.
- Experience supporting AI/ML pipelines.
- Familiarity with relational and NoSQL databases.
- Experience with Kubernetes monitoring tools such as Prometheus and Grafana.