Archie provides smart workplaces & coworking spaces with an all-in-one software in order to manage offices and enable employees to work from anywhere.
You can find us here: https://archieapp.co
ROLE
We are hiring a talented Senior Devops who will assist the CTO in helping scale the architecture. The role will also take responsibility for different certification requirements (SOC2 - ISO27001)
You will be working on a product that will shape how people work in the future, and you will be working with cutting-edge technologies and work in team that moves fast.
RESPONSIBILITIES
- Design and scale a reliable, highly available, and secure GCP infrastructure, including redundancy, backups, disaster recovery, and multi-zone architecture.
- Own and improve our Kubernetes, Terraform, Ansible, and ArgoCD infrastructure and deployment processes.
- Streamline CI/CD pipelines, builds, and deployments to make releases faster, safer, and more reliable.
- Improve observability across the platform using Grafana and related monitoring, logging, alerting, and event-management tools.
- Manage and improve Cloudflare, Zero Trust, networking, IAM, and MDM.
- Work with the engineering team on SOC 2 and ISO 27001 compliance, security controls, audits, and infrastructure documentation.
- Automate infrastructure and operational tasks, including developing Golang scripts and tools where appropriate.
- Proactively identify and address scalability, reliability, security, and operational risks.
- Establish and maintain clear documentation, runbooks, backup procedures, and disaster recovery processes.
- Work closely with developers to ensure our applications and infrastructure are scalable, resilient, and easy to operate.
SKILLS & QUALIFICATIONS
- 5+ years of experience in DevOps, SRE, Cloud Infrastructure, or a similar role.
- Strong experience with GCP, Kubernetes, Terraform, and ArgoCD.
- Strong understanding of CI/CD pipelines, build systems, and deployment automation.
- Experience with observability, monitoring, logging, alerting, and event management, ideally with Grafana, Prometheus, and Sentry.
- Strong knowledge of cloud networking, Cloudflare, Zero Trust, IAM, and security best practices.
- Experience designing high-availability, redundant, and disaster-recovery architectures.
- Experience with backup strategies, data recovery, and business continuity.
- Experience supporting SOC 2 and/or ISO 27001 compliance programs.
- Experience with MDM and endpoint security.
- Proficiency in Golang and scripting/automation.
- Experience with Infrastructure as Code and infrastructure automation.
- Strong troubleshooting and incident-response skills.
- Excellent organizational, documentation, and communication skills.
- Proactive mindset with strong ownership and attention to detail.
- Ability to work independently while collaborating closely with software engineering and security teams.