S
Director, Cloud Infrastructure
Sanity
Contract type
Ongoing
Work mode
100% remote
Experience
Lead / principal · Not stated
Level inferred from job title
Job description
Key details
- Set the infrastructure strategy for Sanity's next stage of scale, with a clear roadmap across cloud infrastructure, reliability, deployment, observability, security, cost, and developer experience
- Lead the teams responsible for the shared foundations behind Sanity's engineering velocity: GCP projects and clusters, Kubernetes, networking, service discovery, routing, gateways, CI/CD, infrastructure as code, and observability
- Draw a clean line between Platform and SRE, with Platform owning shared foundations and SRE enabling product teams to run services well
- Raise the reliability bar across production systems, including dashboards, alert severity, paging standards, service ownership, on-call readiness, and incident response
- Make deployment boring in the best way: clear golden paths, production readiness checks, safe rollouts, useful automation, and fewer places engineers need to look before shipping
- Own cloud cost discipline without slowing the business down, making tradeoffs visible and helping teams build with cost in mind
- Shape longer-term architecture for multi-region scale, disaster recovery, data residency, and trust requirements like SOC 2, ISO 27001, PCI, HIPAA, or similar customer expectations
- Hire, coach, and stretch infrastructure leaders and engineers
- Company mission
- Information not specified
Primary stack
Core technologies
KubernetesGoogle Cloud Platform (GCP)AWSTerraform
Benefits
- Information not specified
Requirements & details
- Led Infrastructure, Platform, SRE, Cloud, or Developer Platform teams in a scaling SaaS, cloud, infrastructure, API, data, or developer-tools company
- Operated systems with high request volume, multi-region production, strict uptime expectations, large cloud bills, customer-facing incidents, and trust requirements
- Built or run production platforms with Kubernetes, GCP or AWS, Terraform or similar infrastructure as code, service discovery, networking, API gateways, CDNs, observability, CI/CD, and incident response
- Technically deep enough to debate architecture with senior infrastructure engineers and practical enough to make the call when the perfect answer is wasting time
- Knows how to split central platform ownership from product-team service ownership
- Kubernetes, GCP, AWS, Terraform, CI/CD, CDN, Observability, Networking, API Gateways
- Kubernetes
- Google Cloud Platform (GCP)
- AWS
- Terraform
