All jobs
Save

Director, Cloud Infrastructure

Sanity
Contract type
Ongoing
Work mode
100% remote
Experience
Lead / principal · Not stated

Level inferred from job title

Job description

Key details

  • Set the infrastructure strategy for Sanity's next stage of scale, with a clear roadmap across cloud infrastructure, reliability, deployment, observability, security, cost, and developer experience
  • Lead the teams responsible for the shared foundations behind Sanity's engineering velocity: GCP projects and clusters, Kubernetes, networking, service discovery, routing, gateways, CI/CD, infrastructure as code, and observability
  • Draw a clean line between Platform and SRE, with Platform owning shared foundations and SRE enabling product teams to run services well
  • Raise the reliability bar across production systems, including dashboards, alert severity, paging standards, service ownership, on-call readiness, and incident response
  • Make deployment boring in the best way: clear golden paths, production readiness checks, safe rollouts, useful automation, and fewer places engineers need to look before shipping
  • Own cloud cost discipline without slowing the business down, making tradeoffs visible and helping teams build with cost in mind
  • Shape longer-term architecture for multi-region scale, disaster recovery, data residency, and trust requirements like SOC 2, ISO 27001, PCI, HIPAA, or similar customer expectations
  • Hire, coach, and stretch infrastructure leaders and engineers
  • Company mission
  • Information not specified

Primary stack

Core technologies

KubernetesGoogle Cloud Platform (GCP)AWSTerraform

Benefits

  • Information not specified

Requirements & details

  • Led Infrastructure, Platform, SRE, Cloud, or Developer Platform teams in a scaling SaaS, cloud, infrastructure, API, data, or developer-tools company
  • Operated systems with high request volume, multi-region production, strict uptime expectations, large cloud bills, customer-facing incidents, and trust requirements
  • Built or run production platforms with Kubernetes, GCP or AWS, Terraform or similar infrastructure as code, service discovery, networking, API gateways, CDNs, observability, CI/CD, and incident response
  • Technically deep enough to debate architecture with senior infrastructure engineers and practical enough to make the call when the perfect answer is wasting time
  • Knows how to split central platform ownership from product-team service ownership
  • Kubernetes, GCP, AWS, Terraform, CI/CD, CDN, Observability, Networking, API Gateways
  • Kubernetes
  • Google Cloud Platform (GCP)
  • AWS
  • Terraform

Apply