•Lead a platform-focused Service Enablement team owning cloud infrastructure, developer experience, and AI-native engineering workflows to ensure reliability and scalability.
•Key Responsibilities Guarantee platform reliability and 99.9%+ uptime across EKS clusters while optimizing costs.
•Define safety rails and test harnesses for autonomous AI-driven infrastructure workflows.
•Scale NestJS monorepo and CI/CD using NX, GitHub Actions, and Okteto.
•Run follow-the-sun on-call, incident reviews, and improve SLOs, runbooks, and observability.
•Requirements 10+ years in technology with SaaS/platform and distributed systems experience. 4+ years managing engineering teams with Senior/Staff engineers.
•Hands-on platform/SRE experience with Kubernetes/EKS and CI/CD pipelines.
•Experience improving developer experience and observability practices.
•Comfort with AI-assisted engineering and building safety/verification for autonomous agents.