Member of Technical Staff - Site Reliability
RunlayerEnterprise AI company
New York, United StatesLead
Khosla Ventures
Felicis
NorthBridge PE
Software Engineering
About the role
TL;DR
Own reliability and performance of Runlayer's cloud infrastructure.
- •As our Site Reliability Engineer, you'll own the reliability, performance, and scalability of Runlayer's infrastructure as we grow to serve enterprise customers across cloud and on-prem environments.
- •Key Responsibilities Own reliability and performance of our cloud infrastructure across AWS (ECS, Aurora, CloudWatch) and GCP Manage and optimize Kubernetes clusters and container orchestration Drive database reliability engineering, including performance tuning and scaling Build and maintain CI/CD pipelines for rapid, safe deployments Run incident response and on-call rotations Requirements Strong AWS experience, particularly ECS, Aurora, and CloudWatch GCP experience as we expand cross-cloud Kubernetes and container orchestration expertise
Required skills
AWSECSGoogle CloudKubernetesCI/CDPython
Nice-to-have skills
PythonKubernetesCI/CD
Benefits & perks
Competitive salary and equity, Paid time off, Professional development, Top-tier equipment, Health benefits, Customer interaction opportunities
Tech stack
AWSECSGoogle CloudKubernetesCI/CDPython