Senior DevOps / Site Reliability Engineer (SRE)
LendbuzzAuto Lending company
Tel AvivSenior
83North
Wellington Management
Group 1001
O.G. Venture Partners
MUFG Innovation Partners
Viola Credit
Software Engineering
About the role
TL;DR
Optimize production environments and ensure system stability and scalability.
- •Join Lendbuzz to enhance our production environment's reliability, availability, and performance.
- •Key Responsibilities Own the reliability, availability, and performance of our production environment.
- •Lead high quality effort and long term solutions that will improve stability and resiliency of our production environments.
- •Work closely with dev teams to review architecture, influence system design, and ensure services are built to be observable and scalable.
- •Handle capacity planning, system tuning, and infrastructure cost/performance optimization.
- •Maintain and improve our GitOps-driven deployment pipelines.
- •Requirements 7-8+ years of hands-on experience in an SRE or DevOps role, managing production environments at scale.
- •Deep experience with AWS (architecture, security, and best practices).
- •Strong programming skills in Python, Go, or TypeScript. 3+ years working with Infrastructure as Code
- •AWS CDK or Terraform.
- •Production experience with Kubernetes (EKS) and microservices architecture.
- •Hands-on experience with CI/CD and GitOps tools (GitHub Actions, ArgoCD).
- •Solid understanding of Linux internals and networking fundamentals.
- •Good communication skills in English and Hebrew.
Required skills
AWSPythonGoTypeScriptTerraformKubernetesEKSCI/CDGitHub ActionsArgoCDLinux
Nice-to-have skills
DatadogPrometheusGrafanaHelm
Domain expertise
fintech
Tech stack
AWSPythonGoTypeScriptTerraformKubernetesEKSCI/CDGitHub ActionsArgoCDLinux