Senior Site Reliability and Infrastructure Engineer
TreeswiftVegetation Intelligence company
New York, United States$160,000 - $220,000 USDSenior
Crosslink Capital
Susa Ventures
Inspired Capital
Pathbreaker Ventures
TenOneTen Ventures
Contour Venture Partners
Software Engineering
About the role
TL;DR
Lead SRE/infrastructure engineering to scale and harden Treeswift's platform.
- •Treeswift is seeking its first full-time SRE/infrastructure engineer to lead improvements and scaling of their platform infrastructure.
- •This role will focus on productionizing data pipelines, machine learning training platforms, and web applications.
- •Key Responsibilities Design and implement reliability and observability for high-volume pipeline operations.
- •Own CI/CD guardrails for production changes and safe rollout mechanics.
- •Make machine learning inference operations more reliable and observable.
- •Create operational tooling and continuously improve systems, including runbooks and automation.
- •Requirements 7-10 years of experience in observability, systems/infrastructure engineering, SRE, or DevOps in a cloud environment.
- •Hands-on experience with infrastructure-as-code (Terraform).
- •Experience with container orchestration (Kubernetes and/or ECS).
- •Strong Linux debugging skills and ability to investigate production issues.
Required skills
AWSKubernetesCI/CDTerraformLinuxGitPythonBash
Nice-to-have skills
AirflowECSS3
Domain expertise
aideveloper-tools
Benefits & perks
Competitive salary and equity package, Comprehensive medical, dental, and vision coverage, Life insurance and short- and long-term disability coverage, 16 weeks of fully paid parental leave, Flexible, unlimited paid time off, 401(k) retirement savings plan, Free OneMedical membership, Commuter benefits, Snacks, goodies, and team lunches
Tech stack
PythonBashAWSKubernetesDockerCI/CDAirflowECSS3TerraformLinuxGit