Drive platform reliability and engineering efficiency through automation and best practices.
•As a Site Reliability Engineer, you’ll be helping to drive our product and engineering department forward, ensuring reliability on different parts of the Paddle platform and helping our Engineers to work better and more efficiently.
•Key Responsibilities Develop and maintain tools to maximise engineering efficiency.
•Seek out processes that can be improved with automation.
•Create, maintain and test our system disaster recovery process.
•Handle production incidents and author blameless postmortems.
•Monitor, alert, and track SLOs.
•Requirements A software development background with experience operating production services.
•Experience working across the AWS ecosystem.
•Knowledge of platform and ops concepts such as networking and Linux administration.
•Experience working with microservices and distributed systems at scale.
•Curiosity about AI and how it's reshaping software development.