Senior Site Reliability Engineer
Recorded FutureThreat Intelligence company
Gothenburg, SwedenSenior
Insight Partners
Balderton Capital
Google Ventures
Google
In-Q-Tel
Software Engineering
About the role
TL;DR
Ensures reliability, scalability, and performance of critical systems through automation and infrastructure management.
- •Recorded Future is seeking a highly motivated and experienced Senior Site Reliability Engineer (SRE) to join their growing team.
- •In this role, you will be instrumental in ensuring the reliability, scalability, and performance of their critical systems.
- •Key Responsibilities Ensure the performance, capacity, scalability, reliability, resiliency, security, compliance, support, cost efficiency, SLA, SLOs, RPOs and RTOs for the platform.
- •Perform comprehensive Root Cause Analysis for outages and make systemic improvements.
- •Design, implement, and maintain scalable and reliable infrastructure on AWS.
- •Develop and manage observability solutions using tools such as Grafana, ELK, and Prometheus.
- •Automate infrastructure provisioning and configuration using Terraform and Chef.
- •Requirements 3+ years of experience in a Site Reliability Engineer, DevOps Engineer, or similar role.
- •Extensive hands-on experience with Amazon Web Services (AWS).
- •Expert-level troubleshooting and diagnostic skills.
- •Advanced Linux skills.
- •Strong proficiency in Terraform and Chef.
Required skills
AWSTerraformChefLinuxGrafanaPrometheusKubernetesRabbitMQApache KafkaMongoDBElasticsearch
Domain expertise
cybersecurity
Tech stack
AWSGrafanaPrometheusTerraformChefLinuxKubernetesRabbitMQApache KafkaMongoDBElasticsearch