Senior DBRE to own reliability, scalability, and automation of PostgreSQL, Elasticsearch, and Kafka.
•Cognite is looking for a Senior Database Reliability Engineer (DBRE) to own the reliability, scalability, automation, and operational excellence of their core database infrastructure.
•You will work across PostgreSQL, Elasticsearch, and Kafka on a multi-cloud, Kubernetes-based platform.
•Key Responsibilities Standardize and automate the lifecycle management of 1000+ PostgreSQL instances across Azure, AWS, and GCP using Infrastructure as Code.
•Design, operate, and scale high-performance Elasticsearch clusters in Elastic Cloud (SaaS) and ECK (self-managed Kubernetes) environments.
•Own the reliability, scalability, and performance of Kafka clusters supporting high-throughput, low-latency event streaming.
•Partner closely with Software Engineering, SRE, Platform Engineering, and Product teams to ensure data platforms are highly available, scalable, secure, and resilient.
•Automate provisioning, configuration, patching, upgrades, backups, and operational workflows for database infrastructure.
•Requirements 6+ years of experience in Database Reliability Engineering, Database Engineering, SRE, or Platform Engineering.
•Strong hands-on experience operating PostgreSQL at scale, preferably in cloud-managed environments.
•Strong experience with Elasticsearch, including cluster administration, performance tuning, scaling, and troubleshooting.
•Experience operating databases and stateful workloads on Kubernetes.
•Experience with Infrastructure as Code using Terraform or similar tools.
•Proficiency in scripting or programming using Python, Go, or a similar language.