Senior Site Reliability Engineer
CartaEquity Management company
London, United KingdomSenior
Andreessen Horowitz
Lightspeed Venture Partners
Spark Capital
Tribe Capital
Union Square Ventures
Silver Lake
Software Engineering
About the role
TL;DR
Build and scale infrastructure for reliability and performance.
- •The Company You’ll Join Carta connects founders, investors, and limited partners through world-class software, purpose-built for everyone in venture capital, private equity and private credit.
- •Trusted by 65,000+ companies in 160+ countries, Carta’s platform of software and services lays the groundwork so you can build, invest, and scale with confidence.
- •Carta's Fund Administration platform supports 9,000+ funds and SPVs, representing nearly $185B in assets under management, with tools designed to enhance the strategic impact of fund CFOs.
- •Recognized by Fortune, Forbes, Fast Company, Inc. and Great Places to Work, Carta is shaping the future of private market infrastructure.
- •Together, Carta is creating the end-to-end ERP platform for private markets.
- •Traditional ERP solutions don’t work for Private Funds.
- •Private capital markets need a comprehensive software solution to replace outdated spreadsheets and fragmented service providers.
- •Carta’s software for the Office of the Fund CFO does just that
- •it’s a new category of software to make private markets look more like public markets
- •a connected ERP for private capital.
- •For more information about our offices and culture, check out our Carta careers page.
- •The Problems You’ll Solve At Carta, our employees set out on a mission to unlock the power of equity ownership for more people in more places.
- •We believe that the problems we solve today unlock the opportunities of tomorrow.
- •As a Senior Site Reliability Engineer, you’ll work to: Build and scale our internal platform offerings (compute, storage and networking services) to ensure the reliability, and performance of our applications.
- •Design and implement monitoring, alerting, and incident response systems.
- •Collaborate with application software engineers (as needed) to guide their design and ensure it scales for what Carta needs in the long run.
- •Act as an agent of change and push boundaries to incrementally improve our systems as we expand globally.
- •The Team You’ll Work With You’ll be joining the Infrastructure Engineering team at Carta.
- •The Infrastructure Engineering team is responsible for providing secure, reliable, scalable and performant infrastructure to Carta’s customers and developers.
- •We are Software and Infrastructure Engineers who specialize in cloud computing, networking, systems design and architecture, storage, real time data telemetry, associated automation, tooling and processes.
- •We possess a breadth and depth of knowledge about Carta’s infrastructure and industry wide best practices, that translates into leverage for Carta’s business.
- •About You You are excited by the idea of developing scalable, reliable and efficient infrastructure that powers the entire company.
- •We’re looking for strong communicators who enjoy collaborating to solve complex problems.
- •Familiarity with infrastructure best practices on performance, reliability and security and their associated tools is appreciated.
- •Our stack is Python, Java, Terraform, gRPC, Docker, Kubernetes, Postgres, running on AWS.
- •Come join us! Cloud Platforms: Extensive experience with cloud services such as AWS, Google Cloud Platform, or Azure, including services like EC2, S3, RDS, and Lambda.
- •Experience with Kubernetes or other container orchestration is preferred! Infrastructure as Code (IaC): Proficient in using tools such as Terraform, Ansible, or CloudFormation for managing and provisioning cloud infrastructure.
- •Networking: Experience with networking concepts and tools, including Container Network Interface (CNI), Network policy implementations.
- •Experience with proxies and service mesh is a big plus.
- •Monitoring and Observability: Strong knowledge of monitoring tools and practices, such as Prometheus, Grafana, ELK Stack, or Datadog, and the ability to set up and maintain comprehensive monitoring solutions.
- •Software Development: Proficiency in Python, with the ability to write efficient, maintainable, and scalable code.
- •API Services: Experience in designing, deploying, and maintaining API services, with a strong understanding of RESTful and/or GraphQL API design principles.
- •AI Fluency: You use AI tools in your own day-to-day work in addition to enabling others.
- •You're comfortable building agents to reduce toil and expect this to be a normal part of how you operate.
- •Experience operating CI/CD and its associated best practices is also appreciated though not essential.
Required skills
KubernetesAWSTerraformPythonDockerPostgreSQLCI/CDPrometheusGrafanaDatadogREST APIgRPCAnsible
Nice-to-have skills
Google CloudAzureCloudFormationGraphQL
Domain expertise
developer-tools
Tech stack
PythonJavaTerraformgRPCDockerKubernetesPostgreSQLAWSEC2S3RDSAnsibleCloudFormationPrometheusGrafanaDatadogREST APIGraphQLCI/CD