BH
Brevan HowardNew

Senior Site Reliability Engineer

onsiteFull timeLondon, GBYesterday

Job Overview & Requirements

Senior Site Reliability Engineer (SRE) - GCP/KubernetesAbout the RoleWe are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership.The successful candidate will split their time between providing expert operational support for our critical systems and leading exciting new infrastructure projects. Our mindset is to get the right person not the person with the skills that match our stack. It is important to be able to foresee problems before they show up and create solutions that mitigate them. If you enjoy a challenging environment, implementing "infrastructure as code" principles, and directly seeing the impact of your work, this is the place for you.What You Will DoDesign & Build: Architect, deploy, and maintain highly scalable and reliable infrastructure on Google Cloud Platform (GCP) using Kubernetes and Infrastructure-as-Code tools.Automation: Champion automation across the entire software development lifecycle (SDLC), utilizing IaC, Python and Bash to reduce toil and improve operational efficiency.Infrastructure-as-Code (IaC): Own and evolve our declarative infrastructure using Terraform for cloud resources and Helm for Kubernetes application deployment.Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification.Reliability & Performance: Define, measure, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Participate in on-call rotation (if applicable) and lead post-incident reviews to drive continuous improvement.Collaboration: Work closely with software development teams to provide expert guidance on deployment strategies, scalability concerns, and cloud-native best practices.Ownership: Take full ownership of projects from inception through to production operation, including documentation and knowledge 

Job Snapshot

Employer:Brevan Howard
Location:London, GB
Work Mode:onsite
Employment Type:Full time
Compensation:Competitive
Date Posted:Yesterday
Verified Creator Feed

This job posting is officially syndicated with tracking preserved to ensure prioritized creator-referred application review.

Similar Open Positions

remotefull time💼5-9 yrsSan Francisco, California, US4h ago

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to a...

$165k - $225k / yr
ReactTypeScriptNext.js+3
remotefull time💼5-9 yrsSan Francisco, California, US4h ago

About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world's largest enterprises to the most ambitious startups—use Stripe to a...

$165k - $225k / yr
ReactTypeScriptNext.js+3
remotefull time💼6-12 yrsUS12h ago

About Anthropic Anthropic is an AI safety and research company that builds reliable, beneficial AI systems. We are dedicated to creating state-of-the-art AI while conducting founda...

$240k - $360k / yr
PythonPyTorchRust+3
BH

Senior Site Reliability Engineer

Competitive

Apply Now