Company Logo
Software Engineer

Netflix - 1d ago

Company Logo
Senior Software Engineer

Reddit - 4d ago

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Requirements

  • Strong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.
  • Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.
  • Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.
  • Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.
  • Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)
  • Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.
  • Knowledge of networking protocols(VPC) , load balancing strategies in a distributed systems environment , database query performance tuning and identify cause for lagging.

Nice to Haves

  • Bachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.
  • 7 to 10 years of experience.

What you'll be doing

  • Partner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.
  • Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.
  • Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.
  • Improve production readiness through load testing, performance tuning, capacity forecasting, chaos engineering and reliability reviews.
  • Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.
  • Promote sustainable operations by helping design healthy on-call models, clear escalation paths, and balanced pager responsibilities.

Perks and Benefits

  • Opportunities for growth professionally and personally.
  • Training and development opportunities.
  • Firmwide networks.
  • Benefits, wellness, and personal finance offerings.
  • Mindfulness programs.
AI Summary ✨
Goldman Sachs logo

Goldman Sachs

Birmingham, UK

Experience: Senior
Posted: August 27, 2026
Last seen: an hour ago
Aws
Azure
Docker
Gcp
Java
Javascript
Kubernetes
Nodejs
Python
Terraform
sitereliability

Why we track Goldman Sachs

Goldman Sachs has large engineering teams in London and other EU cities. They've been investing heavily in technology, and the engineering work goes well beyond traditional finance—platform engineering, cloud infrastructure, developer tools. The pay is competitive with FAANG.

Similar jobs

  • 3 days ago
  • maze logo

    Senior Site Reliability Engineer | SRE (Remote, EU/CET)

    UK, Ireland, Portugal, Spain, Netherlands

    3 days ago
    Remote
  • 9 days ago
  • See all jobs in UK