Skip to Main Content
Location icon
London

Foundation Engineering - SRE Platforms - Site Reliability Engineer – Associate - London

Goldman Sachs
Office & Professional
Office & Professional
Company logo image
Description

Location
London

Hours
Full Time

Salary
Negotiable

About the Role
Goldman Sachs is undertaking one of its most ambitious engineering initiatives: the Consolidated Trade Ledger (CTL), a complete reimagining of the front-to-back architecture supporting every trade executed by the firm. This flagship, cloud-native platform is central to the firm's core technology strategy and is designed to deliver the capacity, extensibility, scalability, and innovation needed to power Global Markets' growth for the next two decades while driving significant operational efficiencies. We are seeking a Site Reliability Engineer to build, operate, and continuously improve this critical service. The role blends software engineering, systems engineering, and production expertise to enhance the reliability, scalability, observability, incident response, and operational efficiency of the CTL platform.

Requirements

Experience
- Strong programming skills in one or more modern languages, with Java or Go strongly preferred, including experience building maintainable automation beyond simple scripts.
- Solid understanding of networking, messaging, distributed systems, data structures, algorithms, and software design fundamentals.
- Hands-on experience with observability tooling such as Prometheus, Grafana, ELK, or OpenTelemetry for metrics, logging, tracing, and dashboarding.
- Proven ability to investigate production issues, identify root causes, and implement durable engineering fixes to improve system behavior and reduce repeat incidents.
- Experience supporting mission-critical production systems, preferably within financial services.
- Familiarity with public cloud platforms, especially GCP, cloud-native architecture, Kubernetes, microservices, or service mesh technologies.
- Experience with relational databases and/or data-intensive platforms.
- Knowledge of progressive delivery techniques such as canary releases, blue/green deployments, feature flags, or chaos/game days.
- Experience developing AI-assisted operations capabilities including alert enrichment, anomaly detection, triage support, or runbook automation.

About you
Excellent written and verbal communication skills with the ability to translate complex technical issues into clear updates for technical, business, and senior stakeholders. You demonstrate a calm and disciplined approach to incident leadership, clear communication, and a commitment to blameless learning. You possess a strong engineering mindset focused on automation, coding, debugging, testing, scalable design, and systems thinking.

Qualifications
Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.

Expiry date: 04/11/2026
Foundation Engineering - SRE Platforms - Site Reliability Engineer – Associate - London
Company:
Goldman Sachs
Job Type:
Full-time
Location:
London