Site Reliability Engineer, Cryptography, Access and Identity Services


Location
London
Hours
Full Time
Salary
Competitive, commensurate with experience
About the Role
Are you interested in working with the latest cloud computing technologies and becoming a core part of the largest cloud infrastructure on the planet? AWS is seeking a Systems Engineer to build and operate services for our customers, automate service operations and deployment methods, and provide mentoring for junior team members. You will work in a large-scale, high availability environment, building and operating critical Cryptography, Access and Identity services for our customers. This role is designed for someone with a strong engineering background and a passion for driving efficiency, quality, and process improvements within service operations. As part of an operations team focused on automation, you will utilize AI tooling and develop solutions to automate manual tasks.
As a Systems Engineer at Amazon, you will use your Linux skills to troubleshoot, innovate fixes and workarounds, maintain software updates, and provide data and metrics to manage capacity and efficiency. You will apply your networking knowledge to identify and resolve connectivity issues. This role offers the opportunity to develop a broad range of skills, take ownership, and gain knowledge across a complex technology stack.
Top reasons to join our team:
- Learn how to build and operate distributed systems at massive scale
- Be a catalyst to deliver disruptive products that are growing rapidly
- Solve unique, large-scale problems across many AWS services
- Build and influence the tools and utilities running AWS internal services
- Deliver solutions that reduce manual toil across the organization
Key Responsibilities
Amazon fosters a collaborative, purposeful, and enthusiastic environment where we "Work Hard, Have Fun, Make History." Systems Engineers are expected to be leaders by becoming subject matter experts on multiple AWS services. They develop, build, deploy, operate, sustain, and grow their services in cloud production environments. They utilize trends and metrics to identify improvement opportunities, help refine procedures for their team and internal customers, and consistently deliver customer-impacting changes. Engineers handle new and ambiguous problem domains while maintaining high standards.
A Day in the Life
Typical daily activities include building infrastructure and services for customers, investigating root causes of customer issues, analyzing metrics, and consulting with senior engineers. We strive to minimize out-of-hours support by implementing Operational Excellence best practices and automating manual processes.
About the Team
AWS values diverse experiences and encourages candidates from all backgrounds to apply. We foster an inclusive culture that promotes curiosity, connection, and collaboration. Our employee-led affinity groups and inclusion events empower our people and fuel innovation. We provide mentorship, career growth opportunities, and support work-life harmony to help you succeed both professionally and personally.
Experience
- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration
- Experience working with Linux
- Experience in systems engineering
- Experience in one or more programming or scripting languages such as Python, Java, Perl, PHP, Ruby, Bash, or Shell
About You
You are passionate about automation, efficiency, and quality improvements. You thrive in a fast-paced, large-scale environment and enjoy solving complex problems. You are a proactive learner and collaborator who embraces new challenges and ambiguous problem domains. You value diversity and inclusion and are eager to contribute to a supportive team culture.
Qualifications
- Knowledge of TCP/IP and networking protocols such as HTTP and DNS
- Experience designing and developing scripts to automate operational tasks, ensuring maintainability, scalability, and security
- Experience working in a 24/7 production environment
- Experience with service-oriented architecture and web services

