Location
London
Hours
Full Time
Salary
Competitive, commensurate with experience
About the Role
At GSK, we aim to supercharge our data capability to better understand our patients and accelerate vaccine and medicine discovery. The Onyx Research Data Platform organization is a major investment by GSK R&D and Digital & Tech, designed to transform our ability to leverage data, knowledge, and prediction to find new medicines.
We are a full-stack team including product and portfolio leadership, data engineering, infrastructure and DevOps, data/metadata/knowledge platforms, and AI/ML and analysis platforms. Our mission is to build a next-generation, metadata- and automation-driven data experience for GSK's scientists, engineers, and decision-makers, increasing productivity and reducing time spent on "data mechanics".
The Compute Platform Engineering team is building a first-class platform of toolchains and workflows that accelerate application development, scale computational experiments, and integrate computation with project metadata, logs, experiment configuration, and performance tracking across Cloud and High-Performance Computing (HPC) environments. This metadata-forward, CI/CD-driven platform supports the entire application and analysis lifecycle including interactive development (notebooks), large-scale batch processing, observability, and production deployments.
As a Senior Compute Platform Engineer, you will be a leading technical contributor who can transform loosely defined business or technical problems into well-defined solutions and execute them at a high level. You will focus on metrics for both impact and operational performance, model best practices in software development, mentor junior team members, ensure robustness of services, and act as an escalation point for operational issues. You will be deeply familiar with your specialization tools and actively engaged with the open-source community, potentially contributing improvements.
Key Responsibilities:
- Design, build, and operate tools, services, and workflows that solve key business problems
- Develop key components of a hybrid on-prem/cloud compute platform for interactive and scalable batch computing
- Establish processes and workflows to transition HPC users to the new platform
- Manage code-driven environment, application, and container/image builds with CI/CD-driven deployments
- Consult science users on application scalability to petabytes of data, leveraging deep software engineering and hardware infrastructure knowledge
- Optimize design and execution of complex solutions in large-scale distributed computing environments
- Produce well-engineered software with automated tests, documentation, and operational strategies
- Ensure consistent application of platform abstractions for logging and lineage
- Participate in code reviews and promote coding best practices
- Adhere to QMS framework and CI/CD best practices and contribute to their continuous improvement
- Provide leadership and mentorship to team members
Experience
- Bachelor's degree in Data Engineering, Computer Science, Software Engineering, or related discipline
- 6+ years of professional experience
- Proficient in Python
- Experience with Cloud platforms
- Experience with High Performance Computing (HPC)
Preferred Qualifications
- Deep knowledge of at least one common programming language such as Python, C++, or Java, including toolchains for documentation, testing, and operations/observability
- Expertise in modern software development tools and practices (e.g., git/GitHub, DevOps tools, metrics/monitoring)
- Strong cloud expertise (AWS, Google Cloud, Azure), including infrastructure-as-code and scalable compute technologies like Google Batch and Vertex
- Experience with CI/CD implementations using git and common CI/CD stacks (Azure DevOps, CloudBuild, Jenkins, CircleCI, GitLab)
- Proficiency with Docker, Kubernetes, and CNCF ecosystem tools including Helm
- Familiarity with low-level build tools (make, CMake) and automated build systems (spack, easybuild)
- Experience with workflow orchestration tools such as Argo Workflow, Airflow, Nextflow, Snakemake, VisTrails, or Cromwell
- Skilled in application performance tuning and optimization in parallel and distributed computing, including MPI, OpenMP, Gloo
- Deep understanding of underlying systems (hardware, networks, storage) and their impact on performance
- Experience working in agile software development environments using tools like Jira and Confluence
- Engagement with open-source communities in the high-performance applications space, including potential contributions
About You
You are a highly skilled software engineer with a passion for building scalable, robust compute platforms. You have strong problem-solving skills, a focus on quality and best practices, and enjoy mentoring others. You thrive in collaborative environments and are eager to contribute to cutting-edge scientific computing solutions that accelerate discovery and innovation.
GSK










