Location
London
Hours
Full Time
Salary
Competitive, commensurate with experience
About the Role
At GSK, we aim to supercharge our data capability to better understand our patients and accelerate the discovery of vaccines and medicines. The Onyx Research Data Platform organization is a major investment by GSK R&D and Digital & Tech, designed to deliver a step-change in leveraging data, knowledge, and prediction to find new medicines.
We are a full-stack team including product and portfolio leadership, data engineering, infrastructure and DevOps, data/metadata/knowledge platforms, and AI/ML and analysis platforms. Our mission is to build a next-generation, metadata- and automation-driven data experience for GSK's scientists, engineers, and decision-makers, increasing productivity and reducing time spent on "data mechanics".
The Compute Platform Engineering team is building a first-class platform of toolchains and workflows that accelerate application development, scale computational experiments, and integrate computation with project metadata, logs, experiment configuration, and performance tracking across Cloud and High-Performance Computing (HPC) environments. This metadata-forward, CI/CD-driven platform supports the entire application and analysis lifecycle including interactive development (notebooks), large-scale batch processing, observability, and production deployments.
A Compute Platform Engineer II is a technical contributor who transforms loosely defined business or technical problems into well-defined specifications and executes them at a high level. They focus on metrics for impact and operations, model best practices in software development including code quality, documentation, DevOps, and testing, and ensure robustness of services. They serve as an escalation point for operations of existing services, pipelines, and workflows and engage with open-source communities relevant to their specialization.
Key Responsibilities:
- Design, build, and operate tools, services, and workflows that solve key business problems
- Develop key components of a hybrid on-prem/cloud compute platform for interactive and scalable batch computing
- Establish processes to transition HPC users and teams to the new platform
- Manage code-driven environments, container/image builds, and CI/CD-driven deployments
- Consult scientific users on application scalability to petabytes of data, leveraging deep understanding of software engineering, algorithms, and hardware infrastructure
- Optimize design and execution of complex solutions in large-scale distributed computing environments
- Produce well-engineered software with automated tests, documentation, and operational strategies
- Ensure consistent application of platform abstractions for logging and lineage
- Participate in code reviews and promote coding best practices
- Adhere to QMS framework and CI/CD best practices and contribute to their continuous improvement
Experience
- Bachelor's degree in Data Engineering, Computer Science, or Software Engineering
- 4+ years of professional experience
- Proficient in Python
- Experience with Cloud platforms
- Experience with High Performance Computing (HPC)
Preferred Qualifications
- Proficiency in at least one additional programming language such as Go, C++, Scala, or Java, including toolchains for documentation, testing, and operations/observability
- Expertise in modern software development tools and practices (e.g., git/GitHub, DevOps tools, metrics/monitoring)
- Cloud expertise (AWS, Google Cloud, Azure) including infrastructure-as-code and scalable compute technologies like Google Batch and Vertex
- Experience with CI/CD implementations using common stacks (Azure DevOps, CloudBuild, Jenkins, CircleCI, GitLab)
- Expertise with Docker, Kubernetes, CNCF ecosystem, and deployment tools such as Helm
- Experience with low-level build tools (make, CMake) and automated build systems (spack, easybuild)
- Experience with workflow orchestration tools such as Argo Workflow, Airflow, and scientific workflow tools like Nextflow, Snakemake, VisTrails, or Cromwell
- Application performance tuning and optimization experience in parallel and distributed computing paradigms and communication libraries (MPI, OpenMP, Gloo)
- Deep understanding of underlying systems (hardware, networks, storage) and their impact on application performance
- Demonstrated excellence in agile software development environments using tools like Jira and Confluence
- Familiarity with high-performance application tools, techniques, and optimizations including engagement with open-source communities and contributions
About You
You are a highly skilled engineer with a passion for building scalable, robust compute platforms. You have strong problem-solving skills and a commitment to software quality and best practices. You enjoy collaborating with scientific users and cross-functional teams to deliver impactful solutions. You are proactive in adopting and improving development processes and thrive in an agile environment.
GSK










