Start Your Search Here

push notification bell

Would you like to receive notifications about IT & Computers jobs in London?

push notification bell

You have blocked notifications

Oops! You have blocked notifications. Click here for more info

You have blocked notifications, please check your browser settings.

push notification bell

You're currently subscribed to job notifications

Want to change your notifications for job alerts?

push notification bell

Subscribe to notifications

You will no longer receive notifications

Job Search

Harnham - Data & Analytics Recruitment Careers

London / Global

GCP Data Engineer

  • Hybrid

Job Summary

Job Type:
Contract
Work Settings:
Hybrid
Benefits:
Healthcare
Apply Now

Job Description

GCP Data Engineer

£700 - £750 per day inside IR35

6-month contract

Hybrid working in London

We're working with a global healthcare and AI research organisation at the forefront of applying data engineering and machine learning to accelerate scientific discovery. Their work supports large-scale, domain-specific datasets that power research into life-changing treatments.

They're now looking for a GCP Data Engineer to join a multidisciplinary team responsible for building and operating robust, cloud-native data infrastructure that supports ML workloads, particularly PyTorch-based pipelines.

The Role

You'll focus on designing, building, and maintaining scalable data pipelines and storage systems in Google Cloud, supporting ML teams by enabling efficient data loading, dataset management, and cloud-based training workflows.

You'll work closely with ML engineers and researchers, ensuring that large volumes of unstructured and structured data can be reliably accessed, processed, and consumed by PyTorch-based systems.

Key Responsibilities

Design and build cloud-native data pipelines using Python on GCP

Manage large-scale object storage for unstructured data (Google Cloud Storage preferred)

Support PyTorch-based workflows, particularly around data loading and dataset management in the cloud

Build and optimise data integrations with BigQuery and SQL databases

Ensure efficient memory usage and performance when handling large datasets

Collaborate with ML engineers to support training and experimentation pipelines (without owning model development)

Implement monitoring, testing, and documentation to ensure production-grade reliability

Participate in agile ceremonies, code reviews, and technical design discussions

Tech Stack & Experience Must Have

Strong Python development experience

Hands-on experience with cloud object storage for unstructured data

(Google Cloud Storage preferred; AWS S3 also acceptable)

PyTorch experience, particularly:

Dataset management

Data loading pipelines

...

Apply Now

Similar Opportunities

View all jobs