Start Your Search Here

push notification bell

Would you like to receive notifications about jobs in Greater London?

push notification bell

You have blocked notifications

Oops! You have blocked notifications. Click here for more info

You have blocked notifications, please check your browser settings.

push notification bell

You're currently subscribed to job notifications

Want to change your notifications for job alerts?

push notification bell

Subscribe to notifications

You will no longer receive notifications

Job Search

Capitalontap

Greater London / Global

Site Reliability Engineer

Job Summary

Benefits:
Healthcare Pension Travel Enhanced Ternity Cycle To Work
Apply Now

Job Description

SRE at Capital On Tap

At Capital On Tap, we run a hybrid embedded SRE model. We aim to work closely with the teams within Capital On Tap to provide them the best support. Our main objective currently is to gain as much visibility into our platform’s health while offering scalable solutions.

You’ll join our BAR (Banking and Repayments) Team, who are building and scaling payments and card systems, ensuring our financial platform is reliable, accessible, and seamless for users.

What You’ll be doing

  • Manage and automate Azure, Datadog, NGINX & Cloudflare
  • Develop and monitor Kubernetes and Serverless resources
  • Maintain infrastructure code with Terraform & CRDs / Crossplane
  • Improve systems, processes, and technologies; consult stakeholders to enhance platform performance
  • Getting involved in new application architecture & design processes
  • Design solutions to reduce toil, automate repetitive tasks and streamline workflows to reduce manual work and boost team productivity.
  • Create SLIs and SLOs; increase application visibility
  • Align with the Product team on SLAs and core service objectives
  • Collaborate with Platform Engineers for automated solutions and pipelines
  • Enhance user experience with infrastructure and pipeline optimisationSupport CI/CD tools such as Azure Devops, Octopus Deploy and Flux to streamline software delivery
  • Lead incident troubleshooting to safeguard customer experience

Requirements

  • Experience in managing a public cloud (Azure advantageous)
  • Experience in Azure DevOps, Octopus, Flux or other CI/CD tools
  • Experience with Linux and Microsoft Systems
  • Excellent communication skills and ability to collaborate with multiple teams in an agile environment
  • Proficient in contributing to IaC technologies involving expertise in writing, managing, and optimising infrastructure with tools such as Terraform and Pulumi
  • Experience working with a cloud monitoring solution (advantageous to have DataDog)
  • Experience with Kubernetes and Docke
  • Experience in at least one scripting language (Python, PowerShell, Go)

Interview Process

  • 1st stage: 30-minute intro, CV review, and values with Talent Partner
  • 2nd stage: 60 minute “Tech Chat” with Team Manager
  • Final stage: 75-minute Technical Task + 30 minute Interview with Head of Platform Engineering

Diversity & Inclusion

We welcome, consider and encourage applications from anyone who shares our commitment to inclusivity. Join us in creating a space where authenticity thrives, and everyone can do their best work.

Great Work Deserves Great Perks

  • Private Healthcare including dental and opticians services through Vitality
  • Worldwide travel insurance through Vitality
  • Anniversary Rewards (£250, £500, £750, 4-week fully paid sabbatical)
  • Salary Sacrifice Pension Scheme up to 7% match
  • 28 days holiday (plus bank holidays)
  • Annual Learning and Wellbeing Budget
  • Enhanced Parental Leave
  • Cycle to Work Scheme
  • Season Ticket Loan
  • 6 free therapy sessions per year
  • Dog Friendly Offices
  • Free drinks and snacks in our offices

Email if you have any questions.

#J-18808-Ljbffr
Apply Now

Similar Opportunities

View all jobs