Together AI
Greater London / Global
Staff Software Engineer, Kubernetes-native GPU Inference
- £100000
You have blocked notifications
Oops! You have blocked notifications. Click here for more info
You have blocked notifications, please check your browser settings.
You're currently subscribed to job notifications
Subscribe to notifications
You will no longer receive notifications
Greater London / Global
Together AI is seeking a Staff Software Engineer to design a Kubernetes-native control plane that provisions and runs a GPU inference fleet across London and Amsterdam. You’ll build a manifest-driven API, contribute self-service tooling, and optimize scheduling, defragmentation, and capacity balancing to push utilization while preserving latency.
You’ll own the provisioning lifecycle, build reliable pipelines, and partner with the ML platform team to encode cluster shapes as first-class
#J-18808-LjbffrGreater London / Global
Greater London / Global
Greater London / Global
Greater London / Global
Greater London / Global
Greater London / Global