IT & Software

Staff Platform Engineer - GPU Inference Control Plane

Togetherai

London · England · United Kingdom

Together AI is seeking a software engineer to build the Kubernetes-native control plane that provisions and runs our GPU inference fleet. You will design a manifest-driven API where the inference team declares needs such as a cluster or a model deployment, and our controllers reconcile and manage lifecycle behind the scenes.

You will own reliability, performance optimizations, defragmentation, and scheduling improvements, shaping a platform that decouples developers from runtime complexity.

#J-18808-Ljbffr

Reference: WJ-766_22384680

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.