IT & Software

Staff Engineer, Kubernetes GPU Inference Platform

Togetherai

London · England · United Kingdom

Together AI, the AI Native Cloud, is seeking a software engineer to build a Kubernetes-native control plane that provisions and runs our GPU inference fleet.

You will design a manifest-driven API where the inference team declares what they need, whether that's a cluster, a model deployment, or a capacity change, and our controllers handle the reconciliation, provider/runtime selection, and lifecycle management underneath, so the inference team never has to know or care which specific serving

#J-18808-Ljbffr

Reference: WJ-766_22200716

Apply now

Continue on the employer's official application - the same link they use for every candidate.

More jobs

Find more on GigBlows

This role is listed on GigBlows for discovery and search. Hiring decisions and applications are handled by the employer or their chosen application system.