Staff Platform Engineer - GPU Inference Control Plane
Togetherai
Together AI is seeking a software engineer to build the Kubernetes-native control plane that provisions and runs our GPU inference fleet. You will design a manifest-driven API where the inference team declares needs such as a cluster or a model deployment, and our controllers reconcile and manage lifecycle behind the scenes.
You will own reliability, performance optimizations, defragmentation, and scheduling improvements, shaping a platform that decouples developers from runtime complexity.
#J-18808-LjbffrReference: WJ-766_22384680