Staff Engineer, Kubernetes GPU Inference Platform
Togetherai
Together AI, the AI Native Cloud, is seeking a software engineer to build a Kubernetes-native control plane that provisions and runs our GPU inference fleet.
You will design a manifest-driven API where the inference team declares what they need, whether that's a cluster, a model deployment, or a capacity change, and our controllers handle the reconciliation, provider/runtime selection, and lifecycle management underneath, so the inference team never has to know or care which specific serving
#J-18808-LjbffrReference: WJ-766_22200716