Jack & Jill
US - California - San Francisco
View Company Profile /
<< Go Back
You will design and manage multi-region GPU deployments and architect EKS clusters to optimize workload placement. Additionally, you will collaborate with researchers to optimize CUDA hot paths and reduce latency for real-time AI conversations.
Requirements: The ideal candidate has hands-on experience deploying large-scale GPU inference workloads on providers like AWS or CoreWeave. You must demonstrate deep expertise in Kubernetes/EKS and a proven track record of senior technical leadership.
Key Skills: GPU Inference, Kubernetes, EKS, CUDA, Cloud Infrastructure, Multi-region Deployment, Backend Services, Service Routing, Cluster Management, Technical Leadership, AWS, CoreWeave
Benefits: Equity
© 2026 engineeringjobs.net, Inc. All Rights Reserved.
Terms of Service | Privacy
Powered by JOBBEX