Jack & Jill
US - California - San Francisco
View Company Profile /
<< Go Back
Own and manage multi-region GPU inference infrastructure and EKS clusters to power real-time AI conversations. Collaborate with researchers to optimize CUDA hot paths and reduce cold-start times for global deployments.
Requirements: Requires hands-on experience deploying large-scale GPU workloads on providers like AWS or CoreWeave. Deep expertise in Kubernetes/EKS and a proven track record of senior technical leadership are essential.
Key Skills: GPU Inference, Kubernetes, EKS, CUDA, Infrastructure Design, Cloud Computing, Service Routing, Cluster Management, Technical Leadership
Benefits: Equity
© 2026 engineeringjobs.net, Inc. All Rights Reserved.
Terms of Service | Privacy
Powered by JOBBEX