NVIDIA AI
US - Washington - Seattle
View Company Profile /
<< Go Back
You will design, build, and maintain large-scale AI infrastructure platforms to support training, inferencing, and Agentic AI workloads. Additionally, you will optimize system efficiency, troubleshoot failures from application to hardware levels, and define reliability metrics.
Requirements: Candidates must have at least 8 years of experience in software infrastructure for large-scale AI systems and a bachelor's degree in a technical field. Proficiency in distributed systems, Kubernetes, and programming languages like Python and Golang is required.
Key Skills: AI Infrastructure, Distributed Systems, Golang, Python, C/C++, Kubernetes, Observability, Prometheus, Loki, ELK, AI Training, Inferencing, Debugging, Root Cause Analysis, Cloud-native Infrastructure, Data Infrastructure
Benefits: Equity, Benefits
© 2026 engineeringjobs.net, Inc. All Rights Reserved.
Terms of Service | Privacy
Powered by JOBBEX