NVIDIA AI
US - New York - Santa Clara
View Company Profile /
<< Go Back
Design and develop agent-side systems to collect GPU health, host telemetry, and BMC data across NVIDIA's GPU fleet. Contribute to open-source software and collaborate with SRE and security teams to improve fleet observability.
Requirements: Requires 5+ years of software engineering experience with strong proficiency in Go and Linux systems. Candidates should have experience with Kubernetes, telemetry systems, and a willingness to work with Rust.
Key Skills: Go, Rust, Linux Systems, Kubernetes, Docker, Telemetry, Observability, OpenTelemetry, Prometheus, Redfish, BMC, GPU Diagnostics, Cuda, InfiniBand, Systemd, Open Source Contribution
Benefits: Equity, Benefits
© 2026 engineeringjobs.net, Inc. All Rights Reserved.
Terms of Service | Privacy
Powered by JOBBEX