NVIDIA AI
US - California - Sonoma County
View Company Profile /
<< Go Back
Develop libraries, code generators, and GPU kernel technologies to accelerate AI inference for large language models. Design and optimize kernels and extensible abstractions for LLM serving engines and just-in-time compilers.
Requirements: Requires a Master's or PhD in Computer Science or Electrical Engineering with 6+ years of experience in ML/DL systems development. Must have strong proficiency in Python, C/C++, and GPU kernel optimization using CUDA or Triton.
Key Skills: GPU Kernel Development, CUDA C/C++, Python, C++, Matrix Multiplication, PyTorch, JAX, TensorFlow, vLLM, SGLang, Triton, MLIR, Apache TVM, LLM Inference, Deep Learning Frameworks, Compiler Design
Benefits: Equity
© 2026 engineeringjobs.net, Inc. All Rights Reserved.
Terms of Service | Privacy
Powered by JOBBEX