Deep Learning Software Engineer, Inference - New College Grad 2026

NVIDIA AI
US - California - Orange County
View Company Profile / << Go Back

  • Job Type: Full time
  • 30+ days ago

Job Description

Design, build, and optimize GPU-accelerated software for high-performance deep learning inference and large-scale model serving. Focus on improving performance for LLM and Generative AI models across various NVIDIA accelerators and contributing to open-source inference libraries.

Requirements: Candidates should be pursuing or have recently completed a Master's or PhD in Computer Engineering, Computer Science, or a related field. Strong proficiency in C/C++ and software design is required, with experience in GPU programming or deep learning optimization being a significant plus.

Key Skills: Deep Learning Inference, C/C++, Python, CUDA, OAI Triton, CUTLASS, NCCL, Performance Optimization, LLM, Generative AI, Software Design, Agile, GPU Programming, Model Serving, NVSHMEM, PyTorch

Benefits: Equity, Benefits




Fast Track Upload