Senior Software Engineer, Deep Learning Inference - TensorRT

NVIDIA AI
US - California - Santa Clara
View Company Profile / << Go Back

  • Job Type: Full time
  • 12 days ago

Job Description

Develop and scale robust inferencing software and components for NVIDIA's TensorRT SDK to accelerate deep learning models. Collaborate with GPU architects and experts to build graph parsers, optimizers, and tools for model deployment.

Requirements: Requires a degree in Computer Science or a related field and at least 3 years of software development experience. Candidates must have strong proficiency in modern C++ standards and a solid grasp of machine learning and computer architecture.

Key Skills: C++, Python, Deep Learning Inference, TensorRT, CUDA, OpenCL, Machine Learning, Computer Architecture, Data Structures, Algorithms, PyTorch, TensorFlow, ONNX Runtime, Compiler Development, Performance Benchmarking, GPU Kernel Programming

Benefits: Equity




Fast Track Upload