Software Engineer, AI Compute Infrastructure

HeyGen
US - California - Los Angeles
View Company Profile / << Go Back

  • Job Type: Full time
  • 30+ days ago

Job Description

Build and scale the foundational compute infrastructure and large-scale AI job frameworks required for generative video models. Optimize GPU utilization and enhance observability tools to ensure reliable, high-throughput, and low-latency performance.

Requirements: Requires a minimum of 5 years of industry experience in MLOps, AI infrastructure, or HPC systems with a degree in Computer Science or Engineering. Must be proficient in Python, C++, and distributed computing frameworks like Kubernetes and Ray.

Key Skills: Python, C++, Kubernetes, Ray, PyTorch, TensorFlow, JAX, CUDA, NCCL, MLOps, HPC, Apache Spark, LanceDB, Distributed Computing, GPU Optimization, Generative AI

Benefits: Competitive salary, Benefits package




Fast Track Upload