Senior MLOps Engineer - DSX Enablement
NVIDIA
NVIDIA
Join NVIDIA's DSX Enablement team as a Senior MLOps Engineer, partnering with leading AI companies and open-source communities. You'll develop cutting-edge AI solutions, advise on infrastructure needs for ML workloads, and help customers resolve complex full-stack AI/ML system challenges. This role is key to driving the success of internal and external customer initiatives, including LLM performance and supporting new hardware in open-source frameworks.
Develop and deploy custom AI solutions on cloud platforms, including distributed training, inference optimization, and MLOps pipelines. Serve as the primary technical liaison for customers and partners, guiding joint projects and resolving production issues. Collaborate with infrastructure software and accelerated framework development teams. Profile and tune large-scale training and inference workloads to enhance performance and reduce costs. Create open-source tools and reference architectures for scalable machine learning and AI system management.
A Bachelor's, Master's, or Ph.D. in Computer Science, Engineering, or a related field, or equivalent experience is required. Minimum 8 years of experience in technical roles such as data science, data engineering, or ML engineering, preferably with large-scale production systems. Proven AI/ML experience across the entire lifecycle, from exploration to production. Proficiency in systems topics including Linux, schedulers, Kubernetes, distributed filesystems, and networking. Strong scripting and programming skills in bash, Python, and systems programming in C++, Go, or Rust. Experience with ML/DL frameworks for training and inference. Excellent communication and presentation skills for technical and leadership audiences. A solid track record of engineering discipline and project execution.
Nvidia
IT Consulting