The point where experts and best companies meet
Share
What you’ll be doing:
Implement deep learning models from multiple data domains (CV, NLP/LLMs, ASR, TTS, RecSys and others) in multiple DL frameworks (PyT, JAX, TF2, DGL and others)
Implement and test new SW features (Graph Compilation, reduced precision training) that use the most recent HW functionalities.
Analyze, profile, and optimize deep learning workloads on state-of-the-art hardware and software platforms.
Collaborate with researchers and engineers across NVIDIA, providing guidance on improving the design, usability and performance of workloads.
Lead best-practices for building, testing, and releasing DL software
What we need to see:
5+ years of experience in DL model implementation and SW Development
BSc, MS or PhD degree in Computer Science, Computer Architecture, Mathematics, Physics or related technical field or equivalent experience
Excellent Python programming skills, extensive knowledge of at least one DL Framework
Strong problem solving and analytical skills
Algorithms and DL fundamentals
Ways to stand out from the crowd:
Experience in performance measurements and profiling
Experience with running large-scale workloads in HPC clusters
Knowledge and love for DevOps/MLOps practices for Deep Learning-based product’s development.
Solid understanding of Linux environments and containerization technologies such as Docker
GPU programming experience (CUDA or OpenCL) is a plus but not required.
These jobs might be a good fit