Resume and JobRESUME AND JOB
NVIDIA logo

Senior Deep Learning Performance Engineer - Training at Scale

NVIDIA

Engineering Jobs

Senior Deep Learning Performance Engineer - Training at Scale

full-timePosted: Aug 22, 2025

Job Description

We are looking for senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out of Deep Learning training, inference and NVIDIA AI Services. We are working across all layers of the hardware/software stack, from GPU architecture to Deep Learning Framework, to achieve peak performance. This role offers an opportunity to directly impact the hardware and software roadmap in a fast-growing company that leads the AI revolution. Join the team building software used by the entire world. Work with world class software engineers to implement blazingly fast SOTA deep learning models that help understanding the end-to-end performance of NVIDIA’s DL software and hardware stack. Work on most powerful, enterprise-grade GPU clusters capable of hundreds of Peta FLOPS and on unreleased hardware before anyone in the world.What you’ll be doing:Implement deep learning models from multiple data domains (CV, NLP/LLMs, ASR, TTS, RecSys and others) in multiple DL frameworks (PyT, JAX, TF2, DGL and others)Implement and test new SW features (Graph Compilation, reduced precision training) that use the most recent HW functionalities.Analyze, profile, and optimize deep learning workloads on state-of-the-art hardware and software platforms.Collaborate with researchers and engineers across NVIDIA, providing guidance on improving the design, usability and performance of workloads.Lead best-practices for building, testing, and releasing DL softwareWhat we need to see:5+ years of experience in DL model implementation and SW DevelopmentBSc, MS or PhD degree in Computer Science, Computer Architecture, Mathematics, Physics or related technical field or equivalent experienceExcellent Python programming skills, extensive knowledge of at least one DL FrameworkStrong problem solving and analytical skillsAlgorithms and DL fundamentalsWays to stand out from the crowd:Experience in performance measurements and profilingExperience with running large-scale workloads in HPC clustersKnowledge and love for DevOps/MLOps practices for Deep Learning-based product’s development.Solid understanding of Linux environments and containerization technologies such as DockerGPU programming experience (CUDA or OpenCL) is a plus but not required.NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and forward-thinking people in the world working for us. If you're creative and autonomous, we want to hear from you! We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Locations

  • Poland (Remote)

Salary

Estimated Salary Rangemedium confidence

12,000,000 - 24,000,000 INR / yearly

Source: ai estimated

* This is an estimated range based on market data and may vary based on experience and qualifications.

Skills Required

  • Performance Analysisintermediate
  • Performance Optimizationintermediate
  • Deep Learning Trainingintermediate
  • Deep Learning Inferenceintermediate
  • NVIDIA AI Servicesintermediate
  • GPU Architectureintermediate
  • Deep Learning Frameworkintermediate
  • Hardware/Software Stackintermediate
  • Software Engineeringintermediate
  • Deep Learning Modelsintermediate
  • Computer Visionintermediate
  • NLPintermediate
  • LLMsintermediate
  • ASRintermediate
  • TTSintermediate
  • Recommendation Systemsintermediate
  • PyTorchintermediate
  • JAXintermediate
  • TensorFlow 2intermediate
  • DGLintermediate
  • Graph Compilationintermediate
  • Reduced Precision Trainingintermediate
  • Hardware Functionalitiesintermediate
  • Workload Profilingintermediate
  • Workload Optimizationintermediate
  • State-of-the-Art Hardwareintermediate
  • State-of-the-Art Software Platformsintermediate
  • Collaborationintermediate
  • Research Guidanceintermediate
  • Design Improvementintermediate
  • Usability Improvementintermediate
  • Performance Improvementintermediate
  • Best Practices Leadershipintermediate
  • Software Buildingintermediate
  • Software Testingintermediate
  • Software Releasingintermediate
  • Deep Learning Model Implementationintermediate
  • Software Developmentintermediate
  • Python Programmingintermediate
  • DL Framework Knowledgeintermediate
  • Problem Solvingintermediate
  • Analytical Skillsintermediate
  • Algorithmsintermediate
  • Deep Learning Fundamentalsintermediate
  • Performance Measurementsintermediate
  • Profilingintermediate
  • Large-Scale Computingintermediate

Target Your Resume for "Senior Deep Learning Performance Engineer - Training at Scale" , NVIDIA

Get personalized recommendations to optimize your resume specifically for Senior Deep Learning Performance Engineer - Training at Scale. Takes only 15 seconds!

AI-powered keyword optimization
Skills matching & gap analysis
Experience alignment suggestions

Check Your ATS Score for "Senior Deep Learning Performance Engineer - Training at Scale" , NVIDIA

Find out how well your resume matches this job's requirements. Get comprehensive analysis including ATS compatibility, keyword matching, skill gaps, and personalized recommendations.

ATS compatibility check
Keyword optimization analysis
Skill matching & gap identification
Format & readability score

Tags & Categories

Poland

Answer 10 quick questions to check your fit for Senior Deep Learning Performance Engineer - Training at Scale @ NVIDIA.

Quiz Challenge
10 Questions
~2 Minutes
Instant Score

Related Books and Jobs

No related jobs found at the moment.

NVIDIA logo

Senior Deep Learning Performance Engineer - Training at Scale

NVIDIA

Engineering Jobs

Senior Deep Learning Performance Engineer - Training at Scale

full-timePosted: Aug 22, 2025

Job Description

We are looking for senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out of Deep Learning training, inference and NVIDIA AI Services. We are working across all layers of the hardware/software stack, from GPU architecture to Deep Learning Framework, to achieve peak performance. This role offers an opportunity to directly impact the hardware and software roadmap in a fast-growing company that leads the AI revolution. Join the team building software used by the entire world. Work with world class software engineers to implement blazingly fast SOTA deep learning models that help understanding the end-to-end performance of NVIDIA’s DL software and hardware stack. Work on most powerful, enterprise-grade GPU clusters capable of hundreds of Peta FLOPS and on unreleased hardware before anyone in the world.What you’ll be doing:Implement deep learning models from multiple data domains (CV, NLP/LLMs, ASR, TTS, RecSys and others) in multiple DL frameworks (PyT, JAX, TF2, DGL and others)Implement and test new SW features (Graph Compilation, reduced precision training) that use the most recent HW functionalities.Analyze, profile, and optimize deep learning workloads on state-of-the-art hardware and software platforms.Collaborate with researchers and engineers across NVIDIA, providing guidance on improving the design, usability and performance of workloads.Lead best-practices for building, testing, and releasing DL softwareWhat we need to see:5+ years of experience in DL model implementation and SW DevelopmentBSc, MS or PhD degree in Computer Science, Computer Architecture, Mathematics, Physics or related technical field or equivalent experienceExcellent Python programming skills, extensive knowledge of at least one DL FrameworkStrong problem solving and analytical skillsAlgorithms and DL fundamentalsWays to stand out from the crowd:Experience in performance measurements and profilingExperience with running large-scale workloads in HPC clustersKnowledge and love for DevOps/MLOps practices for Deep Learning-based product’s development.Solid understanding of Linux environments and containerization technologies such as DockerGPU programming experience (CUDA or OpenCL) is a plus but not required.NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and forward-thinking people in the world working for us. If you're creative and autonomous, we want to hear from you! We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Locations

  • Poland (Remote)

Salary

Estimated Salary Rangemedium confidence

12,000,000 - 24,000,000 INR / yearly

Source: ai estimated

* This is an estimated range based on market data and may vary based on experience and qualifications.

Skills Required

  • Performance Analysisintermediate
  • Performance Optimizationintermediate
  • Deep Learning Trainingintermediate
  • Deep Learning Inferenceintermediate
  • NVIDIA AI Servicesintermediate
  • GPU Architectureintermediate
  • Deep Learning Frameworkintermediate
  • Hardware/Software Stackintermediate
  • Software Engineeringintermediate
  • Deep Learning Modelsintermediate
  • Computer Visionintermediate
  • NLPintermediate
  • LLMsintermediate
  • ASRintermediate
  • TTSintermediate
  • Recommendation Systemsintermediate
  • PyTorchintermediate
  • JAXintermediate
  • TensorFlow 2intermediate
  • DGLintermediate
  • Graph Compilationintermediate
  • Reduced Precision Trainingintermediate
  • Hardware Functionalitiesintermediate
  • Workload Profilingintermediate
  • Workload Optimizationintermediate
  • State-of-the-Art Hardwareintermediate
  • State-of-the-Art Software Platformsintermediate
  • Collaborationintermediate
  • Research Guidanceintermediate
  • Design Improvementintermediate
  • Usability Improvementintermediate
  • Performance Improvementintermediate
  • Best Practices Leadershipintermediate
  • Software Buildingintermediate
  • Software Testingintermediate
  • Software Releasingintermediate
  • Deep Learning Model Implementationintermediate
  • Software Developmentintermediate
  • Python Programmingintermediate
  • DL Framework Knowledgeintermediate
  • Problem Solvingintermediate
  • Analytical Skillsintermediate
  • Algorithmsintermediate
  • Deep Learning Fundamentalsintermediate
  • Performance Measurementsintermediate
  • Profilingintermediate
  • Large-Scale Computingintermediate

Target Your Resume for "Senior Deep Learning Performance Engineer - Training at Scale" , NVIDIA

Get personalized recommendations to optimize your resume specifically for Senior Deep Learning Performance Engineer - Training at Scale. Takes only 15 seconds!

AI-powered keyword optimization
Skills matching & gap analysis
Experience alignment suggestions

Check Your ATS Score for "Senior Deep Learning Performance Engineer - Training at Scale" , NVIDIA

Find out how well your resume matches this job's requirements. Get comprehensive analysis including ATS compatibility, keyword matching, skill gaps, and personalized recommendations.

ATS compatibility check
Keyword optimization analysis
Skill matching & gap identification
Format & readability score

Tags & Categories

Poland

Answer 10 quick questions to check your fit for Senior Deep Learning Performance Engineer - Training at Scale @ NVIDIA.

Quiz Challenge
10 Questions
~2 Minutes
Instant Score

Related Books and Jobs

No related jobs found at the moment.