NVIDIA Logo

NVIDIA

Senior Deep Learning Performance Engineer - Training at Scale

Job Posted 14 Days Ago Posted 14 Days Ago
Be an Early Applicant
Remote
5 Locations
Senior level
Remote
5 Locations
Senior level
We seek a Senior Deep Learning Performance Engineer to optimize DL training and inference performance, implementing models across various frameworks and collaborating with teams.
The summary above was generated by AI

We are looking for senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out of Deep Learning training, inference and NVIDIA AI Services. We are working across all layers of the hardware/software stack, from GPU architecture to Deep Learning Framework, to achieve peak performance. This role offers an opportunity to directly impact the hardware and software roadmap in a fast-growing company that leads the AI revolution.  Join the team building software used by the entire world. Work with world class software engineers to implement blazingly fast SOTA deep learning models that help understanding the end-to-end performance of NVIDIA’s DL software and hardware stack. Work on most powerful, enterprise-grade GPU clusters capable of hundreds of Peta FLOPS and on unreleased hardware before anyone in the world.

What you’ll be doing:

  • Implement deep learning models from multiple data domains (CV, NLP/LLMs, ASR, TTS, RecSys and others) in multiple DL frameworks (PyT, JAX, TF2, DGL and others)

  • Implement and test new SW features (Graph Compilation, reduced precision training) that use the most recent HW functionalities.

  • Analyze, profile, and optimize deep learning workloads on state-of-the-art hardware and software platforms.

  • Collaborate with researchers and engineers across NVIDIA, providing guidance on improving the design, usability and performance of workloads.

  • Lead best-practices for building, testing, and releasing DL software

What we need to see:

  • 5+ years of experience in DL model implementation and SW Development

  • BSc, MS or PhD degree in Computer Science, Computer Architecture, Mathematics, Physics or related technical field or equivalent experience

  • Excellent Python programming skills, extensive knowledge of at least one DL Framework

  • Strong problem solving and analytical skills

  • Algorithms and DL fundamentals

Ways to stand out from the crowd:

  • Experience in performance measurements and profiling

  • Experience with running large-scale workloads in HPC clusters

  • Knowledge and love for DevOps/MLOps practices for Deep Learning-based product’s development.

  • Solid understanding of Linux environments and containerization technologies such as Docker

  • GPU programming experience (CUDA or OpenCL) is a plus but not required.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and forward-thinking people in the world working for us. If you're creative and autonomous, we want to hear from you! We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Top Skills

Cuda
Dgl
Docker
Jax
Opencl
Python
PyTorch
TensorFlow

NVIDIA London, England Office

13th Floor One Angel Court, London, United Kingdom, EC2R 7HJ

Similar Jobs

9 Hours Ago
Easy Apply
Remote
London, England, GBR
Easy Apply
Junior
Junior
Social Impact • Software
As a Solutions Engineer, you'll drive technical sales processes, support customer needs, communicate effectively, and collaborate within a team to provide solutions that enhance digital accessibility.
Top Skills: AgileAndroidCi/CdCSSCucumberCypressHTMLiOSJavaScriptWebdriver Io
Yesterday
Remote
Hybrid
Belfast, County Antrim, Northern Ireland, GBR
Mid level
Mid level
Artificial Intelligence • Cloud • Information Technology • Sales • Security • Software • Cybersecurity
As a DevOps Engineer II, you'll build and maintain services, collaborate with teams to manage security solutions, and enhance the data platform.
Top Skills: AWSDockerKafkaKubernetesPythonSparkTerraformTimescaledb
2 Days Ago
Remote
United Kingdom
Mid level
Mid level
Software • Analytics • Hospitality
The Regional Solutions Engineer provides pre-sales support for IDeaS solutions, working closely with sales and account management teams while utilizing technical knowledge to optimize sales opportunities and customer satisfaction.
Top Skills: Business/Data IntelligenceData ManagementSaaSSystem Integration

What you need to know about the London Tech Scene

London isn't just a hub for established businesses; it's also a nursery for innovation. Boasting one of the most recognized fintech ecosystems in Europe, attracting billions in investments each year, London's success has made it a go-to destination for startups looking to make their mark. Top U.K. companies like Hoptin, Moneybox and Marshmallow have already made the city their base — yet fintech is just the beginning. From healthtech to renewable energy to cybersecurity and beyond, the city's startups are breaking new ground across a range of industries.
By clicking Apply you agree to share your profile information with the hiring company.

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account