Senior Deep Learning Performance Engineer - Training at Scale

Remote Full-time
We are looking for senior engineers who are mindful of performance analysis and optimization to help us squeeze every last clock cycle out of Deep Learning training, inference and NVIDIA AI Services. We are working across all layers of the hardware/software stack, from GPU architecture to Deep Learning Framework, to achieve peak performance. This role offers an opportunity to directly impact the hardware and software roadmap in a fast-growing company that leads the AI revolution. Join the team building software used by the entire world. Work with world class software engineers to implement blazingly fast SOTA deep learning models that help understanding the end-to-end performance of NVIDIA’s DL software and hardware stack. Work on most powerful, enterprise-grade GPU clusters capable of hundreds of Peta FLOPS and on unreleased hardware before anyone in the world. What you’ll be doing: • Implement deep learning models from multiple data domains (CV, NLP/LLMs, ASR, TTS, RecSys and others) in multiple DL frameworks (PyT, JAX, TF2, DGL and others) • Implement and test new SW features (Graph Compilation, reduced precision training) that use the most recent HW functionalities. • Analyze, profile, and optimize deep learning workloads on state-of-the-art hardware and software platforms. • Collaborate with researchers and engineers across NVIDIA, providing guidance on improving the design, usability and performance of workloads. • Lead best-practices for building, testing, and releasing DL software What we need to see: • 5+ years of experience in DL model implementation and SW Development • BSc, MS or PhD degree in Computer Science, Computer Architecture, Mathematics, Physics or related technical field or equivalent experience • Excellent Python programming skills, extensive knowledge of at least one DL Framework • Strong problem solving and analytical skills • Algorithms and DL fundamentals Ways to stand out from the crowd: • Experience in performance measurements and profiling • Experience with running large-scale workloads in HPC clusters • Knowledge and love for DevOps/MLOps practices for Deep Learning-based product’s development. • Solid understanding of Linux environments and containerization technologies such as Docker • GPU programming experience (CUDA or OpenCL) is a plus but not required. NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most brilliant and forward-thinking people in the world working for us. If you're creative and autonomous, we want to hear from you! We are an equal opportunity employer and value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status. Apply tot his job
Apply Now

Similar Opportunities

Experienced Registered Behavior Technician for In-Home ABA Therapy - Atlanta, GA

Remote Full-time

Immediate Hiring: Experienced Registered Behavioral Technician (RBT) for Clinic-Based ABA Therapy Services

Remote Full-time

Experienced Registered Behavioral Technician (RBT) - ABA Therapy for Children with Autism Spectrum Disorder

Remote Full-time

Experienced Registered Nurse - Telehealth: Providing Remote Care Coordination and Patient Support

Remote Full-time

Experienced Substitute Teacher for Riverside County Schools - Join Scoot Education's Innovative Team

Remote Full-time

Experienced Substitute Teacher for San Bernardino County - Flexible Schedules & Competitive Pay

Remote Full-time

Experienced School Year Instructional Coach for High-Dosage Tutoring Programs in Edgewater Park, NJ

Remote Full-time

Experienced School Year Tutor for K-8 Students in Math and Literacy - Mickleton, NJ

Remote Full-time

Experienced Secondary Social Studies Teacher for Kansas - Flexible Hybrid Remote Arrangement

Remote Full-time

USPS Office Helper

Remote Full-time

Experienced Remote Live Chat Support Specialist – Customer Service and Technical Troubleshooting Expert

Remote Full-time

Mortgage Loan Officer ( Licensed ) - In Office or Remote

Remote Full-time

Math/Science Tutor – Teaching Certificate

Remote Full-time

Experienced Remote Data Entry Specialist – Accurate Data Management and Entry for arenaflex

Remote Full-time

**Experienced Customer Service Representative – Remote Work Opportunities with blithequark**

Remote Full-time

Analyst/ Associate- Equity Research, Biotech

Remote Full-time

Experienced Patient Billing Customer Service Support Representative – Remote Work Opportunity with Flexible Shift Schedule and Professional Growth at blithequark

Remote Full-time

Experienced Remote Amazon Data Entry Specialist – Pharmaceutical Industry Opportunity with Flexible Hours and Professional Growth

Remote Full-time

Entry Level Remote Chat Support Agent – Customer Service Representative for blithequark – $35/hr – Work from Home Opportunity

Remote Full-time

Entry-Level Live Chat Support Specialist – No Experience Required, Earn $25-$35/Hour, and Join the blithequark Team for a Fulfilling Remote Career

Remote Full-time
← Back to Home