[Remote] Generative AI Inference Engineer

Remote Full-time
Note: The job is a remote job and is open to candidates in USA. Stability AI is seeking passionate Machine Learning Engineers to join their Inference team, focusing on the creative applications of generative AI models. The role involves leading the design and development of customer-facing multi-modal ML inference systems and collaborating with various teams to optimize and deploy cutting-edge models. Responsibilities • Lead efforts to drive the design, development of customer-facing multi modal ML inference systems • Work with the Platform and Inference teams on building inference systems for the next generation of models, where you will work on areas such as optimization, model tuning and deployment • Partner with leading cloud providers to deliver hosted Stability AI inference solutions • Be a strategic thought partner for leaders across the organization on driving business impact through machine learning • Be part of the team to bring new Stability models and pipelines into existence • Prototype and productionize inference platform improvements and new features Skills • 7+ years working on productionizing machine learning systems, including inference pipeline development • Expert level knowledge on writing and running python services at scale • 5+ years working on python scientific stack, pyTorch and at least one high-performance inference framework (e.g. Triton and TensorRT) • Deep understanding of Diffusion Architecture • Experience profiling and optimizing deep neural networks on Nvidia GPUs, using profiling tools such as NVIDIA Nsight • Experience with python-based image manipulation/encoding/decoding frameworks, such as OpenCV • Experience deploying to cloud orchestration systems such as Kubernetes and cloud providers such as AWS, GCP, and Azure • Experience with Docker • Ability to rapidly prototype solutions and iterate on them with tight product deadlines • Strong communication, collaboration, and documentation skills • Experience with the open-source ML ecosystem (HuggingFace, W&B, etc.) Company Overview • Stability AI is an artificial intelligence company focused on developing open-source generative AI models. It was founded in 2019, and is headquartered in London, England, GBR, with a workforce of 51-200 employees. Its website is Apply tot his job

Apply tot his job

Apply To this Job
Apply Now

Similar Opportunities

Experienced Registered Behavior Technician for In-Home ABA Therapy - Atlanta, GA

Remote Full-time

Immediate Hiring: Experienced Registered Behavioral Technician (RBT) for Clinic-Based ABA Therapy Services

Remote Full-time

Experienced Registered Behavioral Technician (RBT) - ABA Therapy for Children with Autism Spectrum Disorder

Remote Full-time

Experienced Registered Nurse - Telehealth: Providing Remote Care Coordination and Patient Support

Remote Full-time

Experienced Substitute Teacher for Riverside County Schools - Join Scoot Education's Innovative Team

Remote Full-time

Experienced Substitute Teacher for San Bernardino County - Flexible Schedules & Competitive Pay

Remote Full-time

Experienced School Year Instructional Coach for High-Dosage Tutoring Programs in Edgewater Park, NJ

Remote Full-time

Experienced School Year Tutor for K-8 Students in Math and Literacy - Mickleton, NJ

Remote Full-time

Experienced Secondary Social Studies Teacher for Kansas - Flexible Hybrid Remote Arrangement

Remote Full-time

USPS Office Helper

Remote Full-time

Customer Service Representative, UT (5/11 Start)

Remote Full-time

Catering Coordinator - Cleveland Museum of Art

Remote Full-time

VIP Guest Services Manager, Disney Cruise Line

Remote Full-time

**Experienced Customer Service Representative - Healthcare Call Center (Remote)**

Remote Full-time

Software Engineer

Remote Full-time

Director - Diversity, Equity, and Inclusion

Remote Full-time

Construction Project Coordinator / Estimator VA

Remote Full-time

Clinical Reviewer – SCA (Remote – RN/LPN), Anywhere

Remote Full-time

Experienced Live Chat Support Agent – Delivering Exceptional Customer Experiences in a Dynamic Remote Environment at arenaflex

Remote Full-time

Sports Analytics Data Scientist

Remote Full-time
← Back to Home