Golang Developer with DevOps/LLM Experience - Remote / Telecommute

Remote Full-time
Job Description: Required Skills: • Proficiency in Golang for building scalable and performant backend services. • Deep experience building services in modern cloud environments on distributed systems (i.e., containerization (Kubernetes, Docker), infrastructure as code, CI/CD pipelines, APIs, authentication and authorization, data storage, deployment, logging, monitoring, alerting, etc.) • Experience working with Large Language Models (LLMs), particularly hosting them to run inference. • Strong verbal and written communication skills. • Candidates job will involve communicating with local and remote colleagues about technical subjects and writing detailed documentation. • Experience with building or using benchmarking tools for evaluating LLM inference for various models, engine, and GPU combinations. • Familiarity with various LLM performance metrics such as prefill throughput, decode throughput, TPOT, and TTFT. • Experience with one or more inference engines: e.g., vLLM, SGLang, and Modular Max. • Familiarity with one or more distributed inference serving frameworks: e.g., llm-d, NVIDIA Dynamo, and Ray Serve etc. • Experience with client and NVIDIA GPUs, using software like CUDA, ROCm, AITER, NCCL, Client, etc. • Knowledge of distributed inference optimization techniques - tensor/data parallelism, KV cache optimizations, smart routing etc. • Develop and maintain an inference platform for serving large language models optimized for the various GPU platforms they will be run on. • Work on complex AI and cloud engineering projects through the entire product development lifecycle (PDLC) - ideation, product definition, experimentation, prototyping, development, testing, release, and operations. • Build tooling and observability to monitor system health, and build auto tuning capabilities. • Build benchmarking frameworks to test model serving performance to guide system and infrastructure tuning efforts. • Build native cross platform inference support across NVIDIA and client GPUs for a variety of model architectures. • Contribute to open source inference engines to make them perform better on DigitalOcean cloud. Apply tot his job
Apply Now

Similar Opportunities

Experienced Registered Behavior Technician for In-Home ABA Therapy - Atlanta, GA

Remote Full-time

Immediate Hiring: Experienced Registered Behavioral Technician (RBT) for Clinic-Based ABA Therapy Services

Remote Full-time

Experienced Registered Behavioral Technician (RBT) - ABA Therapy for Children with Autism Spectrum Disorder

Remote Full-time

Experienced Registered Nurse - Telehealth: Providing Remote Care Coordination and Patient Support

Remote Full-time

Experienced Substitute Teacher for Riverside County Schools - Join Scoot Education's Innovative Team

Remote Full-time

Experienced Substitute Teacher for San Bernardino County - Flexible Schedules & Competitive Pay

Remote Full-time

Experienced School Year Instructional Coach for High-Dosage Tutoring Programs in Edgewater Park, NJ

Remote Full-time

Experienced School Year Tutor for K-8 Students in Math and Literacy - Mickleton, NJ

Remote Full-time

Experienced Secondary Social Studies Teacher for Kansas - Flexible Hybrid Remote Arrangement

Remote Full-time

USPS Office Helper

Remote Full-time

Product Manager (HR & Healthcare IT Applications) -- GHODC5692915

Remote Full-time

Experienced Full Stack Data Entry Specialist – Remote Content Data Management at Blithequark (Part/Full Time, $72,000/Year)

Remote Full-time

**Experienced Remote Data Entry Specialist – Flexible Work Arrangement for blithequark**

Remote Full-time

**Experienced HR Business Partner – Strategic Talent Management and Organizational Development for Disney's Streaming Services**

Remote Full-time

**Job Title:** Experienced Bilingual Customer Service Representative - Spanish/English Support Specialist for Remote Work at blithequark

Remote Full-time

Communications Specialist and School Board Liaison in Minnesota in Monticello Public School District (job Id: 1690600270)

Remote Full-time

Software Engineering Manager II – Mobile Experience

Remote Full-time

**Job Title:** Experienced Data Entry and Claims Specialist – Customer-Focused Claims Processing and Record Management

Remote Full-time

Experienced Customer Support Representative – Full-Time Remote Live Chat Agent for Innovative E-Commerce Leader

Remote Full-time

Financial Advisor

Remote Full-time
← Back to Home