Link copied to clipboard!
Back to Jobs
Generative AI Inference Engineer at Stability AI
Stability AI
Austin, TX
Information Technology
Posted 0 days ago
Job Description
Generative AI Inference Engineer<Remote>About the role:We are seeking passionate Machine Learning Engineers to join our Inference team focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools likeComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.Responsibilities: Lead efforts to drive the design development of customer-facing multi modal ML inference systems.Work with the Platform and Inference teams on building inference systems for the next generation of models where you will work on areas such as optimization model tuning and deployment.Partner with leading cloud providers to deliver hosted Stability AI inference solutions.Be a strategic thought partner for leaders across the organization on driving business impact through machine learningBe part of the team to bring new Stability models and pipelines into existencePrototype and productionize inference platform improvements and new featuresQualifications:7 years working on productionizing machine learning systems including inference pipeline developmentExpert level knowledge on writing and running python services at scale5 years working on python scientific stack pyTorch and at least one high-performance inference framework (e.g. Triton and TensorRT)Deep understanding of Diffusion ArchitectureExperience profiling and optimizing deep neural networks on Nvidia GPUs using profiling tools such as NVIDIA NsightExperience with python-based image manipulation/encoding/decoding frameworks such as OpenCVExperience deploying to cloud orchestration systems such as Kubernetes and cloud providers such as AWS GCP and AzureExperience with DockerAbility to rapidly prototype solutions and iterate on them with tight product deadlinesStrong communication collaboration and documentation skillsExperience with the open-source ML ecosystem (HuggingFace W&B etc.)Equal Employment Opportunity:We are an equal opportunity employer and do not discriminate on the basis of race religion national origin gender sexual orientation age veteran status disability or other legally protected statuses. Key Skills ASP.NET,Health Education,Fashion Designing,Fiber,Investigation Employment Type : Full Time Experience: years Vacancy: 1
Resume Suggestions
Highlight relevant experience and skills that match the job requirements to demonstrate your qualifications.
Quantify your achievements with specific metrics and results whenever possible to show impact.
Emphasize your proficiency in relevant technologies and tools mentioned in the job description.
Showcase your communication and collaboration skills through examples of successful projects and teamwork.