Generative AI Inference Engineer
Posted 14 hours ago
Job Description
Generative AI Inference Engineer
<Remote>
About the role:
We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.
Responsibilities:
- Lead efforts to drive the design, development of customer-facing multi modal ML inference systems.
- Work with the Platform and Inference teams on building inference systems for the next generation of models, where you will work on areas such as optimization, model tuning and deployment.
- Partner with leading cloud providers to deliver hosted Stability AI inference solutions.
- Be a strategic thought partner for leaders across the organization on driving business impact through machine learning
- Be part of the team to bring new Stability models and pipelines into existence
- Prototype and productionize inference platform improvements and new features
Qualifications:
- 7+ years working on productionizing machine learning systems, including inference pipeline development
- Expert level knowledge on writing and running python services at scale
- 5+ years working on python scientific stack, pyTorch and at least one high-performance inference framework (e.g. Triton and TensorRT)
- Deep understanding of Diffusion Architecture
- Experience profiling and optimizing deep neural networks on Nvidia GPUs, using profiling tools such as NVIDIA Nsight
- Experience with python-based image manipulation/encoding/decoding frameworks, such as OpenCV
- Experience deploying to cloud orchestration systems such as Kubernetes and cloud providers such as AWS, GCP, and Azure
- Experience with Docker
- Ability to rapidly prototype solutions and iterate on them with tight product deadlines
- Strong communication, collaboration, and documentation skills
- Experience with the open-source ML ecosystem (HuggingFace, W&B, etc.)
Equal Employment Opportunity:
We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or other legally protected statuses.

Stability AI
1 job posted
About the job
Similar Jobs
21d
Applied AI Engineer – Generative AI
Kodiak
$150K - $250KSan Francisco, CAApplied AI Engineer – Generative AI
Kodiak
$150K - $250KSan Francisco, CA21d2d
Senior Generative AI Engineer
Datadog
Paris, FranceSenior Generative AI Engineer
Datadog
Paris, France2d3d
Sr. SW Engineer- Generative AI
Visa
$111K - $172KAustin, TXSr. SW Engineer- Generative AI
Visa
$111K - $172KAustin, TX3d25d
Senior Machine Learning Engineer (Multimodal Generative AI)
Code and Theory
Bengaluru, Karnataka, IndiaSenior Machine Learning Engineer (Multimodal Generative AI)
Code and Theory
Bengaluru, Karnataka, India25d24d
Research Engineer, Multimodal Generative AI (Image/Video)
DeepMind
$166KKirklandSeattleResearch Engineer, Multimodal Generative AI (Image/Video)
DeepMind
$166KKirklandSeattle24d25d
Cloud Hardware Dev Engineer (AWS Generative AI & ML Servers), AWS Generative AI & ML Servers
Amazon
US, WACloud Hardware Dev Engineer (AWS Generative AI & ML Servers), AWS Generative AI & ML Servers
Amazon
US, WA25d25d
Quality Analyst — Generative AI Data Quality
Welocalize
IndiaQuality Analyst — Generative AI Data Quality
Welocalize
India25d12d
Quality Analyst — Generative AI Data Quality
Welocalize
Gurgaon, IndiaQuality Analyst — Generative AI Data Quality
Welocalize
Gurgaon, India12d5d
Applied Scientist, Generative AI/ML
Adobe
$152K - $265KSan Jose, CAApplied Scientist, Generative AI/ML
Adobe
$152K - $265KSan Jose, CA5d25d
Sr. Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering
Amazon
US, TXSr. Systems Development Engineer (AWS Generative AI & ML Servers), AWS HW Engineering
Amazon
US, TX25d
Looking for something different?
Browse all AI jobsFree AI job alerts
Get the latest AI jobs delivered to your inbox every week. Free, no spam.