Stability AI Jobs

Generative AI Inference Engineer

Stability AI

Generative AI Inference Engineer

Reposted 5 Hours Ago

Remote

Hiring Remotely in United States

Expert/Leader

Remote

Hiring Remotely in United States

Expert/Leader

Lead the design and development of ML inference systems, focusing on generative AI models and optimization techniques for production environments.

The summary above was generated by AI

Generative AI Inference Engineer

About the role:

We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.

Responsibilities:

Lead efforts to drive the design, development of customer-facing multi modal ML inference systems.
Work with the Platform and Inference teams on building inference systems for the next generation of models, where you will work on areas such as optimization, model tuning and deployment.
Partner with leading cloud providers to deliver hosted Stability AI inference solutions.
Be a strategic thought partner for leaders across the organization on driving business impact through machine learning
Be part of the team to bring new Stability models and pipelines into existence
Prototype and productionize inference platform improvements and new features

Qualifications:

7+ years working on productionizing machine learning systems, including inference pipeline development
Expert level knowledge on writing and running python services at scale
5+ years working on python scientific stack, pyTorch and at least one high-performance inference framework (e.g. Triton and TensorRT)
Deep understanding of Diffusion Architecture
Experience profiling and optimizing deep neural networks on Nvidia GPUs, using profiling tools such as NVIDIA Nsight
Experience with python-based image manipulation/encoding/decoding frameworks, such as OpenCV
Experience deploying to cloud orchestration systems such as Kubernetes and cloud providers such as AWS, GCP, and Azure
Experience with Docker
Ability to rapidly prototype solutions and iterate on them with tight product deadlines
Strong communication, collaboration, and documentation skills
Experience with the open-source ML ecosystem (HuggingFace, W&B, etc.)

Equal Employment Opportunity:

We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or other legally protected statuses.

Similar Jobs

Datadog

Account Executive

5 Hours Ago

Easy Apply

Remote or Hybrid

Easy Apply

135K-150K Annually

Senior level

135K-150K Annually

Senior level

Artificial Intelligence • Cloud • Security • Software • Cybersecurity

The Strategic Account Executive closes new business with major clients, focusing on cloud solutions and building customer relationships.

Top Skills: It InfrastructureSaaS

Airwallex

Senior Payroll Partner

5 Hours Ago

Remote or Hybrid

Senior level

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI

The Senior Payroll Partner will manage global payroll operations, ensuring compliance, delivering insights, driving system improvements, and collaborating with HR and legal teams on payroll-related issues.

Top Skills: Compensation Analysis ToolsHrisPayroll Vendor Management Software

Airwallex

Director GTM Engineering

5 Hours Ago

Remote or Hybrid

Expert/Leader

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI

The Director GTM Engineering will own the GTM architecture, lead a high-caliber engineering pod, and streamline integrations across marketing and revenue systems, improving automation and data governance.

Top Skills: AIAnalyticsData WarehouseDemandbaseMarketoSalesforce

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus