Stability AI Logo

Stability AI

Generative AI Inference Engineer

Reposted 13 Days Ago
Remote
Hiring Remotely in United States
Expert/Leader
Remote
Hiring Remotely in United States
Expert/Leader
Lead the design and development of ML inference systems, focusing on generative AI models and optimization techniques for production environments.
The summary above was generated by AI

Generative AI Inference Engineer

<Remote> 

About the role: 

We are seeking passionate Machine Learning Engineers to join our Inference team, focusing on the creative applications of generative AI models. The ideal candidate will have substantial experience developing and running inference for multi-modal models. A deep understanding of diffusion model architectures and familiarity with workflow tools like ComfyUI are a big plus. You will be expected to leverage and push the boundaries of state-of-the-art inference optimization techniques for multi-modal generative models. This role offers the opportunity to work alongside top researchers and engineers, utilizing cutting-edge high-performance computing resources to make a significant impact in the rapidly evolving field of generative AI.

Responsibilities:  

  • Lead efforts to drive the design, development of customer-facing multi modal ML inference systems.
  • Work with the Platform and Inference teams on building inference systems for the next generation of models, where you will work on areas such as optimization, model tuning and deployment.
  • Partner with leading cloud providers to deliver hosted Stability AI inference solutions.
  • Be a strategic thought partner for leaders across the organization on driving business impact through machine learning
  • Be part of the team to bring new Stability models and pipelines into existence
  • Prototype and productionize inference platform improvements and new features 

Qualifications:

  • 7+ years working on productionizing machine learning systems, including inference pipeline development
  • Expert level knowledge on writing and running python services at scale
  • 5+ years working on python scientific stack, pyTorch and at least one high-performance inference framework (e.g. Triton and TensorRT)
  • Deep understanding of Diffusion Architecture
  • Experience profiling and optimizing deep neural networks on Nvidia GPUs, using profiling tools such as NVIDIA Nsight
  • Experience with python-based image manipulation/encoding/decoding frameworks, such as OpenCV
  • Experience deploying to cloud orchestration systems such as Kubernetes and cloud providers such as AWS, GCP, and Azure
  • Experience with Docker
  • Ability to rapidly prototype solutions and iterate on them with tight product deadlines
  • Strong communication, collaboration, and documentation skills
  • Experience with the open-source ML ecosystem (HuggingFace, W&B, etc.)

Equal Employment Opportunity:

We are an equal opportunity employer and do not discriminate on the basis of race, religion, national origin, gender, sexual orientation, age, veteran status, disability or other legally protected statuses.

Similar Jobs

2 Hours Ago
Remote or Hybrid
Pennsylvania, USA
26K-28K Hourly
Senior level
26K-28K Hourly
Senior level
Digital Media • Information Technology • News + Entertainment
Inbound telesales role selling Comcast Business services to small and mid-size businesses. Manage CRM pipeline, follow consultative sales process, meet quotas, promote bundled connectivity and voice solutions, handle longer sales cycles, provide product feedback, and deliver excellent customer experience. Must work variable schedules including nights and weekends.
Top Skills: Advanced VoiceCRMExcelFiber NetworkingMobile Sales
2 Hours Ago
Remote or Hybrid
Arizona, USA
Senior level
Senior level
Digital Media • Information Technology • News + Entertainment
Develop and execute territory strategy to acquire and manage mid-market and enterprise multi-location customers. Prospect, present, negotiate, and close complex Comcast Business solutions; build partnerships and maintain customer relationships to drive retention and revenue. Coordinate internal teams to ensure service delivery and maintain accurate sales records.
Top Skills: Business ContinuityCustomer Premise EquipmentCybersecurityDisaster RecoveryEthernetLanLayer 1Layer 2Layer 3Man TechnologiesManaged Security SolutionsNetwork SecuritySdwanVoipVpnWanWdm
2 Hours Ago
Remote or Hybrid
Pennsylvania, USA
91K-212K Annually
Expert/Leader
91K-212K Annually
Expert/Leader
Digital Media • Information Technology • News + Entertainment
Lead enterprise customer retention strategy and operations for Business Customer Loyalty. Own retention KPIs (save rate, churn reduction, revenue preservation), manage multi-layer call center teams, optimize offers and processes, drive coaching and quality programs, partner cross-functionally to resolve complex cases, and support budgeting and staffing to improve long-term customer retention.

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account