AssemblyAI Logo

AssemblyAI

Senior Research Engineer

Reposted 20 Days Ago
Be an Early Applicant
Easy Apply
Remote
Hiring Remotely in USA
240K-275K Annually
Senior level
Easy Apply
Remote
Hiring Remotely in USA
240K-275K Annually
Senior level
This role involves optimizing large-scale distributed training and inference systems, implementing deep learning optimizations, and collaborating across teams to enhance AI models.
The summary above was generated by AI
About AssemblyAI

AssemblyAI builds the best-in-class Speech AI models powering the next generation of voice applications. Our models serve 600M+ inference calls monthly, process 1M+ hours of audio daily, and power 2 billion+ end-user experiences—from voice agents and meeting assistants to contact centers and medical scribes. Companies like Zoom, Granola, Fireflies, Cluely, and Calabrio rely on AssemblyAI to ship production-ready voice AI.

We're at an inflection point in Speech AI. We released Universal-Streaming in mid-2025, and it has quickly earned its place as the model offering the best accuracy-latency-cost tradeoff on the market. The adoption has been significant: we now process ~1.5M streaming hours per week, with 25x usage growth in the last six months alone. Our research team drives these advances and ships with relentless velocity. Since releasing Universal-Streaming, we've already launched keyterms prompting feature and multilingual support—with more significant improvements on the roadmap.

We've raised $115M+ from Accel, Insight Partners, Y Combinator's AI Fund, Patrick and John Collison, Nat Friedman, and Daniel Gross. We're a remote team building one of the next great AI companies—and we're looking for researchers who will shape its future.

About the Role

We are seeking a highly skilled Senior Research Engineer to collaborate closely with both Research and Engineering teams. The role involves diagnosing and resolving bottlenecks across large-scale distributed training, data processing, and inference systems, while also driving optimizations for existing high-performance pipelines.

The ideal candidate possesses a deep understanding of modern deep learning systems, combined with strong engineering expertise in areas such as layer-level optimization, large-scale distributed training, streaming, low-latency and asynchronous inference, inference compilers, and advanced parallelization techniques.

This is a cross-functional role requiring strong technical rigor, attention to detail, intellectual curiosity, and excellent communication skills. The position is embedded within the Research team and is responsible for developing and refining the technical foundation that enables cutting-edge research and translates its outcomes into production, bridging research and production engineering.

What You'll Do
  • Investigate and mitigate performance bottlenecks in large-scale distributed training and inference systems.
  • Develop and implement both low-level (operator/kernel) and high-level (system/architecture) optimization strategies.
  • Translate research models and prototypes into highly optimized, production-ready inference systems.
  • Explore and integrate inference compilers such as TensorRT, ONNX Runtime, AWS Neuron and Inferentia, or similar technologies.
  • Design, test, and deploy scalable solutions for parallel and distributed workloads on heterogeneous hardware.
  • Facilitate knowledge transfer and bidirectional support between Research and Engineering teams, ensuring alignment of priorities and solutions.
What You'll Need
  • Strong expertise in the Python ecosystem and major ML frameworks (PyTorch, JAX).
  • Experience with lower-level programming (C++ or Rust preferred).
  • Deep understanding of GPU acceleration (CUDA, profiling, kernel-level optimization); TPU experience is a strong plus.
  • Proven ability to accelerate deep learning workloads using compiler frameworks, graph optimizations, and parallelization strategies.
  • Solid understanding of the deep learning lifecycle: model design, large-scale training, data processing pipelines, and inference deployment.
  • Strong debugging, profiling, and optimization skills in large-scale distributed environments.
  • Excellent communication and collaboration skills, with the ability to clearly prioritize and articulate impact-driven technical solutions.

Pay Transparency:

AssemblyAI strives to recruit and retain exceptional talent from diverse backgrounds while ensuring pay equity across our team. Our salary ranges are set to be competitive for our size, stage, and industry, and reflect just one component of the full compensation, benefits, and rewards we offer.

Salary determinations consider a variety of factors, including relevant experience, technical depth, skills demonstrated during the interview process, and maintaining internal equity with peers on the team. The range shared below represents a general expectation for the posted position. However, we are open to considering candidates who may fall above or below the outlined experience level—in those cases, we will communicate any adjustments to the expected salary range.

The range provided applies to candidates located in the United States. For candidates outside of the U.S., compensation ranges may differ; any adjustments will be communicated throughout the interview process.

Salary range: $210,000 - $309,000 

The expected base compensation for this role is listed above. Our total compensation package includes competitive equity grants, 100% employer-paid benefits, and the flexibility of being fully remote. 401k match up to 4% for US-based full time team members.

Working at AssemblyAI

We are a small but mighty group of startup veterans and experienced AI researchers with over 20 years of expertise in Machine Learning, Speech Recognition, and NLP. As a fully remote team, we’re looking for people to join our team who are ambitious, curious, and lead with integrity. We’re still in the early days of AI and of AssemblyAI’s journey, and are looking for teammates who won’t just fit in, but will help us define and build our company culture. 

We’re committed to creating a space where our employees can bring their full selves to work and have equal opportunity to succeed. No matter your race, gender identity or expression, sexual orientation, religion, origin, ability, age, veteran status, if joining this mission speaks to you, we encourage you to apply!

Using AI to Interview:

If you’re selected for an interview, please review this resource to better understand how AssemblyAI approaches the use of AI in our interview process.

GDPR privacy notice:

Candidates from the EU should review this job applicant privacy notice before applying. 

Keep Exploring AssemblyAI:Keep Exploring AssemblyAI:

Check us out on YouTube!

Learn more about AI models for speech recognition

Speech-to-Text | Speech Understanding | LLM Gateway | Try the Playground

Our $50M Series C fundraise

Top Skills

Aws Neuron
C++
Cuda
Inferentia
Jax
Onnx Runtime
Python
PyTorch
Rust
Tensorrt

Similar Jobs

8 Days Ago
Remote
United States
151K-239K Annually
Senior level
151K-239K Annually
Senior level
Artificial Intelligence • Software
Lead the development of a differentiable CFD framework for optimizing chemical reactor designs, involving scalable GPU engineering and numerical optimization. Collaborates with AI developers to validate and document simulations.
Top Skills: CudaDockerGitGoogle Cloud BatchJaxK8SMpiPyTorchSlurm
9 Hours Ago
In-Office or Remote
Santa Clara, CA, USA
224K-431K Annually
Senior level
224K-431K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
Design and post-train generative AI foundation models, contribute to development of training infrastructure, collaborate across teams, and mentor junior engineers. Expertise in large scale training and familiarity with transformer architectures required.
Top Skills: CpDdpFsdpJaxPythonPyTorchRaySparkTpZero
12 Days Ago
Easy Apply
Remote
United States
Easy Apply
136K-170K Annually
Senior level
136K-170K Annually
Senior level
Aerospace • Big Data • Greentech • Hardware • Social Impact
As a Senior Operations Research Engineer, you'll architect software solutions, improve satellite operations, and mentor team members in mission planning for satellites.
Top Skills: C++DockerJenkinsJIRALinuxPython

What you need to know about the Colorado Tech Scene

With a business-friendly climate and research universities like CU Boulder and Colorado State, Colorado has made a name for itself as a startup ecosystem. The state boasts a skilled workforce and high quality of life thanks to its affordable housing, vibrant cultural scene and unparalleled opportunities for outdoor recreation. Colorado is also home to the National Renewable Energy Laboratory, helping cement its status as a hub for renewable energy innovation.

Key Facts About Colorado Tech

  • Number of Tech Workers: 260,000; 8.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lockheed Martin, Century Link, Comcast, BAE Systems, Level 3
  • Key Industries: Software, artificial intelligence, aerospace, e-commerce, fintech, healthtech
  • Funding Landscape: $4.9 billion in VC funding in 2024 (Pitchbook)
  • Notable Investors: Access Venture Partners, Ridgeline Ventures, Techstars, Blackhorn Ventures
  • Research Centers and Universities: Colorado School of Mines, University of Colorado Boulder, University of Denver, Colorado State University, Mesa Laboratory, Space Science Institute, National Center for Atmospheric Research, National Renewable Energy Laboratory, Gottlieb Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account