AssemblyAI Logo

AssemblyAI

Senior Research Engineer

Reposted 7 Days Ago
Be an Early Applicant
Remote
Hiring Remotely in USA
240K-275K Annually
Senior level
Remote
Hiring Remotely in USA
240K-275K Annually
Senior level
This role involves optimizing large-scale distributed training and inference systems, implementing deep learning optimizations, and collaborating across teams to enhance AI models.
The summary above was generated by AI
About AssemblyAI

At AssemblyAI, we’re building at the forefront of Speech AI, creating powerful models for speech-to-text and speech understanding available through a straightforward API. With more than 200,000 developers building on our API and over 5,000 paying customers, AssemblyAI is helping unlock and support the next generation of powerful, meaningful products built with AI. 

Progress in AI is moving at an unprecedented pace– and our team is made up of experts in AI research that are focused on making sure that our customers are able to stay on the cutting edge, with production-ready AI models that are constantly updating and improving as our team continues to improve accuracy, latency, and what’s possible with Speech AI. Our models consistently rank highest in industry benchmarks for accuracy, outperforming models from Google and Amazon, and up to 30% fewer hallucinations than OpenAI’s Whisper. Our models power more than 2 billion end-user experiences each day, helping companies better understand customer feedback, run more productive meetings with automated meeting notes, and helping improve childhood literacy via ed tech tools. 

We’ve raised funding by leading investors including Accel, Insight Partners, Y Combinator’s AI Fund, Patrick and John Collision, Nat Friedman, and Daniel Gross. We’re a remote team looking to build one of the next great AI companies, and are looking for driven, talented people to help us get there!

About the Role

We are seeking a highly skilled Senior Research Engineer to collaborate closely with both Research and Engineering teams. The role involves diagnosing and resolving bottlenecks across large-scale distributed training, data processing, and inference systems, while also driving optimizations for existing high-performance pipelines.

The ideal candidate possesses a deep understanding of modern deep learning systems, combined with strong engineering expertise in areas such as layer-level optimization, large-scale distributed training, streaming, low-latency and asynchronous inference, inference compilers, and advanced parallelization techniques.

This is a cross-functional role requiring strong technical rigor, attention to detail, intellectual curiosity, and excellent communication skills. The position is embedded within the Research team and is responsible for developing and refining the technical foundation that enables cutting-edge research and translates its outcomes into production, bridging research and production engineering.

What You'll Do
  • Investigate and mitigate performance bottlenecks in large-scale distributed training and inference systems.
  • Develop and implement both low-level (operator/kernel) and high-level (system/architecture) optimization strategies.
  • Translate research models and prototypes into highly optimized, production-ready inference systems.
  • Explore and integrate inference compilers such as TensorRT, ONNX Runtime, AWS Neuron and Inferentia, or similar technologies.
  • Design, test, and deploy scalable solutions for parallel and distributed workloads on heterogeneous hardware.
  • Facilitate knowledge transfer and bidirectional support between Research and Engineering teams, ensuring alignment of priorities and solutions.
What You'll Need
  • Strong expertise in the Python ecosystem and major ML frameworks (PyTorch, JAX).
  • Experience with lower-level programming (C++ or Rust preferred).
  • Deep understanding of GPU acceleration (CUDA, profiling, kernel-level optimization); TPU experience is a strong plus.
  • Proven ability to accelerate deep learning workloads using compiler frameworks, graph optimizations, and parallelization strategies.
  • Solid understanding of the deep learning lifecycle: model design, large-scale training, data processing pipelines, and inference deployment.
  • Strong debugging, profiling, and optimization skills in large-scale distributed environments.
  • Excellent communication and collaboration skills, with the ability to clearly prioritize and articulate impact-driven technical solutions.

Pay Transparency:

AssemblyAI strives to recruit and retain exceptional talent from diverse backgrounds while ensuring pay equity for our team. Our salary ranges are based on paying competitively for our size, stage, and industry, and are one part of many compensation, benefit, and other reward opportunities we provide.

There are many factors that go into salary determinations, including relevant experience, skill level, qualifications assessed during the interview process, and maintaining internal equity with peers on the team. The range shared below is a general expectation for the function as posted, but we are also open to considering candidates who may be more or less experienced than outlined in the job description. In this case, we will communicate any updates in the expected salary range.

The provided range is the expected salary for candidates in the U.S. Outside of those regions, there may be a change in the range which will be communicated to candidates throughout the interview process.

Salary range: $210,000 - $275,000

Working at AssemblyAI

We are a small but mighty group of startup veterans and experienced AI researchers with over 20 years of expertise in Machine Learning, Speech Recognition, and NLP. As a fully remote team, we’re looking for people to join our team who are ambitious, curious, and lead with integrity. We’re still in the early days of AI and of AssemblyAI’s journey, and are looking for teammates who won’t just fit in, but will help us define and build our company culture. 

We’re committed to creating a space where our employees can bring their full selves to work and have equal opportunity to succeed. No matter your race, gender identity or expression, sexual orientation, religion, origin, ability, age, veteran status, if joining this mission speaks to you, we encourage you to apply!

Using AI to Interview:

If you’re selected for an interview, please review this resource to better understand how AssemblyAI approaches the use of AI in our interview process.

Keep Exploring AssemblyAI:

Check us out on YouTube!

Learn more about AI models for speech recognition

Core Transcription | Audio Intelligence | LeMUR | Try the Playground

Our $50M Series C fundraise

Top Skills

Aws Neuron
C++
Cuda
Inferentia
Jax
Onnx Runtime
Python
PyTorch
Rust
Tensorrt

Similar Jobs

20 Days Ago
Remote
United States
150K-200K Annually
Senior level
150K-200K Annually
Senior level
Cybersecurity
The Senior Security Engineer will design and build security tools, collaborate with teams, and contribute to AI/ML security research. Responsibilities include security solution architecture, secure implementation, and effective communication of technical concepts.
Top Skills: C++Ci/CdGitGoJavaPythonRust
2 Days Ago
Remote
USA
190K-217K Annually
Senior level
190K-217K Annually
Senior level
Artificial Intelligence
The Senior Research Engineer will enhance JAX frameworks, optimize JAX inference for speech models, and bridge research and engineering teams to improve performance and scalability of AI systems.
Top Skills: C++CudaFlaxGpuJaxOptaxPythonRustTpuXla
6 Days Ago
In-Office or Remote
New York, NY, USA
176K-252K Annually
Senior level
176K-252K Annually
Senior level
Music
The Senior Research Engineer collaborates with researchers to enhance AI music technologies, optimize model performance, and integrate solutions into production environments while maintaining high code quality.
Top Skills: AWSGoogle Cloud PlatformAzurePyTorch

What you need to know about the Boston Tech Scene

Boston is a powerhouse for technology innovation thanks to world-class research universities like MIT and Harvard and a robust pipeline of venture capital investment. Host to the first telephone call and one of the first general-purpose computers ever put into use, Boston is now a hub for biotechnology, robotics and artificial intelligence — though it’s also home to several B2B software giants. So it’s no surprise that the city consistently ranks among the greatest startup ecosystems in the world.

Key Facts About Boston Tech

  • Number of Tech Workers: 269,000; 9.4% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Thermo Fisher Scientific, Toast, Klaviyo, HubSpot, DraftKings
  • Key Industries: Artificial intelligence, biotechnology, robotics, software, aerospace
  • Funding Landscape: $15.7 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Summit Partners, Volition Capital, Bain Capital Ventures, MassVentures, Highland Capital Partners
  • Research Centers and Universities: MIT, Harvard University, Boston College, Tufts University, Boston University, Northeastern University, Smithsonian Astrophysical Observatory, National Bureau of Economic Research, Broad Institute, Lowell Center for Space Science & Technology, National Emerging Infectious Diseases Laboratories

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account