You and the process

With a person
Anna, in-house recruiter
9 years hiring engineers
30 minutes with a real recruiter
They read your CV with you, on a call, and say where the offers are being lost.
Didn’t find what you were looking for? Tell us what to build
Be the first to open itNo views yet
Decagon

Research Engineer, Audio and Speech

  • Office
  • 0-3 years

Salary

$200,000 - 400,000/ year

AI summary

For members

The whole posting in a few lines. Sign up to read it here and on every role you open.

Sign Up to Read

Description

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experiences.

Our technology enables industry-defining enterprises like Avis Budget Group, Block’s Cash App and Square, Chime, Oura Health, and Hunter Douglas to deploy AI agents that power personalized, deeply satisfying interactions across voice, chat, email, SMS, and every other channel.

We’re building a future where customer experiences are being redefined from support tickets and hold music to faster resolutions, richer conversations, and deeper relationships. We’re proud to be backed by world-class investors who share that vision, including a16z, Accel, Bain Capital Ventures, Coatue, and Index Ventures, along with many others.

We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values — Just Get It Done, Invent What Customers Want, Winner’s Mindset, and The Polymath Principle — shape how we work and grow as a team.

Read more about the Speech Research Team's work:

The Research team develops the model and decision-making stack that powers Decagon’s conversational agents for enterprise support. We research, adapt, and implement state-of-the-art techniques in model training, prompting, orchestration, and evaluation in order to make our agents more accurate, robust, and efficient in real-world deployments.

Our goal is to push the frontier of applied conversational AI: agents that reliably understand nuanced intent, track long context, and take the right actions under uncertainty. We measure success the way customers feel it: higher resolution rates, better user satisfaction, and consistent behavior at scale.

About the Role

As a Research Engineer focused on Audio and Speech, you’ll be responsible for building the models and agent harnesses that power Decagon’s real-time voice agents and taking them all the way from idea to production. Your work will advance multimodal and full-duplex systems that can listen, reason, speak, and respond naturally in real time.

We’re looking for strong engineers who want to build the next generation of AI voice agents. People here own their work end-to-end, ship real improvements, and are trusted to make high-impact technical decisions.

In this role, you will

Design and build next-generation agent harnesses optimized for streaming speech, turn-taking, interruptions, overlapping speech, and continuous interaction

Research and train multimodal and full-duplex models that jointly understand audio, reason, and generate speech

Improve speech recognition, voice activity detection, endpointing, and speech generation across diverse speakers, environments, domains, and languages

Build evaluations and use production calls to ship measurable improvements in accuracy, latency, naturalness, and task outcomes

Optimize end-to-end inference for responsiveness, throughput, stability, and cost, partnering with Voice Platform and Infrastructure teams to deploy at scale

Your background looks something like this

2+ years of experience in speech, audio ML, multimodal ML, or production machine learning

Experience developing or adapting autoregressive, diffusion, flow-matching, or codec-based speech models

Hands-on experience with streaming agent systems, low-latency inference, production model serving, and evaluation on real-world audio

Fluency in Python and a modern deep-learning framework such as PyTorch, with strong foundations in machine learning and signal processing

A track record of taking research ideas from prototype to reliable, measurable production impact

Even better if you have

Familiarity with speech-to-speech or full-duplex models

Experience with telephony, multilingual speech, noisy-channel robustness, speaker adaptation, or expressive speech generation

Compensation

$200K – $400K + Offers Equity

Benefits

  • We proudly offer the following benefits for our full-time employees:
  • Medical, Dental, and Vision benefits for you and your family
  • Life Insurance and Disability Benefits
  • Retirement Plan (e.g., 401K, pension)
  • Parental Leave
  • Fertility and family building benefits through Carrot
  • Monthly stipend to support your wellness, lifestyle, and work-life balance
  • Daily lunches and snacks in the office to keep you at your best
  • Take what you need vacation policy (subject to local requirements; UK employees receive 25 days of statutory leave)
  • These benefits are described in more detail in Decagon’s policies, may vary by location, and can change at any time according to applicable compensation and benefits plans.

Where you’d work

From the office

About the company

Decagon

  • Industry: AI

Offices in San Francisco, United States, New York City, United States

Also hiring in Australia and London, United Kingdom

Your chances

Worth a look before you spend an evening tailoring a CV for it.

  • 21 checks run
  • 4 red flags

Still hiring?

14 checks

1 red flag

How crowded?

7 checks

3 red flags

Fits Me

How well does this role fit you?

Answer a few questions or drop your CV, and every role gets a fit score with the reasons, this one first.

  • Your field
  • Level
  • Stack
  • Work model
  • Salary floor
  • Must-haves
Details12 facts · Role, Location, Compensation, Employment, Company
Tech stack
  • Python
  • PyTorch
Type
Full-time
Equity
Equity offered
Industry
AI
Specialty
ML
Region
United States
Show 6 more factsShow less

Role

Category
Data & Analytics
Specialty
ML
Experience
2+ years
Tech stack
  • Python
  • PyTorch

Location

Work model
Office
Region
United States
Offices
  • San Francisco, United States
  • New York City, United States

Compensation

Salary
$200,000 - 400,000 / year
Pay period
Annual
Equity
Equity offered

Employment

Type
Full-time

Company

Industry
AI

Something wrong with this vacancy?

Similar vacancies

  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Data & AI Engineer Intern

    Salary by agreement

    • Hybrid · Paris
    • Junior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    AI Engineer

    $80,000 - 210,000 / year

    • Office · CA, Austin
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Senior AI/ML Engineer

    $200,000 - 260,000 / year

    • Office · San Francisco
    • Senior
  • Early Window: Be the first to open itNo views yetCloses in
    Company hidden

    Member of Technical Staff, ML Infra

    Salary by agreement

    • Office · San Francisco
    • Senior

Share this vacancy

What's wrong with it?

The employer never sees who reported.

Reason