Jobs in speech technology & voice AI

Browse open ASR, TTS, speech-analytics and voice-AI roles by specialty. Get a weekly digest of who's hiring, and read the guides on salaries, interviews, and breaking in.

Curated, not scraped
Weekly hiring digest
Free career guides
Browse roles by specialty Get the weekly digest

Open roles this week

From the latest weekly digest — see all open roles →

Principal Research & Engineering, Realtime Voice AIPosted Oct 6
Inflection AI · Palo Alto, CA · $400K–$550K base + equity
A hands-on lead who sets the technical roadmap for Inflection's real-time voice stack, the voice side of Pi. That covers streaming ASR, TTS, speech-to-speech, speech LLMs, turn-taking, barge-in and latency, plus build-vs-buy-vs-train decisions, with a 1,000-GPU cluster for experiments. The posting asks for evaluation that goes past WER, scoring interruption handling, emotional fit and task success. You'd also mentor a team while staying close to the code.
View posting →
Speech Science Technology ManagerPosted Oct 6
Motorola Solutions (Theatro) · Richardson, TX · $130K–$178K + incentive bonus
Leads the speech team behind Theatro, Motorola's voice-driven communication platform for retail store staff. The role architects the real-time pipeline (voice-command detection, VAD, ASR, TTS, noise suppression, echo cancellation, keyword spotting) and tunes ASR engines such as Whisper, Deepgram, Cerence and Sensory for a fixed command set. It also works with hardware teams on microphone arrays. Requires 7+ years in speech with at least 2 in management. The posting has been up 30+ days, so apply soon.
View posting →
Research Scientist II, Speech AI LabPosted Oct 6
Adobe Research · San Francisco, CA · $187.1K–$270.95K in California ($142.7K–$270.95K US range) + bonus
A senior audio research scientist for Adobe's speech generative AI and multimodal work, covering speech and audio generation, audio representations and large-scale training. The lab expects independent research leadership: first-author papers, patents and prototypes that product teams can ship. PhD preferred (a research-focused master's is accepted), and 3+ years of industry research is strongly preferred.
View posting →
Research Engineer, Language — Wearables Polyglot AIPosted Oct 6
Meta · Redmond, WA + 3 other locations · $154K–$217K + bonus + equity
Voice LLMs for speech recognition, translation and synthesis on Reality Labs wearables, deployed both on servers and on the device itself. Datasets and evaluation frameworks are part of the job, not an afterthought. Requires 5+ years building speech or language models and shipping them to production. Multilingual and edge-deployment experience are listed as pluses.
View posting →
Research EngineerPosted Oct 6
Sesame · San Francisco, Bellevue or New York (onsite) · $190K–$320K + stock options
An evaluation-first role on the team building Sesame's voice companion. You'd own the offline and live eval pipelines for its speech and multimodal models, build the dataset-curation tooling and monitoring, and scale training and inference for LLM-sized workloads. Requires expert-level PyTorch and evaluation metrics that "actually predict user happiness." This is a different opening from the Research Scientist role we listed on September 8.
View posting →
Audio ML Engineer (Research)Posted Oct 6
HARMAN · Northridge, CA (hybrid) · $134.25K–$196.9K
Perception models for HARMAN's Intelligent Audio research group: quality prediction, artifact detection, acoustic scene classification and listener-preference modeling. The models have to fit embedded and cloud budgets, using quantization, pruning and distillation where needed. Asks for 5+ years of applied ML, at least 2 of them on audio, speech or acoustics. The posting says shipped product impact counts for more than credentials.
View posting →

Companies hiring in speech AI

Teams we track for the weekly digest — from big labs to Series-A startups

Tech Giants

  • Amazon (Alexa)
  • Google (Assistant)
  • Apple (Siri)
  • Microsoft (Azure)
  • Meta AI Research

Speech AI Startups

  • AssemblyAI
  • Deepgram
  • Speechmatics
  • Rev.ai
  • Otter.ai

Enterprise

  • Nuance Communications
  • SoundHound
  • Verint
  • Twilio
  • Cisco

Research Labs

  • OpenAI
  • Anthropic
  • AI2 (Allen Institute)
  • DeepMind
  • FAIR (Meta)

Browse by Specialty

Explore opportunities in specific areas of speech technology

Get the weekly Speech AI jobs digest

New ASR, TTS, voice-AI and speech-analytics roles from ~30 companies, plus salary and hiring notes. One email a week, free.

One email a week
No spam, ever
Unsubscribe anytime

Or download the free Speech AI Career Starter Kit — the whole field in one PDF.

What companies are looking for

ASR / Speech Recognition

  • Kaldi, ESPnet, Whisper
  • Wav2Vec, HuBERT
  • Acoustic & language modeling
  • End-to-end ASR architectures
  • Speaker diarization
  • Streaming recognition

Natural Language Processing

  • BERT, GPT, T5
  • Transformers, attention mechanisms
  • NER, intent classification
  • Dialogue systems
  • LLM fine-tuning
  • Hugging Face ecosystem

Text-to-Speech

  • Tacotron, FastSpeech
  • WaveNet, HiFi-GAN
  • Neural vocoders
  • Voice cloning
  • Prosody modeling
  • Multi-speaker synthesis

Audio ML

  • Audio feature extraction
  • Spectral analysis
  • Source separation
  • Audio classification
  • Noise suppression
  • Room acoustics modeling

Market data

$180K–$230K

Avg. total comp, Senior Speech Engineer (6–9 yrs)

Source: SpeechTechJobs Salary Guide, 2026

33.5%

Projected data-scientist job growth, 2024–2034

Source: U.S. Bureau of Labor Statistics

8.4B devices

Voice-enabled devices active worldwide, 2026

Source: Juniper Research

30+

Speech-AI companies tracked in our weekly digest

Source: SpeechTechJobs, updated weekly

Don't miss the next role

Get the week's speech-tech and voice-AI openings in one email.

Get the weekly digest