Career guides, technical tutorials, and industry insights for speech engineers
Subscribe via RSSA principal real-time voice AI lead at Inflection AI at $400–550K base, a new-grad speech role on ByteDance Seed at $218–388K, speech research at Adobe, Meta and Sesame, real-time voice systems at ASAPP, LILT, Cloaked and Wispr, a remote technical enablement role at Deepgram, plus HARMAN and Motorola Solutions.
A big-money week: Grok voice-model engineering at xAI up to $450K, two audio and voice roles at Decagon at $200–400K, multimodal audio research at Dolby, real-time conversational AI at Amazon AGI, audio inference at Cohere, speech data and evaluation at Innodata, plus applied ASR at ClickUp, Speak, Clera, Qualitate, Assort Health and ALTEN.
It's the 27th Interspeech, not the 37th — 28 September to 1 October 2026 at ICC Sydney. What's actually on, whether it's worth the flight, how it differs from ICASSP, and the 2027 São Paulo deadline.
Toronto, 16–21 May 2027, papers due 23 September 2026 with no extension expected. Every important date, what the acceptance rate actually is, whether it counts as tier 1, and ICASSP vs Interspeech.
A research-heavy week: a TTS research director and data-science research staff at Deepgram, audio-to-audio research at Google DeepMind, audio ML at David AI up to $360K, speech synthesis at Amazon Prime Video, codec and DSP research at Qualcomm, voice-model serving at Together AI, plus real-time speech engineering at Cantina, Guava and Ford.
A mixed week: spatial-audio research at Dolby, voice-agent engineering at Wispr and Retell AI, a speech-analytics pipeline role at Gong, edge and platform work at Deepgram and Speechmatics, plus product and design roles at Descript, Rev and Picovoice.
A research-heavy week: generative-audio and speech research at Spotify, Google, Sesame and Rad AI; real-time voice-agent engineering at Decagon, Amazon, NRG, Weave and PolyAI; audio and platform infrastructure at Meta and Speechify. Eleven roles, most remote-friendly.
Most teams that want to "hire a Whisper developer" actually need a short contract. Contractor vs in-house vs agency, what to screen for, five interview questions, and where to find candidates.
ASRU is biennial (odd years), so there's no ASRU 2026. ASRU 2025 was in Honolulu; the next is ASRU 2027. The even-year IEEE workshop is SLT 2026 in Palermo (Dec 13–16). Plus Interspeech 2026 and the full calendar.
Twelve speech-tech roles this week: ASR and TTS research at Deepgram, ElevenLabs, NVIDIA, Otter and Rime; voice-agent and in-car roles at Cartesia, SoundHound and Cerence; plus healthcare speech at DeepScribe. Most remote-friendly.
Yes — but "Kaldi" now means two projects. Where classic Kaldi and next-gen Kaldi (k2, icefall, sherpa) still beat Whisper, and what it means if Kaldi shows up in a job description.
How the two revenue-intelligence companies differ in structure, product and engineering work — and how to get real compensation numbers before you decide.
Comprehensive interview prep for speech analytics roles. Technical questions on diarization, sentiment analysis, NLP, system design, and coding. Real answers from Gong, Chorus, and CallMiner engineers.
Where speech engineers get hired for conversation intelligence, contact-center analytics and meeting AI — grouped by what each company builds, with careers links.
Complete guide to speaker diarization technology and careers. Learn how "who spoke when" systems work, required skills, salaries ($150K-$240K), and companies hiring in 2026.
Complete guide to Whisper AI career opportunities. Discover salaries ($140K-$230K), required skills, companies hiring, and how to land a Whisper specialist role in 2026.
Complete guide to Kaldi speech recognition toolkit. Learn what Kaldi is, how it works, when to use it vs Whisper, and career opportunities for Kaldi engineers in 2026.
Step-by-step guide to launching a speech tech career without a PhD. Learn the skills, build a portfolio, and land your first ASR job in 2026. Real advice from self-taught engineers.
Technical comparison of the three major ASR frameworks. Learn when to use Kaldi (production), Whisper (general-purpose), or Wav2Vec 2.0 (research/low-resource).
Find fully remote ASR roles at top companies. Learn which companies hire remotely, salary expectations ($130K-$240K), time zone requirements, and how to land remote speech tech jobs.
Complete salary breakdown for speech recognition engineers. Discover pay ranges by experience level, location, and company type. Base + equity + bonus data for $130K-$280K+ total comp.
Master speech recognition interviews with 40+ real questions. Covers ASR fundamentals, acoustic modeling, language models, system design, and coding challenges at top companies.
Where speech-recognition engineers get hired — dedicated ASR-API companies, big labs, voice-agent startups and vertical players. What each builds, the roles they open, and where to apply.
Submit your profile to get matched with top companies hiring speech engineers.
Get the weekly digest