Skip to main content
Back to jobs

Research Engineer, Voice

External
Inflection AI logoInflection Ai · Palo Alto, CA
$225K–$325K/yrFull-timeOn-site1mo ago
PyTorchSAFe
Cover LetterConnect

Prepare for this interview

Elite

AI-generated questions, company research, and talking points tailored to this role


About the role

We're looking for a Member of Technical Staff (MTS), Research Engineer focused on voice and audio to help advance the spoken intelligence behind Pi. In this role, you'll work at the intersection of research and production-developing, training, and shipping neural models across the full spectrum of voice: speech synthesis, recognition, audio generation, and real-time spoken dialogue. You'll collaborate closely with ML engineers, product teams, and infrastructure to turn cutting-edge ideas in areas like neural audio codecs, diffusion-based TTS, and multimodal foundation models into the natural, expressive voice experiences that millions of Pi users interact with every day.

Responsibilities

  • Research, develop, and optimize neural models for voice and audio-including text-to-speech, automatic speech recognition, audio generation, and spoken dialogue systems.
  • Build and maintain production-grade training and inference pipelines for voice models, with close attention to latency, naturalness, and scalability.
  • Run experiments end-to-end: data curation, model architecture design, training, evaluation, and ablation studies.
  • Collaborate with ML engineers, product teams, and infrastructure to integrate voice models into Pi's real-time conversational stack.
  • Explore and apply advances in neural audio codecs, diffusion-based synthesis, streaming architectures, and multimodal foundation models to improve Pi's voice experience.
  • Develop robust evaluation frameworks combining perceptual metrics, automated benchmarks, and user-facing quality signals.
  • Contribute to Inflection's research culture through publications, internal reviews, and knowledge sharing.

Requirements

  • 2-5 years of research or engineering experience (including graduate work) in audio, speech, or multimodal ML.
  • Strong proficiency in PyTorch and hands-on experience training and debugging large-scale neural models on GPU/accelerator clusters.
  • Solid understanding of audio and speech fundamentals spectrograms, mel features, vocoders, codec-based representations, and signal processing.
  • Demonstrated ability to take a research idea from prototype to production: equally comfortable reading papers and writing efficient, CUDA-aware training loops.
  • Familiarity with modern generative architectures for audio (e.g., diffusion models, autoregressive codecs, flow-matching) and their trade-offs.
  • Clear, collaborative communication able to distill complex research into actionable insights for cross-functional partners.
  • Have a bachelor's degree or equivalent in Computer Science, Electrical Engineering, Linguistics, or a related field; MS or PhD strongly preferred.
  • Employee Pay Disclosures

Benefits

Inflection AI values and supports our team's mental and physical health. We are focused on building a positive, safe, inclusive and inspiring place to work. Our benefits include:Diverse medical, dental and vision options401k matching programUnlimited paid time offParental leave and flexibility for all parents and caregiversSupport of country-specific visa needs for international employees living in the Bay AreaHealth insuranceDental insuranceVision insurance401(k)Equity / stock optionsParental leave

Additional Information

About Inflection AI Inflection AI is a Public Benefit Corporation empowering people with human-centered, emotionally intelligent AI. We're shaping the future of AI by combining emotional intelligence (EQ) and raw intelligence (IQ) to elevate people's potential. Inflection AI created Pi, the world's first emotionally intelligent AI, to help people work through decisions, emotions, and challenges. Pi is a personal AI agent powered by Inflection AI's foundation model, proving that AI can be personal, empathetic, and contextually aware.


Your Match

How well this role fits your profile.

Company Intel

What employees say

Worked at Inflection AI? Share your experience

Interested in this role?

Apply on the company's website.

Cover LetterConnect