The posting
At ASAPP, our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. We are guided by principles that shape how we think, build, and execute, including deep customer obsession, purposeful speed, ownership, and a relentless focus on outcomes. We work in small, highly skilled teams, prioritize clarity over complexity, and continuously evolve through curiosity, data, and craftsmanship.
We’re building a globally diverse team of technologists and problem solvers who thrive in fast-paced environments, value collaboration, and approach every challenge with a Day 1 mindset. With hubs in New York City, Mountain View, Latin America, and India. If you’re driven by continuous learning, rapid iteration, and the challenge of building in a high-growth startup, this is more than a role—it’s a journey.
We are seeking a Senior Speech Software Engineer to drive both the infrastructure and applied speech intelligence behind our real-time voice AI platform. This is not just a systems role — you will operate at the intersection of speech research, model optimization, and production engineering, ensuring our ASR and TTS systems meet the demanding quality, latency, and reliability requirements of enterprise call centers.
You will help evolve our speech stack to deliver human-like, low-latency voice interactions at massive scale, tuning and adapting modern speech models to perform in noisy, real-world customer environments. You will work closely with Speech Scientists, ML Researchers, and Infrastructure Engineers to bridge cutting-edge speech technology with hardened production systems.
What you'll do
What you'll need
What we'd like to see
- Experience with noise reduction, echo cancellation, VAD, diarization, or other speech enhancement technologies
- Familiarity with forced alignment techniques or phoneme/word-level timing models
- Hands-on experience deploying ML services with Kubernetes, Docker, and cloud platforms (AWS/GCP/Azure)
- Knowledge of event-driven and asynchronous systems (e.g., async I/O, event loops, streaming frameworks)
- Experience analyzing large-scale speech or conversation datasets to drive model or system improvements
ASAPP is committed to creating a diverse environment and is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, disability, age, or veteran status. If you have a disability and need assistance with our employment application process, please email us at [email protected] to obtain assistance. #LI-AG1 #LI-Hybrid



