Deepgram
Backend Software Engineer - Engine Team (Voice Agent)
About this role
Deepgram seeks a backend engineer to design and build scalable voice agent services, including speech processing pipelines, AI model integrations, and telephony systems. You'll work in a fast-paced, AI-first environment building production infrastructure for the voice AI platform.
What you'll do
- Design and implement secure, scalable backend services for speech processing and voice agent orchestration
- Build integrations with telephony providers, RAG systems, and third-party AI models
- Debug and optimize complex distributed systems involving networking, scheduling, and concurrent workloads
- Improve core inference services including networking, model orchestration, and observability
- Partner with Product to design and implement end-to-end features and services
- Customize backend services rapidly to support customer requirements
What they're looking for
- Rust (or C/C++) programming
- Python development
- Low-latency AI model orchestration
- Audio processing
- UNIX/Linux systems administration
- Git version control
- Distributed systems debugging
- API design and integration
Benefits
- Medical, dental, vision insurance with annual wellness stipend
- Mental health support
- Unlimited PTO and flexible schedule
- Parental leave
- Life, short-term, and long-term disability insurance
- 12 paid US company holidays
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Deepgram
Deepgram builds voice AI and speech processing technology, offering platforms that power real-time speech recognition and audio intelligence across cloud, edge, and embedded environments. The company is hiring applied ML engineers, backend engineers, embedded AI engineers, and customer-facing roles (customer success and pre-sales engineers) to scale its AI infrastructure and drive enterprise adoption.
- Website
- deepgram.com
Likely interview questions
- Tell us about your experience building low-latency, production-grade backend services. How have you approached scaling and optimizing for real-time performance?
- Describe a complex system issue you've debugged that involved networking, concurrency, or distributed systems. Walk us through your debugging approach.