Backend Software Engineer - Engine Team (Voice Agent)
- Company
- Deepgram
- Location
- Remote
- Work type
- Full Time
- Posted
- 2026-08-10
Job description
What You’ll Do
Improve Deepgram’s core inference services including areas in networking, speech processing, model orchestration, and observability
Develop integrations with cutting edge in-house, third-party, and open-source AI models for perception and managing conversational dynamics
Debug complex system issues that include networking, scheduling, and highly concurrent workloads
Rapidly customize backend services to support our customer needs
Partner with Product to design and implement new services, features, and/or products end to end
You’ll Love This Role If You
Thrive in a fast-paced, impact-driven environment where learning new skills on-the-fly is not only encouraged but a regular necessity
Enjoy balancing decisions about product and feature maturity to decide when to make minimally invasive changes versus when to incorporate detailed design work
It’s Important To Us That You Have
3+ years of experience in an industry role
Programming experience in Rust (or C, C++), with competence in Python
Excellent communication and organizational skills, both written and verbal.
A high level of experience and understanding of version control; preferably git.
Comprehensive experience with UNIX-style systems.
It Would Be Great If You Had
Experience with low-latency, multi-model orchestration for AI-enabled applications
Experience with audio processing
Notice: We're aware of individuals impersonating Deepgram recruiters. All legitimate Deepgram recruiting communication comes from an @deepgram.com email address. If you've received a message claiming to be Deepgram, please forward it to [email protected].
Skills Required
3+ years of experience in an industry role
Programming experience in Rust (or C, C++)
Competence in Python
High level of experience and understanding of version control (git preferred)
Comprehensive experience with UNIX-style systems
Experience with low-latency, multi-model orchestration for AI-enabled applications
Experience with audio processing