Sr. AI Engineer (AI/ML Inference)
full-time
senior
Posted 19 hours ago
Before you apply
Build my evidence-backed draft — free Apply on company site →Paste your relevant resume section or 2–4 true bullets. See supported requirements and honest gaps. No account and no application sent.
About this role
About Dialpad Dialpad is the AI platform for customer experience, built to resolve customer problems in real time across voice and digital. Our AI agents learn from your best human agents and improve with every interaction, helping organizations understand their customers, deliver better experiences, increase operational efficiencies, and build a lasting competitive advantage.
Unlike legacy systems built to route and answer, or standalone agentic bot vendors built to deflect, Dialpad was built to resolve. Our AI agents and human agents operate on a single platform with shared context, allowing Agentic AI to resolve issues, advance deals, and eliminate busywork through automation while seamlessly handing conversations to humans when needed, with full context preserved.
Market-leading brands, including Randstad, Motorola Solutions, Netflix, the San Diego Padres, the Colorado Rockies Baseball Club, and Cal Athletics, trust Dialpad. Dialpad is backed by Andreessen Horowitz, GV, ICONIQ Capital, and T-Mobile.
Being a Dialer At Dialpad, AI isn’t just a feature; it’s how our teams do their best work every day. We put powerful AI tools in every employee’s hands so they can move faster, think bigger, and achieve more.
We believe every conversation matters. And we’ve built the platform that turns those conversations into insight and action, for our customers and ourselves.
We look for people who are intensely curious and hold themselves to a high bar. Our ambition is significant, and achieving it requires a team that operates at the highest level. We seek individuals who embody our core traits: Scrappy, Curious, Optimistic, Persistent, and Empathetic .
Your role As a Sr. AI Engineer: Systems, you’ll serve as an embedded senior back-end engineer on our Speech Team, owning the production systems that turn speech models and third-party capabilities into reliable, scalable experiences for Dialpad’s AI voice agents. You’ll work at the intersection of speech, ML infrastructure, and product engineering: productionizing models, enabling self-hosted inference, integrating external APIs, and building the operational foundations required for strong uptime and latency SLAs. You’ll partner closely with the MLOps (Inference) team while bringing deep ownership of the speech domain, helping the team move quickly from promising model or vendor capability to safe, observable, and cost-effective production. This role offers broad technical influence and the opportunity to shape how Dialpad operates real-time speech systems at scale.
This position reports to our Senior Manager, AI Speech, and offers the opportunity to be based in our Canada Hub locations.
What you’ll do
Productionization & Service Ownership: Own the path from speech model or third-party capability to production, building the APIs, services, deployment workflows, and integration layers that make it safe and easy for the Speech Team to ship improvements.
Self-Hosted Inference & Scaling: Productionize and operate self-hosted speech models, optimizing serving architecture, resource utilization, concurrency, autoscaling, and cost so they can meet the demands of real-time voice agents.
Third-Party APIs & Provider Resilience: Integrate and maintain third-party speech APIs behind durable abstractions, with clear failover, capacity planning, version management, and vendor-performance monitoring.
Reliability, SLOs & Observability: Build the monitoring, alerting, dashboards, health checks, and incident-response practices needed to meet uptime, latency, and quality SLAs for customer-facing speech systems.
Release & Evaluation Infrastructure: Partner with Speech and MLOps engineers to enable shadow traffic, staged rollouts, model and artifact versioning, rollback-safe releases, and candidate-versus-incumbent comparisons.
Cross-Functional Leadership & Mentorship: Work closely with the MLOps (Inference) team and partner teams across speech, platform, telephony, and product to set technical direction, mentor engineers, and turn model advances into reliable production impact.
Skills you’ll bring
Systems & Backend Engineering: Strong software engineering fundamentals and proficiency in Python, plus experience designing maintainable APIs, services, and integration layers. We’re open to candidates who are strongest in backend/platform engineering or who have grown from ML into systems.
Production ML & Streaming: 5+ years of experience building or operating production software, including ML-backed systems, real-time services, speech applications, streaming media, or other latency-sensitive systems.
Model Serving & Inference: Hands-on experience deploying, scaling, and troubleshooting ML models in production, including model serving, inference optimization, resource management, and safe model and version rollouts.
Cloud & Distributed Systems: Experience with cloud infrastructure and distributed systems, as well as familiarity w
Similar Jobs
Related searches:
On-site Jobs
Senior Jobs
On-site Senior Jobs
Senior Machine LearningSenior AI InfrastructureSenior AI Agents & RAGSenior Backend & SystemsSenior Generative AI
AI Jobs in Vancouver
Machine Learning in VancouverAI Infrastructure in VancouverAI Agents & RAG in VancouverBackend & Systems in VancouverGenerative AI in Vancouver
fine-tuningmlopsclouddistributed-systemsagentsinference
Get jobs like this delivered weekly
Free AI jobs newsletter. No spam.