Sr. Software Engineer, AI / ML Inference Platform
Dialpad
- Location
- Buenos Aires, Argentina
- Work model
- On-Site
- Level
- Senior
- Posted
- 2h ago
Skills
About this role
About Dialpad Dialpad is the AI platform for customer experience, built to resolve customer problems in real time across voice and digital. Our AI agents learn from your best human agents and improve with every interaction, helping organizations understand their customers, deliver better experiences, increase operational efficiencies, and build a lasting competitive advantage.
Unlike legacy systems built to route and answer, or standalone agentic bot vendors built to deflect, Dialpad was built to resolve. Our AI agents and human agents operate on a single platform with shared context, allowing Agentic AI to resolve issues, advance deals, and eliminate busywork through automation while seamlessly handing conversations to humans when needed, with full context preserved.
Market-leading brands, including Randstad, Motorola Solutions, Netflix, the San Diego Padres, the Colorado Rockies Baseball Club, and Cal Athletics, trust Dialpad. Dialpad is backed by Andreessen Horowitz, GV, ICONIQ Capital, and T-Mobile.
Being a Dialer At Dialpad, AI isn’t just a feature; it’s how our teams do their best work every day. We put powerful AI tools in every employee’s hands so they can move faster, think bigger, and achieve more.
We believe every conversation matters. And we’ve built the platform that turns those conversations into insight and action, for our customers and ourselves.
We look for people who are intensely curious and hold themselves to a high bar. Our ambition is significant, and achieving it requires a team that operates at the highest level. We seek individuals who embody our core traits: Scrappy, Curious, Optimistic, Persistent, and Empathetic.
Your role
We are hiring a Senior Software Engineer to build the shared AI / ML platform that takes Dialpad’s model-backed capabilities from training through production inference.
The AI / ML Platform team builds and operates GPU training infrastructure, model evaluation and lifecycle tooling, and production inference systems running on NVIDIA GPUs in GCP. We provide the common engineering foundations that allow ASR, NLP, and other AI teams to train, evaluate, release, operate, and continually improve models at enterprise scale.
Inference is an important center of gravity for this role: turning trained models into reliable, observable, efficient production services. The work is intentionally end-to-end, however, because production outcomes are shaped by decisions made throughout the model lifecycle. You will work across training clusters, model artifacts, evaluation and release workflows, serving runtimes, production operations, and feedback loops.
You will also serve as a senior engineering partner to ASR and NLP scientists. You will help teams reason about reproducibility, evaluation, scalability, hardware and runtime constraints, latency, reliability, cost, and release safety while there is still time to influence the design. You will not be expected to conduct original ML research, but you must understand training, data, evaluation, and model behavior well enough to help translate scientific work into dependable enterprise ML systems.
This is an implementation-heavy engineering role, not an operations support position.