Senior engineer building Dialpad's shared AI/ML platform that takes models from training to production inference, including GPU training infrastructure on GCP, model lifecycle tooling, and low-latency NVIDIA GPU inference services. Day-to-day is hands-on backend/infra work in Python or Go, Kubernetes, and model-serving runtimes, partnering with ASR and NLP scientists.
Sign in to see your match