Tech jobs
Job listings
ML Ops Engineer (EMEA Remote)
Build and operate scalable ML inference platforms using vLLM/TGI/Triton to serve AI models with low latency and high GPU efficiency for a next-gen cloud startup.
Staff Machine Learning Scientist – Personalization (Open to Remote)
Lead the design, development, and deployment of recommender systems and personalization models for Penguin Random House’s digital platforms to improve book discovery and customer engagement.
AI / LLM Inference Engineer
Engineer AI/LLM inference on GPU clusters: benchmark, tune, and optimize model serving with vLLM, Triton, or TensorRT-LLM to hit latency, throughput, and memory targets.
Data Scientist (Маркетинг)
Builds uplift and RFM models in Python to optimize marketing campaigns and A/B tests for an e-commerce company.
Staff AI Platform Engineer, Infrastructure Services
Build and scale SentinelOne’s AI Gateway infrastructure (Kong-based) to route, secure, and monitor AI coding assistant traffic, while operating self-hosted LLM stacks and driving reliability across Kubernetes and CI/CD systems.
Senior AI Platform Engineer, Infrastructure Services
Build and scale SentinelOne’s AI Gateway infrastructure (Kong-based) to route, secure, and monitor AI coding assistant traffic, while operating self-hosted LLM stacks and driving reliability across Kubernetes and CI/CD tooling.
Software Engineer, GPU Cluster Infrastructure
Build and operate a unified GPU compute platform for AI training and inference, hiding cloud complexity behind Kubernetes operators and schedulers that manage multi-cloud NVIDIA clusters.

Senior Forward Deployed Engineer I (AI Infra)
Senior engineer in Bengaluru who partners with AI-native companies to deploy, optimize, and scale production AI systems on DigitalOcean’s AI-Native Cloud, focusing on inference, agentic workloads, and platform tooling.
Data Scientist
Build and deploy ML models for media, ad-tech and e-commerce: recommendations, real-time personalization, pricing and bidding engines that directly lift CTR, eCPM and revenue.
Machine Learning Inference Manager
Own and optimize the ML inference platform that powers real-time, near-real-time, and batch predictions for sports data, ensuring low-latency, cost-efficient serving across on-premise GPUs and cloud.
AI Engineer (Serving)
Build and optimize production-grade LLM serving systems using frameworks like vLLM and Triton to run large-scale models reliably and efficiently.
ML Engineer (ML/LLM Ops)
Build and operate ML/LLM platforms for a Korean neobank, focusing on stable, scalable, and secure model training, deployment, and serving using tools like MLflow, Kubeflow, Triton, and vLLM.
ML Engineer (Platform)
Builds and operates a machine-learning platform for a securities app, focusing on LLM serving, gateway systems, and MLOps tooling in Kubernetes.
ML Engineer (Product)
Builds and deploys AI/ML systems for banking services like fraud detection, lending models, and LLM-based agents in a high-traffic, regulated environment.
Member of Technical Staff, Accelerator Systems
Optimize and port AI inference/training kernels (SGLang, Miles) across NVIDIA/AMD GPUs, TPUs, CPUs, and emerging accelerators to maximize performance on heterogeneous hardware.
Staff Machine Learning Engineer, Generative AI, Voice & Speech
Staff ML Engineer specializing in generative AI for voice and speech, designing scalable AI-powered features and infrastructure to enable product innovation at Weave. Focuses on audio/voice models, LLMs, RAG, and distributed systems for large-scale B2B applications.
Staff ML - GenAI Engineer, Voice & Speech
Build and lead ML infrastructure for voice/GenAI at scale, enabling teams to ship AI-powered features while democratizing ML tooling for developers.