Tech jobs
Job listings
Member of Technical Staff — Performance
Performance Engineer at RadixArk in Palo Alto optimizes LLM inference and training systems for latency, throughput, and cost efficiency across production workloads using SGLang, Miles, and GPU/TPU infrastructure.
Member of Technical Staff — Product
Builds and maintains developer-facing tools (APIs, SDKs, dashboards) on top of AI inference/training infrastructure like SGLang and Miles, collaborating with product and research teams to create intuitive interfaces.
Member of Technical Staff — Backend/API Platform
Build and maintain the backend and API platform powering SGLang and Miles, including REST/gRPC services, authentication, multi-tenancy, and monitoring for AI infrastructure.
Member of Technical Staff — Developer Experience
Builds and engages the technical community around SGLang and Miles by creating content, speaking at events, and collaborating with developers to optimize AI infrastructure.
Member of Technical Staff — Developer Technology
Optimize and accelerate LLM inference and training systems like SGLang and Miles by profiling GPU performance, writing custom kernels, and enabling new models on modern hardware.
Senior Solution Engineer – GPU & AI Infrastructure
Designs and deploys GPU-powered AI and HPC infrastructure on Civo’s cloud platform, optimizing Kubernetes and bare-metal performance for AI workloads.
Sr. Computer Vision Engineer (Deep Learning)
Design and deploy deep-learning perception models for ADAS in commercial EVs, optimizing neural networks for embedded automotive hardware.
Staff Machine Learning Engineer
Build and scale ML systems for autonomous driving, focusing on reinforcement learning, generative models, and evaluation workflows to improve Waymo’s self-driving technology.
Staff Machine Learning Engineer (Gaia)
Build and improve Wayve’s Gaia video world model, training large-scale ML models to predict future driving frames and generate synthetic scenarios for autonomous vehicles.
Staff Machine Learning Engineer
Build and scale ML systems for autonomous driving, focusing on reinforcement learning, generative models, and evaluation workflows to improve self-driving behavior.
Senior Backend Engineer — AI Training & Experimentation
Senior Backend Engineer builds and scales distributed systems for AI training, experiment management, and developer tooling at a platform company.
Senior Software Engineer, ML Infrastructure Platform
Build and maintain ML infrastructure for Nuro’s autonomous driving models, including distributed GPU training, data pipelines, and Kubernetes-native orchestration to enable safe, scalable self-driving vehicles.
Software Engineer, ML Infrastructure Platform
Build and maintain ML infrastructure for Nuro’s autonomous driving models, including distributed GPU training, data pipelines, and orchestration systems to keep autonomy development running smoothly.
Ingénieur(e) IA Générative - Hébergement De Modèles LLM
Design and build a cloud-native SaaS platform for data management and MLOps that supports 5G core networks, using Python, Go, Kubernetes, and hyperscale cloud services.
Ingénieur(e) IA Générative - Hébergement De Modèles LLM
Build and run a multi-tenant GPU cluster platform for LLM training and inference using Ray, NVIDIA stacks, and Kubernetes, while exposing OpenAI-compatible endpoints and developer tooling.
Senior Principal AI Engineer (Remote)
Build and optimize distributed training systems for large neural networks (LLMs, diffusion, SSMs) across GPU clusters, focusing on throughput, stability, and fault tolerance using PyTorch, Megatron-LM, and DeepSpeed.
Senior Machine Learning Engineer
Build and scale ML systems that connect, price, and surface travel inventory across Expedia’s global marketplace, improving outcomes for travelers and partners.
Research Robotics/Computer Vision Engineer
Researches and builds 3D computer vision and SLAM systems for autonomous robots, focusing on perception pipelines, navigation, and manipulation in real-world environments.