Tech jobs
Job listings
Staff Machine Learning Engineer, Generative AI, Voice & Speech
Staff ML Engineer specializing in generative AI for voice and speech, designing scalable AI-powered features and infrastructure to enable product innovation at Weave. Focuses on audio/voice models, LLMs, RAG, and distributed systems for large-scale B2B applications.
Staff ML - GenAI Engineer, Voice & Speech
Build and lead ML infrastructure for voice/GenAI at scale, enabling teams to ship AI-powered features while democratizing ML tooling for developers.
Software Developer (Golang & Kubernetes)
Builds and evolves RackAI, a cloud-native AI platform, using Go and Kubernetes to design scalable backend services, operators, and AI-driven tooling while shaping an AI-first engineering culture.

Member of Technical Staff - Low Level & Kernels Capabilities
Build low-level reinforcement learning environments targeting hardware features like GPUs, FPGAs, and vector ISAs to stress-test AI models, using C/C++/CUDA and Python.

Member of Technical Staff - Machine Learning Capabilities
Build reinforcement learning environments and reward functions to train frontier AI models on real-world ML tasks, blending research and engineering with Python, PyTorch/JAX, and LLM infrastructure.
DL Performance Software Engineer - LLM Inference
Build and optimize GPU kernels and inference frameworks (e.g., vLLM) to accelerate large language model serving, integrating research into production-grade, open-source software.
Manager Engineering-5th Fleet (OCONUS)
RELOCATION ASSISTANCE: No relocation assistance available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Secret TRAVEL: Yes, 25% of the Time Description At Northrop Grumman, our employees have incredible…
Manager, Sales Research & Ad Effectiveness
OVERVIEW OF THE COMPANY Fox Corporation Under the FOX banner, we produce and distribute content through some of the world’s leading and most valued brands, including: FOX News Media, FOX Sports, FOX Entertainment, FOX…
Senior Security Machine Learning / AI Engineer
Build and fine-tune AI models for security alert triage and risk scoring using enterprise telemetry, then deploy a multi-model routing layer that keeps costs predictable while improving accuracy over time.
Machine Learning Ops Engineer
Build and deploy LLM-powered tools, RAG systems, and agentic workflows for autonomous aircraft systems, focusing on retrieval quality, tool integrations, and production-grade AI infrastructure on Kubernetes.
Machine Learning Performance Engineer, Compiler
Optimize ML inference for edge accelerators and GPUs, focusing on transformer-based models for low-power in-vehicle compute by improving compilers, runtimes, and kernels.
Senior AI Engineer
Senior AI Engineer at DV Trading builds and deploys custom AI models for proprietary trading, fine-tuning open-weight LLMs and operating on-prem inference infrastructure to reduce costs and latency.
Senior ML Manager - Voice Teams (x/f/m)
What we do At Doctolib, we are building AI-powered healthcare solutions that make a real difference in the lives of millions of patients and healthcare professionals every day. Our AI organization is at the heart of…
Senior MLOps
Build and run the ML platform that trains, serves, monitors, and retires models in batch and real-time on CPU/GPU, using Kubernetes, Triton, MLflow, Prometheus and CI/CD.
MLOps Engineer
MLOps engineer builds and runs Kubeflow/ClearML-based ML pipelines, deploys high-load inference services on Kubernetes, and supports data scientists with GitOps and observability tooling.
AI/DL 솔루션 아키텍트 (AI/DL Solution Architect)
Designs end-to-end AI/deep-learning architectures for geospatial platforms, customizing models like LLM/RAG and deploying on GPU/cloud infrastructure for defense, government, and commercial clients.
Member of Technical Staff, Software
Build the software that automates a biology lab, turning scientific intent into executed experiments and structured, AI-ready data with full provenance.
Senior Machine Learning Engineer, ML Infrastructure- Online
Build and scale Unity Vector’s online ML inference platform, optimizing model serving for low-latency, high-reliability production systems using PyTorch, Triton, Kubernetes, and Ray.