Tech jobs
Job listings
Senior Manager, Software Engineering - NIM Factory
NVIDIA is the platform upon which every new AI‑powered application is built! We are seeking a deeply technical, hands‑on Senior Engineering Manager to lead the NVIDIA Inference Microservices (NIM) Factory team. You…
Senior Software Engineer, Agentic Infrastructure
Build and maintain agentic infrastructure for aerospace systems, including local LLM inference, multi-agent systems, and production-grade compute environments to support rocket development and factory operations.
Staff Software Engineer, Agentic Infrastructure
Build and maintain agentic infrastructure for aerospace systems, including local LLM inference, multi-agent systems, and production-grade compute environments to support rocket development and manufacturing.
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Lead AI inference optimization and deployment for autonomous-vehicle models, tuning low-latency pipelines on embedded and cloud GPUs with CUDA, TensorRT, and LLM serving stacks.
Sovereign Engineering Platform SRE
Build and operate a secure Kubernetes-based platform for AI engineering tools, implementing GitOps and observability while ensuring sovereignty, reliability, and auditability in a high-security environment.
Senior DevSecOps Engineer
Build and secure AWS/Azure cloud infrastructure using IaC, CI/CD automation, and Kubernetes hardening, embedding security-by-design across the SDLC and integrating Microsoft Defender for Cloud.
Member of Technical Staff — Inference-Multi-Hardware
Optimize and port AI inference/training kernels across NVIDIA, AMD, TPUs, and emerging accelerators, designing portable abstractions for SGLang and Miles.
Runtime Engineer
What MatX is Building MatX is building custom silicon for large-language-model inference and training, with HW/SW co-design across ISA, RTL, simulator, compiler, and kernels so each layer benefits from the others. The…
Senior Software Developer, LLM Infrastructure
Build and maintain the ML platform that trains, deploys, and monitors AI/ML models across Wealthsimple’s investing, crypto, and other financial products.
MLOps Platform Developer / Full-Stack AI Engineer
Build and run a full-stack AI platform: React/TypeScript web app, Postgres telemetry, LLM fine-tuning and vLLM serving, plus MQTT/BMS integrations for heat-network decarbonisation.
Senior AI Compute Infrastructure Engineer
Engineer and optimize GPU clusters for AI training and inference, build scheduling/orchestration systems, and improve performance, cost, and reliability of AI infrastructure.
AI Infrastructure Engineer
Optimize LLM inference performance on Intel GPUs by profiling bottlenecks, writing custom kernels, and contributing to open-source serving frameworks like vLLM and SGLang.
Senior Software Engineer 3 - (Python, Kubernetes, Helm)
Build and maintain AI inference infrastructure using Python, Kubernetes, and Helm to deploy and optimize LLM services for internal and customer use.
Senior Devops - F/H/NB
Optimize AI model inference for low-latency production deployment, manage Kubernetes-based ML infrastructure, and collaborate with data-science teams to scale and automate systems handling thousands of concurrent requests.
Applied AI Engineer
Build and ship AI features end-to-end: design agent workflows, fine-tune models, and turn outputs into reliable production systems using Python, PyTorch/JAX, and modern LLMs.
Engineering Manager, DevOps
Lead a DevOps team to design, deploy, and scale cloud and Kubernetes infrastructure for an enterprise AI platform, focusing on reliability, CI/CD, and observability.
Senior Software Developer, ML Ops (Remote)
Build and maintain ML infrastructure to deploy and scale AI models in production, using Python, Kubernetes, and Ray. Own platform services that enable data scientists to move models from experimentation to live systems.
Forward Deployed AI Engineer
Build and deploy enterprise-grade AI/ML solutions with customers, moving prototypes to production while advising on GenAI strategies and scalable architectures.
AI Enterprise Architect
Design and lead end-to-end AI infrastructure, from GPU clusters and high-performance networking to MLOps and governance frameworks, ensuring scalable, compliant, and production-ready AI systems.