Tech jobs
Job listings
Product Manager - AI SW Infrastructure (AI Software Stack / Compiler / Platform) - US
Lead the AI software stack, compiler, and platform roadmap for a next-gen AI infrastructure startup, defining toolchains, framework integrations, and runtime engines to optimize AI model execution on custom hardware.
Technical Product Manager, AI Inference & Software
Own the AI inference and software product roadmap, translating model innovations into technical requirements and serving-stack features while collaborating with engineering and GTM teams.
Senior Software Engineer 3 - (AWS, Kubernetes, Platform)
Build and maintain AWS and Kubernetes infrastructure that powers the organization's AI capabilities, ensuring reliability, security, and scalability for AI services.
Member of Technical Staff
Build core data-query infrastructure for AI workloads, focusing on multimodal data (images, video, text) and distributed systems using Rust/C++/Python/Go.
Senior Engineer, Platform Engineering and Architecture
Lead the architecture and engineering of a bank-grade AI & Agentic Platform, designing distributed systems, agent protocols, LLM infrastructure, and cloud-native stacks to support secure, scalable agent workloads across the enterprise.
Senior Software Developer, ML Ops
Build and maintain the ML platform that trains, deploys, and monitors AI/ML models across Wealthsimple’s products using Kubernetes, Ray, FastAPI, and MLflow.
Sovereign Engineering Platform SRE
Designs, builds, and maintains Kubernetes-based AI engineering platforms for secure, sovereignty-sensitive environments, focusing on GitOps, observability, and CI/CD for AI-assisted SDLC workflows.
AI Software Engineer
Develop and optimize deep-learning frameworks for AMD GPUs, focusing on GPU kernels, distributed inference, and compiler tech to improve training/inference performance.
Principal Software Quality Engineer, GPU & Machine Learning
Lead ROCm software validation for AMD Instinct GPUs, defining test architecture, CI/CD pipelines, and release gates for AI/ML and HPC workloads across multi-node systems.
Senior Software Engineer, AI
Build and optimize the Triton compiler and runtime for AMD GPUs to enable scalable, distributed AI workloads across multi-GPU systems.
Software Engineer, AI Triton Kernels
Develop high-performance Triton/Gluon GPU kernels for AI models, optimizing matmul, attention, and transformer layers to maximize throughput on AMD Instinct accelerators.
Founding Product Designer
Design the visual identity and developer-facing interfaces for vLLM, an AI inference engine, from brand assets to product UIs and launch campaigns.
TECHNICAL MANAGER-- GPU CLOUD & AI INFRASTRUCTURE
TECHNICAL MANAGER - GPU CLOUD & AI INFRASTRUCTURE Location: Singapore Employment Type: Full-Time, Permanent Monthly Salary: S$15,000-S$20,000 Reporting To: Head of GPU Cloud and AI Infrastructure Travel…
Software Engineer, Inference Runtime
Build and optimize LM Studio’s inference runtime for on-device and cloud AI, integrating new engines and models while improving performance across CPU/GPU targets.
Sr. Software Engineer - AI Triton Kernels
Build and optimize high-performance Triton/Gluon kernels for AI models on AMD GPUs, collaborating with compiler and hardware teams to maximize throughput and efficiency.
Developer Advocate, MAX Inference & Serving
Evangelize Modular’s MAX AI inference and serving platform through technical content, benchmarks, and community engagement to help developers deploy models efficiently.
AI Engineer / Research Scientist (Senior, Staff), Explainable AI
Design and build explainable AI systems that help users understand and contest AI outputs, working with LLMs, RAG, and agent frameworks in a hybrid Austin or Reston office.
Senior AI Infrastructure Engineer
Design and operate Kubernetes-based GPU infrastructure and AI serving platforms for large-scale model training, inference, and autonomous agents using Python, Go, or Rust.
Staff AI Infrastructure Engineer
Design and operate Kubernetes-based GPU infrastructure and AI serving platforms for large-scale model training, inference, and autonomous agents using Python, Go, or Rust.
Sr. Software Engineer - AI Triton Communication
Build and optimize Triton’s distributed GPU communication and execution stack for AMD Instinct accelerators, enabling scalable multi-GPU AI training and inference.