Tech jobs
Job listings
Machine Learning Engineer
Build and deploy large-scale machine learning models and AI agents to optimize Micron’s semiconductor manufacturing workflows using distributed training and GPU optimization techniques.
GPU Performance Architect
Architect GPU performance by analyzing game/AI workloads, identifying bottlenecks, and proposing hardware improvements using C/C++, Python, and GPU APIs like Vulkan/CUDA.
Deep Learning Software Engineer, TensorRT Performance (Remote)
Build and optimize NVIDIA’s deep-learning inference stack (TensorRT, TensorRT-LLM) by profiling models, writing GPU kernels, and integrating OSS frameworks to maximize GenAI performance across datacenter and edge GPUs.
Senior Deep Learning Frameworks CUDA Software Engineer (Remote)
Build and optimize CUDA features for AI frameworks like PyTorch and TRT-LLM, improving multi-GPU performance and distributed runtime for training and inference workloads.
Deep Learning Performance Software Engineer, LLM Inference
Build and optimize GPU-powered inference systems for large language models, improving serving efficiency and performance through kernel tuning and novel algorithms.
Solutions Architect, AI Cloud Services
Design and deploy AI/ML solutions on cloud GPU platforms, collaborating with customers to integrate NVIDIA’s full stack of hardware and software technologies.
Middle Python / MLOps Engineer
Build and scale FastAPI-based REST services, set up CI/CD, and deploy ML models in a high-load ecommerce environment using Python, Kubernetes, and MLOps tooling.
Cloud and AI Platform Architect
Architect and own LSports’ GCP-based cloud platform and its agentic AI layer, designing scalable, secure systems for real-time sports data and LLM-powered automation.
Machine Learning Engineer, Adobe Firefly Services
Senior ML engineer building scalable GenAI inference pipelines and APIs that power Adobe’s Firefly, Photoshop, and other creative tools using PyTorch, CUDA, and diffusion models.
Machine Learning Platform Engineer
Build and maintain the ML platform that deploys, monitors, and scales AI models powering Hadrian’s autonomous factories, using MLflow, Dagster, EKS, and FastAPI.
AI Engineer (ML Systems & Infrastructure)
Builds and optimizes large-scale AI infrastructure, including Kubernetes clusters, RDMA networking, and GPU orchestration to improve efficiency and scalability of AI training/inference systems.
Senior Engineer, Platform Engineering and Architecture
Lead the architecture and engineering of a bank-grade AI & Agentic Platform, designing agentic runtimes, LLM gateways, identity layers, and cloud-native infrastructure to support secure, scalable agent workflows across the enterprise.
AI Infrastructure Engineer
Optimize LLM inference performance on Intel GPUs by profiling bottlenecks, writing custom kernels, and contributing to open-source serving frameworks like vLLM and SGLang.
Principal/Senior Principal Systems Test Engineer
Lead end-to-end system validation for the Triton Program, designing and executing automated/manual tests to verify functional requirements and improve mission readiness.
Systems Test Engineer - Hardware 2/3
Designs and tests hardware systems for aerospace/defense using CATIA V5 to model ground equipment, cables, and racks, and releases drawings via Teamcenter PLM.
Senior Manager, Performance Engineering – Kernel and Software Platforms
NVIDIA’s accelerated computing platform relies on continuous performance excellence at every stage of development. We are seeking an outstanding Performance Analysis Manager to lead an engineering team responsible for…

Senior Machine Learning Engineer
About this opportunity: At Freenome, we are seeking a Senior Machine Learning Research Engineer to join the Machine Learning Science (MLS) team, within the Computational Science department. The ideal candidate has a…
Lead Software Engineer - LLM Ops Platform Reliability
Build and operate scalable LLM serving infrastructure using cloud and Kubernetes, ensuring reliability and performance for AI systems in production.
ML Platform Engineer
Build and maintain the ML platform that deploys, monitors, and scales AI models for Hadrian’s autonomous factories, using MLflow, Dagster, EKS, and FastAPI.
Principal ML Ops Engineer (EMEA Remote)
Build and operate scalable ML inference platforms for an AI-native cloud startup, designing GPU-powered serving systems, deployment pipelines, and observability for real-time AI applications.