Tech jobs
Job listings
Senior Machine Learning Engineer
Build and optimize deep-learning models for autonomous-driving perception and planning, using PyTorch and model-efficiency techniques to speed up real-world deployment.
Senior Machine Learning Engineer
Senior ML Engineer builds and scales computer-vision models for biometric authentication on AWS, using PyTorch/TensorFlow and optimizing low-latency inference pipelines.
Binance Accelerator Program - LLM Recommendation & Agentic AI Engineer
Build LLM-powered recommendation systems and agentic AI for crypto trading and Web3 use cases using PyTorch, RAG, and reinforcement learning.
GPU Software Engineer (CUDA)
Build and optimize CUDA kernels for AI/HPC workloads, tuning GPU memory and compute to maximize throughput in production systems.
ML Platform Engineer
Design and operate high-performance ML inference platforms for serving large models in production, optimizing latency, throughput, and GPU utilization while ensuring reliability and observability.
AI Engineer
Build and deploy production AI systems for a global bio-resources company, including GenAI pipelines, document processing, and computer vision for industrial operations.
Machine Learning Engineer, AI Inference Solutions (University Grad)
Build and optimize ML deployment platforms and inference pipelines for autonomous-vehicle software, shipping PyTorch models to GM’s Super Cruise fleet with real-time latency and safety constraints.
Senior Product Manager – AI Inference Performance
Own the AI inference performance roadmap at NVIDIA, turning deep optimization techniques into products that improve latency, efficiency, and cost per token across the inference stack for LLM deployments.
Senior Manager, Software Engineering - NIM Factory
NVIDIA is the platform upon which every new AI‑powered application is built! We are seeking a deeply technical, hands‑on Senior Engineering Manager to lead the NVIDIA Inference Microservices (NIM) Factory team. You…
Staff Machine Learning Engineer
Develops and optimizes 2D traffic sign detection models for autonomous driving using PyTorch, TensorRT, and ONNX, from data curation to edge deployment.
Machine Learning Engineer
Build and optimize scalable generative AI inference pipelines and APIs for Adobe’s Firefly suite, integrating models like diffusion transformers into Photoshop, Illustrator, and other products.
Staff Software Engineer, ML Acceleration
Builds and optimizes machine-learning training pipelines for autonomous-vehicle models, focusing on GPU acceleration and deployment across hardware platforms using PyTorch, CUDA, and Triton.
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Lead AI inference optimization and deployment for autonomous-vehicle models, tuning low-latency pipelines on embedded and cloud GPUs with CUDA, TensorRT, and LLM serving stacks.
Senior Machine Learning Engineer
Build and scale computer-vision models for biometric identity verification, training and deploying PyTorch/TensorFlow pipelines on AWS with low-latency inference.
Machine Learning System Software Engineer
Develop low-power, high-performance AI system software for Apple’s Neural Engine, optimizing ML models for devices like Vision Pro and iPhone.
Machine Learning Performance Engineer - Offboard Training & Inference
Optimize large-scale ML training and batch inference for autonomy workloads, profiling GPU clusters to maximize throughput and reduce cost per unit of data processed.
Runtime Engineer
What MatX is Building MatX is building custom silicon for large-language-model inference and training, with HW/SW co-design across ISA, RTL, simulator, compiler, and kernels so each layer benefits from the others. The…
Senior AI Compute Infrastructure Engineer
Engineer and optimize GPU clusters for AI training and inference, build scheduling/orchestration systems, and improve performance, cost, and reliability of AI infrastructure.
AI Computer Vision Engineer
Develop real-time object detection and tracking models for UAV aerial video, optimizing inference pipelines for edge hardware like NVIDIA Jetson.