Tech jobs
Job listings
AI Engineer
Build and deploy scalable AI systems that turn physical-world data into enterprise intelligence, focusing on energy efficiency and sustainability across industries.
AI Solutions Engineers - Pre-Sales
Roles & Responsibilities AI Solutions Engineer About the Role We are seeking a dynamic AI Solutions Engineer to join my client's Enterprise Sales team. In this role, you will be the technical authority guiding…
Senior Robotics Software Engineer
Lead the design and development of AI-driven robotics systems for Dyson’s vacuum robots, focusing on perception, sensor fusion, and adaptive cleaning behaviors using C++/Python and ML frameworks.
Senior Machine Learning Engineer
Senior ML Engineer builds and scales computer-vision models for biometric authentication on AWS, using PyTorch/TensorFlow and optimizing low-latency inference pipelines.
Software Engineer, AI Kernels & Performance Optimization — MTIA Software
Build and optimize low-level AI kernels (GEMM, attention, quantization) for Meta’s custom AI chips, shaping hardware-software co-design to maximize performance for billions of users.
AI Frameworks Software Engineer – Model Compression
Develops and optimizes model compression tools like Intel Neural Compressor for LLMs and generative models, focusing on quantization, pruning, and inference acceleration on Intel AI hardware.
ML Platform Engineer
Design and operate high-performance ML inference platforms for serving large models in production, optimizing latency, throughput, and GPU utilization while ensuring reliability and observability.
Embedded AI Engineer
Design and deploy machine learning models optimized for edge devices like mobile SoCs and embedded accelerators, using quantization, pruning, and hardware-aware tuning to meet latency, memory, and energy constraints.
Machine Learning Engineer, AI Inference Solutions (University Grad)
Build and optimize ML deployment platforms and inference pipelines for autonomous-vehicle software, shipping PyTorch models to GM’s Super Cruise fleet with real-time latency and safety constraints.
Senior Product Manager – AI Inference Performance
Own the AI inference performance roadmap at NVIDIA, turning deep optimization techniques into products that improve latency, efficiency, and cost per token across the inference stack for LLM deployments.
Staff Machine Learning Engineer
Develops and optimizes 2D traffic sign detection models for autonomous driving using PyTorch, TensorRT, and ONNX, from data curation to edge deployment.
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Designs and optimizes large-language-model inference systems, quantizing models, deploying LLMs, and scaling GPU-based pipelines for low-latency production use.
AI Systems Architect
Designs and optimizes AI model architectures for specialized hardware, mapping model requirements to efficient silicon implementations and improving performance and scalability.
Senior/Sr. Staff AI Infrastructure Engineer, Inference & Optimization
Lead AI inference optimization and deployment for autonomous-vehicle models, tuning low-latency pipelines on embedded and cloud GPUs with CUDA, TensorRT, and LLM serving stacks.
Applied Scientist, GenAI & ML Systems
Builds and deploys production-grade GenAI/ML systems (LLM/SLM optimization) for multi-agent architectures in cloud environments, focusing on agent orchestration, RAG pipelines, and inference efficiency.
Senior Machine Learning Engineer
Build and scale computer-vision models for biometric identity verification, training and deploying PyTorch/TensorFlow pipelines on AWS with low-latency inference.
Member of Technical Staff — Inference-Multi-Hardware
Optimize and port AI inference/training kernels across NVIDIA, AMD, TPUs, and emerging accelerators, designing portable abstractions for SGLang and Miles.
Machine Learning Performance Engineer - Offboard Training & Inference
Optimize large-scale ML training and batch inference for autonomy workloads, profiling GPU clusters to maximize throughput and reduce cost per unit of data processed.
ML/AI Developer
Builds and optimizes AI-powered legal RAG pipelines using Qdrant, hybrid search, and LLM providers (Anthropic, OpenAI, YandexGPT) to retrieve and rank relevant legal documents from a 23M+ corpus.
Runtime Engineer
What MatX is Building MatX is building custom silicon for large-language-model inference and training, with HW/SW co-design across ISA, RTL, simulator, compiler, and kernels so each layer benefits from the others. The…