Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build and optimize scalable generative AI inference pipelines and APIs for Adobe’s Firefly suite, integrating models like diffusion transformers into Photoshop, Illustrator, and other products.
Memwize is a VC-backed, stealth-mode startup building rack-level AI inference systems . We are seeking a Senior/Principal Machine Learning Compiler Developer to build compiler capabilities from PyTorch through Triton…
Lead AI inference optimization and deployment for autonomous-vehicle models, tuning low-latency pipelines on embedded and cloud GPUs with CUDA, TensorRT, and LLM serving stacks.
You will own complex escalated customer issues and platform incidents from investigation through resolution. You will troubleshoot GPU compute, networking, storage, drivers, CUDA, control plane, and billing issues;…
Serve as the frontline for a public-cloud GPU platform by handling customer tickets, monitoring platform health, responding to alerts, troubleshooting infrastructure and service issues, escalating incidents,…
Overview At Zebra, we are a community of innovators who come together to create new ways of working. United by curiosity and a culture of caring, we develop smart solutions that anticipate our customer’s and partner’s…
Optimize and port AI inference/training kernels across NVIDIA, AMD, TPUs, and emerging accelerators, designing portable abstractions for SGLang and Miles.
Who we are Aurora’s mission is to deliver the benefits of self-driving technology safely, quickly, and broadly. The Aurora Driver will create a new era in mobility and logistics, one that will bring a safer, more…
Develops high-performance kernels, compilers, and communication libraries to optimize AI workloads on GPU clusters, focusing on low-latency and memory efficiency.
Port, optimize, and parallelize HPC simulation codes for government modeling workflows; mentor teams and deliver training on HPC tools and best practices.
Build and optimize systems to deploy large language models like Gemini into production, profiling and accelerating inference pipelines on TPUs/GPUs for global-scale AI applications.
Design, deploy, and maintain large-scale GPU clusters for AI/ML workloads, automate provisioning, and optimize performance using Kubernetes, Slurm, and NVIDIA tools.
Design and implement cloud infrastructure and MLOps solutions for AI teams, advising clients on GPU cloud technologies and optimizing ML pipelines using tools like Terraform, Kubernetes, and Python.
Musashi AI North America, Inc. is a growing hardware and software focused company that builds smart vision solutions for quality assurance in manufacturing environments. Based in Waterloo, Ontario, Musashi AI North…
Build and maintain ML infrastructure to deploy and scale AI models in production, using Python, Kubernetes, and Ray. Own platform services that enable data scientists to move models from experimentation to live systems.
Engineer and operate large-scale GPU-accelerated HPC clusters for AI/ML workloads, focusing on automation, performance tuning, and reliability across on-prem and cloud environments.
Ищем специалиста для разработки и интеграции алгоритмов навигации и восприятия для автономных беспилотников. Ты будешь работать на пересечении робототехники, компьютерного зрения и прикладной математики. Обязанности…
Build and deploy a data-science system that powers UK property insights, using Python, Azure, Docker, and PyTorch.
Build and optimize the MLOps platform powering Revolut’s AI products, focusing on scalable Python backend services, GPU workloads, and PyTorch-based model training/inference.
Own and optimize the CI/CD infrastructure for SGLang, an open-source LLM inference engine, ensuring fast, reliable, and secure test pipelines across multiple GPU hardware pools.
We couldn't check your fit for this role — add a CV to your profile to see it next time.