Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Principal engineer optimizing large-language-model inference at a major bank, designing quantization, speculative decoding, and benchmarking strategies to cut cost and latency for production AI workloads.
The Platform Engineer III will build and maintain AI infrastructure on Google Kubernetes Engine (GKE) to support Generative AI applications. The role involves implementing CI/CD, GitOps, and observability standards while collaborating with data science teams to optimize LLM serving and cloud-native tooling.
The GPU Software Engineer will design and optimize high-performance CUDA kernels for AI and scientific computing workloads. The role involves profiling GPU code, collaborating with ML teams to improve training and inference pipelines, and working with modern accelerator hardware.
An AI Forward Deployed Engineer embedding with enterprise clients to build production-grade applications on proprietary foundation models, involving RAG pipelines, model fine-tuning, agentic workflows, and inference optimization using Python, PyTorch, vector databases, and cloud infra.
At the National Robotics Engineering Center (NREC), it is our engineers and technicians who drive the breakthroughs that define our success. The members of our technical staff collaborate closely with leadership and…
Senior Staff ML Engineer building and scaling ML infrastructure for LLM training, evaluation, and deployment at Moveworks (ServiceNow), working with PyTorch, vLLM, TensorRT-LLM, Python, and C++/GoLang.
Build and scale data pipelines that tag petabytes of autonomous-truck logs, turning raw sensor data into curated driving scenarios for ML and simulation teams using Python, Databricks, AWS, and CI/CD.
Senior Data Scientist building AI retrieval and knowledge systems for rare disease research at the NIH, working across the full stack from PostgreSQL vector search and LLM application engineering through Next.js/React interfaces and Kubernetes deployment.
Build and deploy agentic AI systems for rare disease research, designing multi-turn LLM workflows that validate structured outputs and log decisions for review by clinicians and researchers.
As a founding AI engineer, you will build an efficiency layer for production AI systems by developing infrastructure to optimize token usage, latency, and costs. You will work directly with founders to design and deploy scalable backend systems, APIs, and evaluation frameworks for LLM-based applications.
Senior engineer designs and maintains Azure cloud infrastructure and deploys AI/ML workloads for an enterprise client, using AKS, Azure OpenAI, and Python/React stacks.
GalaxEye is seeking a backend engineer to build and maintain ML platforms that process satellite imagery in air-gapped, offline environments. You will design data pipelines and deploy self-hosted ML models to provide geospatial intelligence for defense and intelligence applications.
Ключевые задачи: Разбирать запросы заказчиков и предлагать, каким классом решений их закрывать; Разрабатывать и улучшать RAG-пайплайны: поиск, реранжирование, сборка контекста, цитирование; Поднимать и оптимизировать…
First AI engineer building a production AI efficiency layer (token reduction, cost optimization, context management) using Python/TypeScript, LLM APIs, and cloud infrastructure—on-site in Berlin five days a week.
Founding AI Engineer building the core efficiency layer for production AI systems—reducing token usage and costs via context optimization, caching, and model routing—using Python/TypeScript, cloud infrastructure, and LLM APIs.
Founding AI engineer building a production AI efficiency layer—working on token reduction, context optimization, caching, model routing, and an OpenAI-compatible gateway using Python/TypeScript, LLM APIs, PostgreSQL, Docker, and cloud infrastructure.
Salary: £ 65,360 - £81,700 p.a. Closing date: Tuesday 8 September 2026 Contract type: Permanent Interviews: w/c 21 September 2026 The Wellcome Trust is a global charitable foundation. We improve health for everyone by…
Senior AI/backend engineer builds and scales Python-based FastAPI services, RAG pipelines, and LLM agent systems on AWS/Azure, with cloud-native tooling and security best practices.
KoronaTech – это популярный, интенсивно развивающийся онлайн-сервис денежных переводов и смарт-займов на любые товары и услуги. Благодаря нашим продуктам пользователи в десятках стран мира могут быстро, удобно и…
Вам предстоит заниматься: Вывод ML/LLM-моделей в промышленную эксплуатацию: упаковка, деплой и сопровождение инференс-движков (vLLM, Triton Inference Server, BentoML) на Kubernetes (CPU/GPU); развитие и поддержка…
We couldn't check your fit for this role — add a CV to your profile to see it next time.