Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Senior AI/backend engineer builds and scales Python-based FastAPI services, RAG pipelines, and LLM agent systems on AWS/Azure, with cloud-native tooling and security best practices.
KoronaTech – это популярный, интенсивно развивающийся онлайн-сервис денежных переводов и смарт-займов на любые товары и услуги. Благодаря нашим продуктам пользователи в десятках стран мира могут быстро, удобно и…
Вам предстоит заниматься: Вывод ML/LLM-моделей в промышленную эксплуатацию: упаковка, деплой и сопровождение инференс-движков (vLLM, Triton Inference Server, BentoML) на Kubernetes (CPU/GPU); развитие и поддержка…
Build and own the infrastructure that integrates LLMs into Isomorphic Labs' drug discovery workflows — designing scalable, secure systems for agentic tooling, RAG, model serving, and evaluation frameworks. Core tech: Python, LLM serving stacks (vLLM), cloud infrastructure (IaC), and agentic/MCP tooling.
Мы строим систему управления операционными рисками экосистемы Сбера на основе агентного ИИ. Если тебе тесно в роли просто сильного разработчика, а хочется вести за собой команду единомышленников, реализуя проекты…
AI/ML Engineer II 3-8 Years Experience JOB SUMMARY We are seeking a highly capable AI/ML Engineer II to serve as a strong individual contributor within our Artificial Intelligence and Automation team. The ideal…
Build and integrate AI/LLM features (RAG, agents, search, automation workflows) into the My OPSWAT customer portal, owning the model lifecycle from data pipelines to self-hosted LLM deployment and evaluation.
Designs and sells high-performance GPU and HPC cloud solutions for AI workloads, translating customer needs into feasible architectures while collaborating with engineering teams.
About the Role VESSL AI는 AI 기업을 위한 GPU 인프라를 제공하는 GPU Cloud 기업입니다. 폭발적으로 커지는 AI 컴퓨팅 수요 속에서, GPUaaS(GPU-as-a-Service)를 통해 AI 기업들이 필요한 순간에 필요한 만큼의 GPU 컴퓨팅 파워를 안정적으로 확보하도록 지원합니다. 국내 Upstage, SqueezeBits, Holiday Robotics는…
About the Role VESSL AI의 Backend Software Engineer (Senior)는 GPU 클라우드 플랫폼의 설계와 구현을 리드합니다. VESSL은 여러 데이터센터와 클라우드에 걸쳐 H200·B300부터 GB300 NVL72, Vera Rubin에 이르는 최신 GPU 클러스터를 운영하며, 이를 효과적으로 활용하기 위해 Kubernetes 기반 컨테이너부터 VM,…
About the Role VESSL AI의 Backend Software Engineer 는 GPU 클라우드 플랫폼을 설계하고 구현합니다. VESSL은 여러 데이터센터와 클라우드에 걸쳐 H200·B300부터 GB300 NVL72, Vera Rubin에 이르는 최신 GPU 클러스터를 운영하며, 이를 효과적으로 활용하기 위해 Kubernetes 기반 컨테이너부터 VM, 베어메탈에 이르기까지…
Adentris is seeking a founding engineer to lead the full-stack development and architecture of their AI-powered healthcare compliance platform. You will build and scale data pipelines, manage private LLM inference, and define the technical direction for clinical data normalization and AI evaluation.
The Applied AI Engineer will build and deploy end-to-end AI features by integrating machine learning models into production systems. The role involves optimizing model performance, designing agent workflows, and ensuring reliability using technologies like Python, PyTorch, JAX, and LLMs.
Director-level AI Architect at PwC's Data & Analytics Advisory practice in Bengaluru, designing ML pipelines, LLM serving/GPU infrastructure, and LLMOps solutions for clients using cloud platforms and frameworks like MLflow, DeepSpeed, and LangChain.
Design and deliver production-grade ML and Generative AI systems—including LLMs, RAG, and Agentic AI workflows—using Python, agentic frameworks, and major cloud platforms as an individual contributor on Fortive's AI and Automation team.
GPU Systems Engineer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterprise solutions across the United States. This is a fantastic…
The Machine Learning Infrastructure Engineer will design and maintain high-performance inference platforms for large machine learning models, focusing on systems engineering tasks like request routing, autoscaling, and GPU optimization. The role requires expertise in Python, systems programming languages, and production-grade AI serving frameworks.
Optimize LLM/VLM inference performance on NVIDIA GPUs—profiling workloads, building/tuning CUDA kernels, and improving open-source inference engines like TensorRT-LLM and vLLM.
Build NVIDIA NIM's model customization and deployment lifecycle platform—designing fine-tuning pipelines (LoRA, quantization), evaluation harnesses, and compliance/attestation layers on top of LLM serving infrastructure.
We couldn't check your fit for this role — add a CV to your profile to see it next time.