Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Design, build, and optimize production-grade AI systems, specifically LLM-based agentic workflows on NVIDIA GPUs, using Python and modern frameworks.
About Neurophos The demand for new data centers and AI compute is rapidly outpacing the planet's energy capacity. Digital solutions are hitting a power wall as we approach the physical limits of traditional…
About US:- We turn customer challenges into growth opportunities. Material is a global strategy partner to the world’s most recognizable brands and innovative companies. Our people around the globe thrive by helping…
Builds and optimizes high-performance AI inference systems in C++ to serve millions of users with low latency and high reliability, using GPU engines like vLLM and Triton.
Research Engineer at Lightning AI to optimize AI training and inference workloads on PyTorch Lightning infrastructure, improving performance and scalability for production systems.
Build and optimize PyTorch-based training and post-training pipelines for large language models, improving model quality and training efficiency while collaborating with research and infrastructure teams.
Engineer at VinFast's Speech & Language Processing Center building and fine-tuning ASR models (Vietnamese, English, Bahasa) for cloud and on-device (Android/embedded) use in vehicles. Day-to-day: train/optimize speech models, build ML pipelines and serving infra, and optimize inference with tools like ONNX Runtime, TensorRT, and CUDA.
Senior DevOps/MLOps Engineer at VinFast (EV maker) who builds and runs the infrastructure behind its Agentic AI and VoiceAI systems: deploying LLM model serving on AWS EKS with GPU optimization (vLLM, Triton), plus IaC, CI/CD, auto-scaling and observability for ultra-low-latency production AI services.
Designs and deploys high-performance GPU clusters for AI workloads, owning the full architecture cycle from customer requirements to production deployment.
Alice is hiring an Engineering Manager to lead its Data Science group, which builds GenAI safety models (content moderation, automated LLM red-teaming) and the Databricks/Spark data platform behind them. The role is hands-on people leadership: setting technical direction, owning delivery, and growing a team working in Python, Rust, and PySpark on AWS/Kubernetes.
Contractor role at RavenPack (Marbella, remote with EU timezone) doing hands-on optimization of the search & LLM stack: fine-tuning open-source LLMs (PEFT/LoRA, DPO), inference acceleration (quantization, vLLM, Triton), retrieval improvements, and reproducible delivery via SageMaker and Docker. Initial 3-6 month contract.
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment,…
Senior ML engineer optimizes PyTorch models for edge deployment, iterating from training to production-ready releases under tight latency and memory constraints for autonomous-vehicle systems.
Clinic Manager - Fraser Street, Tauranga At Triton Hearing, we put people at the heart of everything we do and push ourselves to innovate and think differently. About Us At Triton Hearing we’re changing hearing, for…
Senior MLOps Solutions Engineer on the Pure Solutions team who designs and automates end-to-end MLOps pipelines and GPU-accelerated AI/ML reference architectures, integrating Pure Storage platforms (FlashBlade, FlashArray, Portworx) with Kubeflow, MLflow, and Ray. Day to day: CI/CD-driven pipeline automation, Terraform/Ansible-based infrastructure, and optimizing LLM inference with tools like NVID
About Platform Media Platform Media is an entertainment company creating shows, channels and brands across video, social and beyond. It builds worlds around the people, stories and ideas audiences care about most,…
Junior machine learning engineer building and improving computer vision models (YOLO-based object detection/classification) that turn store shelf photos into merchandising insights like SKU detection and planogram compliance. Day to day: training and fine-tuning models, tracking experiments in MLFlow, managing labelling pipelines, and serving models via Triton and Flask.
About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform,…
Get to Know the Team The AI Platform (AIP) team builds and operates the core ML and AI infrastructure that powers Grab. Our stack spans model serving, ML pipelines, data serving, AI infrastructure, and Applied…
We couldn't check your fit for this role — add a CV to your profile to see it next time.