Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Design, build, and optimize production-grade AI systems, specifically LLM-based agentic workflows on NVIDIA GPUs, using Python and modern frameworks.
About Neurophos The demand for new data centers and AI compute is rapidly outpacing the planet's energy capacity. Digital solutions are hitting a power wall as we approach the physical limits of traditional…
Build low-level firmware and drivers for real-time camera and lidar pipelines on autonomous rail vehicles, optimizing Linux/ARM/NVIDIA Jetson stacks for latency and throughput.
Develop and maintain production ML/computer vision systems that sort thousands of batteries per hour using PyTorch and OpenCV, integrating with factory hardware while monitoring model performance and collaborating across engineering and manufacturing teams.
Machine Learning Engineer at Netskope AI Labs building, optimizing, and deploying enterprise-scale AI solutions, particularly focusing on LLM inference optimization using technologies like vLLM, SGLang, and KV Cache optimization.
Lead the AI inference and optimization layer for agentic workflows, tuning models and runtimes for latency, throughput, and memory efficiency on real hardware within a cloud security company.
About the Team Our DD Labs team builds real-time autonomous delivery systems. The Planning & Decision-Making group is investing heavily in deep reinforcement learning to move beyond classical planning, learning…
Build and scale the shared infrastructure that powers DoorDash’s generative AI products, including real-time and batch LLM inference, fine-tuning, and agent platforms across San Francisco, Sunnyvale, and Seattle.
The Role We are looking for a Machine Learning Engineer to help build cutting edge ML infrastructure for building and serving LLM’s at Moveworks. This role will be critical in building, optimizing and scaling…
Build the onboard perception system for an all-electric air taxi, fusing camera, lidar, and radar data to detect obstacles and enable precision landings in real time.
VinFast is a pioneering electric vehicle (EV) company committed to revolutionizing the automotive industry with sustainable and innovative mobility solutions. As a leading player in the EV market, VinFast is dedicated…
Builds and scales a high-throughput AI inference API in Rust, owning model serving, routing, and observability for global developer access.
Builds and optimizes high-performance AI inference systems in C++ to serve millions of users with low latency and high reliability, using GPU engines like vLLM and Triton.
Job Title: Software Engineer II- Robotics & Edge AI Location(s): Tustin, CA Compensation: $95,000- $113,000 About this position : We are looking for a Software Engineer – Robotics & Edge AI Software Development…
Builds and deploys small language models that enforce real-time AI governance policies, undoing harmful agent actions while optimizing for latency and cost in enterprise environments.
Leads the integration and deployment of advanced perception models into production ADAS and autonomous driving systems, optimizing performance on automotive-grade hardware using C++, Python, CUDA, and TensorRT.
Build and deploy production AI systems for customers by designing end-to-end workflows, integrating AI infrastructure, and collaborating with engineering/product teams using Python, Go, React, and cloud-native tools.
Build and optimize PyTorch-based training and post-training pipelines for large language models, improving model quality and training efficiency while collaborating with research and infrastructure teams.
Engineer at VinFast's Speech & Language Processing Center building and fine-tuning ASR models (Vietnamese, English, Bahasa) for cloud and on-device (Android/embedded) use in vehicles. Day-to-day: train/optimize speech models, build ML pipelines and serving infra, and optimize inference with tools like ONNX Runtime, TensorRT, and CUDA.
Build and optimize LLM-based AI agents for VinFast's SLP Center in Hanoi — designing multi-agent orchestration, RAG pipelines, function calling, and evaluation frameworks, plus fine-tuning and inference optimization. Core stack: Python, LangChain/LangGraph, PyTorch, Hugging Face, vLLM, Docker, and cloud platforms.
We couldn't check your fit for this role — add a CV to your profile to see it next time.