Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Who we are looking for: As an Engineer IV, Embedded Systems within the R&D team, you will serve as a technical anchor for our next-generation hardware platforms. You will design, develop, and optimize…
Builds and optimizes large-scale AI infrastructure, including Kubernetes clusters, RDMA networking, and GPU orchestration to improve efficiency and scalability of AI training/inference systems.
Optimize and deploy large-scale on-prem LLM inference systems using NVIDIA H200 GPUs, vLLM, TensorRT-LLM, and OpenShift AI for enterprise private GenAI environments.
Build and deploy enterprise LLM applications, RAG systems, and AI agents using open-source models (DeepSeek, Qwen, Kimi) and frameworks like LangChain and vLLM.
Engineer AI/LLM inference on GPU clusters: benchmark, tune, and optimize model serving with vLLM, Triton, or TensorRT-LLM to hit latency, throughput, and memory targets.
Leads ML/CV development for autonomous navigation and computer-vision systems in unmanned aerial vehicles, optimizing algorithms for embedded Linux and sensor fusion.
Research and develop on-device generative and 3D spatial AI models for mobile platforms, optimizing models for edge deployment and shipping to product.
Build and operate a unified GPU compute platform for AI training and inference, hiding cloud complexity behind Kubernetes operators and schedulers that manage multi-cloud NVIDIA clusters.
Build and deploy ML models for media, ad-tech and e-commerce: recommendations, real-time personalization, pricing and bidding engines that directly lift CTR, eCPM and revenue.
Machine Learning Engineer at Augur developing advanced computer vision capabilities including scene understanding, object detection, tracking, and 3D reconstruction from edge-deployed sensors. Will evaluate and integrate state-of-the-art research, build efficient inference pipelines, and deploy production-grade CV models using Python, PyTorch, and related technologies.
Build and optimize real-time graphics and ML inference pipelines for interactive visual apps, integrating super-resolution and denoising models into DX12/Vulkan renderers.
Build and optimize computer-vision services (OCR, detection, classification) in Python, tune models for production, and collaborate with labeling teams and business stakeholders.
Software Engineer - AI Location: Gurugram, India Department: Engineering At Anaplan, we are a team of innovators focused on optimizing business decision-making through our leading AI-infused scenario planning and…
Builds and evolves RackAI, a cloud-native AI platform, using Go and Kubernetes to design scalable backend services, operators, and AI-driven tooling while shaping an AI-first engineering culture.
Build and maintain data pipelines for AI systems, integrate ML components, and optimize inference workflows in Python.
Principal AI Engineer designs and owns the shared AI architecture for multi-agent marketing systems, retrieval pipelines, and evaluation frameworks that power SMB-focused products at scale.
Build and optimize on-device AI inference software for NVIDIA GPUs, focusing on low-latency, memory efficiency, and deployment on RTX/DGX systems.
Build and maintain scalable CI/CD infrastructure for NVIDIA’s TensorRT Edge-LLM, automating builds, tests, and deployments across embedded and cloud platforms using tools like GitLab, Kubernetes, and Docker.
Research engineer translating biometrics and AI research into scalable production software using ML/DL, computer vision, C++, and Python to build biometric recognition systems at Thales.
Designs and implements deep learning models for autonomous truck perception, planning, and control, using PyTorch/TensorFlow and state-of-the-art neural architectures.
We couldn't check your fit for this role — add a CV to your profile to see it next time.