Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Designs and builds an enterprise-scale Agentic AI platform that lets engineering teams create, deploy, and run secure, governed AI agents using cloud-native tools and AI frameworks.
Designs and deploys private AI systems for document intelligence and legal automation using LLMs, RAG, and GPU infrastructure in a law firm setting.
Develops topology-aware collective algorithms and transport layers for AI hardware, optimizing multi-node performance and co-designing with hardware teams.
Optimize and deploy large language models for fast inference on Amazon’s Trainium accelerators using AWS Neuron’s software stack.
Build secure, AI-powered web apps for a defense contractor, integrating LLMs and RAG pipelines with Python backend, React/Vue frontend, and vector databases in a classified environment.
Build and deploy AI/ML models for Yahoo Mail’s personalization and NLP systems, leveraging PyTorch/TensorFlow and cloud-native architectures at massive scale.
Build and deploy AI/ML systems that personalize Yahoo Mail for 220M+ users, using NLP, GenAI, and agentic workflows while optimizing for scale and latency.
Machine Learning Engineer - IT Location: Bogotá,Colombia Department: IT Work Location: hybrid Como ML Engineer en Mercado Libre, diseñarás y escalarás sistemas innovadores y seguros que resuelven problemas reales y de…
Build and operate internal AI tools (chat, agents, RAG, sandboxes) on GCP/GKE/Vertex AI for engineers, while ensuring compliance with regulated environments like CMMC and FedRAMP.
Algorithm - Serving System Engineer Department: Algorithm Location: Seoul HQ Employment Type: FullTime About the Role 우리 팀은 NPU 기반 inference의 roofline 자체를 개선하는 방법을 연구합니다. 6개월~2년 뒤에 제품에 적용될 수 있는 도전적인 과제를 스스로 정의하고 탐색하며,…
AI Operations Architect & AWS Engineer, Life Sciences Company: Norstella Location: Remote, United States Date Posted: Jul 27, 2026 Employment Type: Full Time Job ID: R-2070 **Description** **Job Description - AI…
Principal engineer builds and scales the AI inference software stack for d-Matrix’s custom compute engine, optimizing hardware-software co-design and collaborating across ML, compilers, and hardware teams.
Build and optimize Cerebras' GPU-based AI inference stack, integrating vLLM, ROCm, and AMD GPUs to deliver ultra-low-latency, high-throughput LLM serving for production workloads.
Description Altamira brings a commercial mindset to solving the most complex national security problems by delivering mission application development, multi-intelligence analysis, and data science technologies and…
Designs and builds real-time, multimodal data pipelines that enrich raw data with ML models to power Apple’s customer-facing services at massive scale.
Build and optimize AI model inference systems to serve millions of users with low-latency GPU workloads using frameworks like vLLM or Triton.
Develops and trains multilingual, multicultural large language models (LLMs) and related AI products, focusing on underrepresented languages and real-world applications.
Builds and runs evaluation frameworks for large language models, focusing on multilingual and multimodal capabilities, using Python and PyTorch.
Build and scale the AI infrastructure that powers CrowdStrike’s LLM-driven security products, including GPU clusters, model-serving pipelines, and MLOps tooling.
Lead co-design of production agentic AI systems with enterprise partners, building RAG, multi-agent, and LLM workflows on NVIDIA’s stack (NeMo, TensorRT-LLM) and shaping product roadmaps.
We couldn't check your fit for this role — add a CV to your profile to see it next time.