Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build and operate the production ML/LLM platform for healthcare workflows, including training, deployment, monitoring, and compliance systems on GCP.
Designs and optimizes large language models and reasoning pipelines for formal code verification, using Python/C++ and PyTorch/JAX in a San Francisco office.
Designs and refines ML data pipelines for distributed training and validation of AI models that enable formally verifiable software, using Python/C++ and frameworks like PyTorch/TensorFlow.
Build and scale ML training and inference infrastructure for autonomous robots, focusing on data pipelines, distributed training, and GPU optimization.
Build and maintain the distributed compute and orchestration platforms powering large-scale ML training and simulation for AI game agents.
Optimize distributed AI training for 100B+ parameter models across 100k+ GPUs, writing CUDA/Triton kernels and co-designing hardware-efficient training recipes.
Build and optimize large-scale feed recommendation systems using ML to improve user engagement and retention for a social media platform.
Build and scale distributed training infrastructure for large-scale AI avatar models, optimizing GPU clusters, parallelism, and multimodal data pipelines for real-time, full-duplex training.
Own the design, training, and deployment of a novel foundation model, including custom CUDA kernels, distributed training pipelines, and cloud-based scaling.
Engineer large-scale LLM pre-training pipelines and distributed GPU clusters using PyTorch, DeepSpeed, or Megatron-LM, optimizing networking, memory, and fault tolerance for month-long runs.
Lead cross-functional programs to scale AI server infrastructure from prototype to volume production, coordinating hardware design, ODM partners, and supply chains to deliver production-ready systems for large-scale AI workloads.
Technical Program Manager at an AI startup leading cross-functional programs to build and deploy custom AI servers, coordinating hardware, software, and supplier teams.
Senior engineer building and optimizing open-source AI frameworks (Megatron Core, NeMo) for large language and multimodal model training, fine-tuning, and deployment on NVIDIA GPUs.
Build and deploy AI models for defense/intelligence missions using Python and PyTorch, from research to production deployment in Linux environments.
Job Description Operates & Maintains a Linux based Distributed Training System Simulation & Stimulation (DTS Sim/Stim) to support Warfighter training. Work location is CCDC (HSV) Redstone Arsenal Roles and…
Designs and maintains GPU clusters, Kubernetes, and high-speed networking/storage for AI/ML workloads, using Terraform, Ansible, and observability tools.
Build and maintain cloud infrastructure, CI/CD pipelines, and Kubernetes clusters to support robotics software, ML workflows, and embedded deployments for a Physical AI startup.
Lightning AI seeks a Senior Software Engineer to design and scale backend systems powering AI development. You will work across distributed training infrastructure, workload orchestration, experiment management,…
The Company: Faraday Future is a California-based technology company focused on the design, engineering, and development of intelligent, connected electric vehicles and related artificial intelligence–enabled…
Build and deploy deep-learning computer-vision systems using CNNs and Vision Transformers for defence and security applications.
We couldn't check your fit for this role — add a CV to your profile to see it next time.