Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build and train ML models that power real-time quantitative trading systems, working with PyTorch/JAX and distributed compute platforms in a fast-paced finance setting.
Build and deploy NLP pipelines using LLMs for contact-center insights, including text classification, summarization, and agentic applications in a fast-moving AI startup.
Own and optimize the CI/CD infrastructure for SGLang, an open-source LLM inference engine, ensuring fast, reliable, and secure test pipelines across multiple GPU hardware pools.
Build and maintain ML pipelines, deploy models via APIs/containers, and monitor performance for government AI projects using cloud-native tools.
Research and build next-gen AI world models that predict and interact with the real world, focusing on conditioning, control, and scalable training/inference systems.
Build and maintain AI-driven trading models and data pipelines for a quantitative hedge fund, translating research into production systems.
About VESSL AI AI researchers spend more time fighting for compute than doing research. Quota walls, rigid contracts, and fragmented supply across clouds and clusters mean teams are stuck waiting on a single provider…
Build and scale ML infrastructure for brain-computer interface R&D, including distributed training pipelines and large-scale data platforms to support neuroscientific modeling and neural decoding.
Technical Program Manager at RadixArk coordinates large-scale AI infrastructure programs, including inference engines, training frameworks, and GPU integration, to deliver systems serving billions of tokens daily.
Designs RadixArk's visual identity and brand system, creating illustrations and collateral using AI tools and code, while collaborating with cross-functional teams to shape the company's public presence.
Product Manager at RadixArk in Palo Alto defining and driving the AI infrastructure product roadmap, collaborating with engineering teams, and engaging with developer communities to shape open AI systems.
About the Role RadixArk is seeking a Product Marketing Manager to own how SGLang, Miles, and our open source infrastructure are positioned and perceived across the market. SGLang already has 20K+ GitHub stars and…
Designs and scales high-performance GPU/TPU clusters for AI training and inference, focusing on distributed systems, scheduling, and performance optimization.
About the Role We're looking for a Head of Business Development to build the BD function at RadixArk from the ground up. The BD team is the institutional memory of this company — maintaining active relationships…
Builds and optimizes high-performance ML systems for TPU hardware using JAX, XLA, and Pallas, focusing on inference/training workloads and compiler/runtime tuning.
Builds and optimizes large-scale AI inference systems for frontier models, focusing on performance, latency, and cost across thousands of GPUs.
Lead Chewy’s Ads Data Science team to build AI-powered ranking, bidding, and optimization systems for onsite and offsite ads, shaping the future of retail media with LLMs and generative AI.
Build and deploy AI models that turn satellite and geospatial data into actionable insights using vision-language models, multimodal systems, and reasoning models in Python with PyTorch/TensorFlow.
Research and develop multilingual large language models from scratch, including data curation, model architecture, and distributed training on GPU clusters.
Performance Engineer at RadixArk in Palo Alto optimizes LLM inference and training systems for latency, throughput, and cost efficiency across production workloads using SGLang, Miles, and GPU/TPU infrastructure.
We couldn't check your fit for this role — add a CV to your profile to see it next time.