Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Leads the build-out of an internal AI/LLM platform on GPU infra: evaluates open-weight models, designs inference routers and agent tooling, and deploys scalable, observable services for corporate users.
About us We are building AI systems that can reason, use tools, and complete meaningful work in the real world. Our team works across model post-training, reinforcement-learning infrastructure, large-scale training,…
Lead Maven AGI’s DevOps team to design, scale, and secure cloud and on-prem Kubernetes infrastructure for an enterprise AI platform serving autonomous customer support.
Lead the AI software stack, compiler, and platform roadmap for a next-gen AI infrastructure startup, defining toolchains, framework integrations, and runtime engines to optimize AI model execution on custom hardware.
Own the AI inference and software product roadmap, translating model innovations into technical requirements and serving-stack features while collaborating with engineering and GTM teams.
Build and maintain AWS and Kubernetes infrastructure that powers the organization's AI capabilities, ensuring reliability, security, and scalability for AI services.
Build core data-query infrastructure for AI workloads, focusing on multimodal data (images, video, text) and distributed systems using Rust/C++/Python/Go.
Lead the architecture and engineering of a bank-grade AI & Agentic Platform, designing distributed systems, agent protocols, LLM infrastructure, and cloud-native stacks to support secure, scalable agent workloads across the enterprise.
Build and maintain the ML platform that trains, deploys, and monitors AI/ML models across Wealthsimple’s products using Kubernetes, Ray, FastAPI, and MLflow.
Designs, builds, and maintains Kubernetes-based AI engineering platforms for secure, sovereignty-sensitive environments, focusing on GitOps, observability, and CI/CD for AI-assisted SDLC workflows.
Immediate need for a talented SRE – ML focus. This is a 06+months contract opportunity with long-term potential and is located in Austin,TX / Sunnyvale, CA(Remote). Please review the job description below and contact…
Develop and optimize deep-learning frameworks for AMD GPUs, focusing on GPU kernels, distributed inference, and compiler tech to improve training/inference performance.
Lead ROCm software validation for AMD Instinct GPUs, defining test architecture, CI/CD pipelines, and release gates for AI/ML and HPC workloads across multi-node systems.
Build and optimize the Triton compiler and runtime for AMD GPUs to enable scalable, distributed AI workloads across multi-GPU systems.
Develop high-performance Triton/Gluon GPU kernels for AI models, optimizing matmul, attention, and transformer layers to maximize throughput on AMD Instinct accelerators.
Design the visual identity and developer-facing interfaces for vLLM, an AI inference engine, from brand assets to product UIs and launch campaigns.
TECHNICAL MANAGER - GPU CLOUD & AI INFRASTRUCTURE Location: Singapore Employment Type: Full-Time, Permanent Monthly Salary: S$15,000-S$20,000 Reporting To: Head of GPU Cloud and AI Infrastructure Travel…
Build and optimize LM Studio’s inference runtime for on-device and cloud AI, integrating new engines and models while improving performance across CPU/GPU targets.
Build and optimize high-performance Triton/Gluon kernels for AI models on AMD GPUs, collaborating with compiler and hardware teams to maximize throughput and efficiency.
Evangelize Modular’s MAX AI inference and serving platform through technical content, benchmarks, and community engagement to help developers deploy models efficiently.
We couldn't check your fit for this role — add a CV to your profile to see it next time.