Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Builds and deploys on-device AI prototypes using C++/Python, PyTorch, and low-level system tools for edge devices like smartphones and IoT.
Optimize and deploy AI models on Qualcomm’s low-power hardware by writing C/C++ kernels, transforming graphs, and applying quantization to boost inference speed and accuracy.
Research and develop efficient on-device AI inference systems, focusing on model compression, hardware-software co-design, and optimization for Qualcomm’s low-power AI accelerators.
Video Codec Design Engineer, up to Staff Location: Taipei, Taipei City, Taiwan Department: ASICS Engineering Work Location: onsite Company: Qualcomm Semiconductor Limited Job Area: Engineering Group, Engineering Group…
Optimize and deploy large language model inference systems for an AI-driven chip-design platform, focusing on throughput, latency, and cost-efficiency across multi-node GPU clusters.
Builds and operates the AI agent infrastructure that automates engineering tasks, runs closed-loop evaluations, and enables autonomous research loops for Nuro’s self-driving platform.
Senior AI Engineer at PwC’s India Acceleration Center, designing and running experiments with LLMs and AI agents to evaluate tools and produce client-ready insights using Python and cloud platforms.
Lead Developer Relations for NVIDIA’s AI platforms in India, engaging ML researchers and startups to adopt GPU-accelerated AI workflows and optimize model training and inference.
Designs and optimizes next-gen satellite ground systems, integrating AI/ML pipelines and automation to process and orchestrate data from advanced constellations for national security missions.
Build and deploy AI models for public-safety devices and cloud, optimizing computer-vision and speech/NLP systems on edge and backend.
Lead the development and optimization of AI/ML models (CNNs, Vision Transformers, LLMs) for AMD’s next-gen compute platforms, including compiler toolchains and performance tuning.
Principal-level engineer optimizing AI model training and inference performance, reliability, and efficiency across large language models, diffusion models, and recommendation systems using frameworks like PyTorch and TensorFlow.
Staff/Principal DevOps Engineer, AI Inference Location: Cambridge, MA USA Department: Software Your Impact at LILA The Staff/Principal DevOps Engineer - AI Inference will drive the design, implementation, and…
Lead end-to-end design and delivery of large-scale AI systems for federal programs, building Python-based LLM and generative AI solutions with cloud-native architectures and MLOps pipelines.
Multiverse Computing Multiverse Computing is a fast-growing deep-tech company founded in 2019 and recognized by CB Insights as one of the 100 most promising AI companies globally. We are the largest quantum software…
Optimize Morph’s inference stack—kernels, serving, routing, and autoscaling—to maximize speed, cost-efficiency, and reliability for open AI models.
Forward Deployment Engineering Manager Location: United States Department: Strategy & Ecosystem About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a…
Senior Machine Learning Engineer, ML Efficiency Team: IT Location: US Commitment: Full-time Workplace Type: remote This position is listed on behalf of a partner company, who manages all applications and next steps.…
Machine Learning Engineer - Large Language Models Team: IT Location: US Commitment: Full-time Workplace Type: remote This position is listed on behalf of a partner company, who manages all applications and next steps.…
AI Engineer, Sr 1 Team: IT Location: South Africa Commitment: Full-time Workplace Type: remote This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking…
We couldn't check your fit for this role — add a CV to your profile to see it next time.