Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Builds and optimizes multilingual speech-to-text models for on-device use, focusing on compression, loanword handling, and evaluation across languages for a government security mission.
Build and optimize Qualcomm’s AI Runtime SDK to deploy large generative models (LLMs, LVMs) efficiently on Snapdragon chips using C/C++ and Python.
Build and optimize Qualcomm’s AI Runtime SDK to run large language and vision models efficiently on Snapdragon chips using C/C++ and hardware accelerators.
Design and deploy AI-powered edge-to-cloud systems for smart cities and IoT, optimizing models for latency, power, and scale across Qualcomm’s hardware and cloud platforms.
Builds and optimizes AI model quantization tooling to deploy neural networks efficiently on Qualcomm’s Snapdragon platforms, using Python/C++ and frameworks like PyTorch.
Research and develop efficient generative AI models (LLMs, multi-modal) for on-device deployment, focusing on reasoning, RAG, and Agentic AI.
Working conditions Location: Copenhagen, Denmark (primarily on-site) Office attendance: 4 days per week Working language: English Work authorization: We are looking for candidates who already have the legal right to…
Build and deploy AI models into scalable services, turning research prototypes into production-ready systems with Python, Airflow, and cloud/on-prem infrastructure.
Build and lead large-scale AI systems for federal programs, designing Python-based LLM solutions, RAG pipelines, and cloud-native architectures while ensuring compliance and security in regulated environments.
Maintains and optimizes a hybrid HPC/AI Linux cluster with GPUs, scheduling tools (SLURM/Kubernetes), and MLOps pipelines to support large-scale model training and inference for researchers.
Own the AI inference and software product roadmap, translating model innovations into technical requirements and serving-stack features while collaborating with engineering and GTM teams.
Build and operate AI model validation, quantization, and safety systems for Akamai’s Inference Cloud, including vulnerability scanning, optimization pipelines, and compliance guardrails.
ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And…
onepot is automating chemistry. Our goal is to enable a self-improvement loop for chemistry by combining AI and advanced robotics. In this loop, AI systems design experiments, robotic systems execute them, and the…
Evangelize Modular’s MAX AI inference and serving platform through technical content, benchmarks, and community engagement to help developers deploy models efficiently.
Write and optimize low-level compute kernels (matmul, attention, quantization) in C++23 for a custom RISC-V chip, using SIMD intrinsics and memory-hierarchy tuning to accelerate LLM inference/training.
Build full-stack AI apps for Micron’s smart factories: ingest sensor data, train models, ship React/FastAPI front-ends, and deploy digital twins in Gazebo/Unity for robotics and predictive maintenance.
Senior Computer Vision Engineer builds and optimizes ML SDKs, ports models to edge/cloud hardware, and automates deployment pipelines using C++ and Python.
Build and optimize NVIDIA’s deep-learning inference stack (TensorRT, TensorRT-LLM) by profiling models, writing GPU kernels, and integrating OSS frameworks to maximize GenAI performance across datacenter and edge GPUs.
Build and ship AI models for Berlitz’s online learning platform, focusing on speech, text, and computer vision to deliver real-time speaking practice, conversation, and feedback.
We couldn't check your fit for this role — add a CV to your profile to see it next time.