Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Data Center AI Systems Engineer (Competitive Analysis) - Riyadh, KSA Location: Riyadh, Saudi Arabia Department: Systems Engineering Work Location: onsite Company: Qualcomm Middle East Information Technology Company LLC…
Designs and builds an enterprise-scale Agentic AI platform that lets engineering teams create, deploy, and run secure, governed AI agents using cloud-native tools and AI frameworks.
Design and evaluate MLOps systems and GPU kernel optimizations for next-gen AI models using JAX, PyTorch, and Triton, collaborating with research teams on large-scale training infrastructure.
Senior engineer optimizing generative AI models (LLMs, diffusion) for inference efficiency using PyTorch, CUDA, TensorRT, and Triton, and deploying them at scale.
Build and deploy AI/ML models for Yahoo Mail’s personalization and NLP systems, leveraging PyTorch/TensorFlow and cloud-native architectures at massive scale.
Leads Developer Relations for robotics and edge AI in India, building demos, running workshops, and guiding teams to prototype and deploy AI-powered systems using NVIDIA’s platforms like Isaac Sim, Jetson, and TensorRT.
Leads strategic AI adoption for India's major conglomerates by engaging executives and technical teams to identify, build, and scale enterprise AI workloads using NVIDIA's platforms and software.
Build and deploy AI/ML systems that personalize Yahoo Mail for 220M+ users, using NLP, GenAI, and agentic workflows while optimizing for scale and latency.
Lead cybersecurity governance, risk, and compliance for the Triton unmanned aircraft system, ensuring Defence ATO and security documentation meet strict standards.
Algorithm - Serving System Engineer Department: Algorithm Location: Seoul HQ Employment Type: FullTime About the Role 우리 팀은 NPU 기반 inference의 roofline 자체를 개선하는 방법을 연구합니다. 6개월~2년 뒤에 제품에 적용될 수 있는 도전적인 과제를 스스로 정의하고 탐색하며,…
Research engineer optimizing large-scale generative AI training systems, focusing on performance, stability, and low-precision techniques for diffusion and multimodal models.
Build and optimize Cerebras' GPU-based AI inference stack, integrating vLLM, ROCm, and AMD GPUs to deliver ultra-low-latency, high-throughput LLM serving for production workloads.
Designs AI-driven systems that recursively optimize compute workloads and hardware, using agentic loops, performance profiling, and reinforcement learning to improve correctness, speed, and efficiency.
Neurosoft Bioelectronics is developing next-generation AI for decoding neural time series (high-density subdural ECoG LFPs). We seek an intern machine learning scientist with interest in sequence modelling, state-space…
Build and optimize cutting-edge large language models and foundation models using decentralized learning methods, focusing on post-training, distributed GPU training, and open-source deployment.
Develops and optimizes AI models for satellite and aerial imagery using generative techniques like Diffusion and GANs, focusing on super-resolution and image translation, and deploys them in on-premise GPU environments.
Designs and builds real-time, multimodal data pipelines that enrich raw data with ML models to power Apple’s customer-facing services at massive scale.
Build and optimize AI model inference systems to serve millions of users with low-latency GPU workloads using frameworks like vLLM or Triton.
Build and scale the AI infrastructure that powers CrowdStrike’s LLM-driven security products, including GPU clusters, model-serving pipelines, and MLOps tooling.
Build and deploy AI models for 3D dental scans using PyTorch/TensorFlow to power remote orthodontic monitoring used by 10,000+ clinicians worldwide.
We couldn't check your fit for this role — add a CV to your profile to see it next time.