Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
- 5+ years of non-internship professional software development experience - 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience - 5+ years of…
Senior/Staff Machine Learning Engineer, Perception Team: Autonomy Software Location: South San Francisco, CA Commitment: Full Time Workplace Type: hybrid Salary: The US base salary range for this full-time position is…
Designs and builds a next-gen, real-time search platform at Apple scale, blending Generative AI with Information Retrieval to power user-facing experiences.
Principal ML Engineer New York, NY (Hybrid) About the Company We're building AI-native enforcement infrastructure for enterprise communication — technology that catches and fixes compliance issues in real time,…
Designs and deploys end-to-end AI infrastructure for enterprises, integrating NVIDIA GPU clusters, high-performance networking, parallel storage, and AI software stacks to support large-scale model training and inference.
AI Engineer, Platforms Location: NTU Main Campus, Singapore Time Type: Full time Job Description AI Singapore (AISG) is a national AI programme launched by the National Research Foundation (NRF), Singapore, to build…
ML/LLM engineer building AI agents and large language models for medtech products, automating medical workflows and assisting doctors and patients.
Lead a team building production-grade computer vision models in Python/PyTorch to improve workplace safety, owning the full ML lifecycle from data to deployment.
Data Center AI Systems Engineer (Competitive Analysis) - Riyadh, KSA Location: Riyadh, Saudi Arabia Department: Systems Engineering Work Location: onsite Company: Qualcomm Middle East Information Technology Company LLC…
Engineer deep-learning inference engines for GigaChat, optimizing CUDA kernels and GPU clusters to deliver sub-second LLM responses at massive scale.
Senior Staff ML Engineer in San Diego builds production-ready inference components from research prototypes, optimizing for latency, memory, and reliability in constrained local environments using Rust, C++, and CUDA.
Design and run AI hardware benchmarks for GPUs, TPUs and custom silicon, analyze throughput, latency and cost-per-token, and work with leading chipmakers to set industry standards.
Benchmark and analyze AI model inference performance across providers, extending methodologies for speed, cost, and accuracy while shaping industry standards.
Design and run AI hardware benchmarks (GPUs, TPUs, custom silicon), analyze inference economics, and partner with chipmakers to set industry standards for performance and cost-per-token.
Develop underwater computer-vision algorithms for fish-farm monitoring using deep learning, 3D vision, and edge deployment on Jetson/GPU hardware.
Build and optimize a multi-cloud LLM inference platform, deploying and managing large language models in production while ensuring reliability, performance, and cost efficiency.
Develops topology-aware collective algorithms and transport layers for AI hardware, optimizing multi-node performance and co-designing with hardware teams.
Senior engineer optimizing generative AI models (LLMs, diffusion) for inference efficiency using PyTorch, CUDA, TensorRT, and Triton, and deploying them at scale.
Build and deploy AI/ML models for Yahoo Mail’s personalization and NLP systems, leveraging PyTorch/TensorFlow and cloud-native architectures at massive scale.
Leads Developer Relations for robotics and edge AI in India, building demos, running workshops, and guiding teams to prototype and deploy AI-powered systems using NVIDIA’s platforms like Isaac Sim, Jetson, and TensorRT.
We couldn't check your fit for this role — add a CV to your profile to see it next time.