Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Research engineer at Graphcore designs and implements AI/ML experiments, optimizes models for hardware efficiency, and publishes findings to advance AI compute technology.
Engineer deep-learning inference engines for GigaChat, optimizing CUDA kernels and GPU clusters to deliver sub-second LLM responses at massive scale.
Senior Staff ML Engineer in San Diego builds production-ready inference components from research prototypes, optimizing for latency, memory, and reliability in constrained local environments using Rust, C++, and CUDA.
Design and run AI hardware benchmarks for GPUs, TPUs and custom silicon, analyze throughput, latency and cost-per-token, and work with leading chipmakers to set industry standards.
Design and run AI hardware benchmarks (GPUs, TPUs, custom silicon), analyze inference economics, and partner with chipmakers to set industry standards for performance and cost-per-token.
Build and scale Snowflake’s Cortex LLM post-training platform, designing distributed GPU scheduling, APIs, and fault-tolerant ML infrastructure to let customers fine-tune open models at enterprise scale.
Build and maintain CrowdStrike’s AI/ML infrastructure, debugging Ray, Spark, Airflow, and Kubernetes pipelines that process billions of daily events to keep the cybersecurity platform running.
About Us Resaro was founded on the belief that AI will change the world in ways we cannot even imagine - but every new technology needs safeguards to advance. We are an independent, third-party AI assurance company: we…
Engage with CAE developers and ISVs to drive adoption of NVIDIA’s accelerated computing platforms, providing technical expertise and shaping product strategy through partnerships and community outreach.
Designs and deploys private AI systems for document intelligence and legal automation using LLMs, RAG, and GPU infrastructure in a law firm setting.
Senior engineer optimizing generative AI models (LLMs, diffusion) for inference efficiency using PyTorch, CUDA, TensorRT, and Triton, and deploying them at scale.
Develops, deploys, and maintains AI/ML models for a defense program using GCP, focusing on LLMs, RAG, and agentic workflows to support military operations.
Leads Developer Relations for robotics and edge AI in India, building demos, running workshops, and guiding teams to prototype and deploy AI-powered systems using NVIDIA’s platforms like Isaac Sim, Jetson, and TensorRT.
Leads strategic AI adoption for India's major conglomerates by engaging executives and technical teams to identify, build, and scale enterprise AI workloads using NVIDIA's platforms and software.
Leads cross-functional compiler development programs at NVIDIA, coordinating teams to deliver GPU-focused software like CUDA, AI, and HPC tools on time using Agile and Waterfall methods.
Build and deploy AI/ML models (NLP, LLMs, computer vision) for national security systems, using Python frameworks and GPU programming to enhance decision-making.
Build and maintain data pipelines and platforms to organize disparate data for mission-driven projects like fraud detection and national intelligence using ETL, cloud infrastructure, and Python.
About the team: Materra is on a mission to radically reduce global waste and move to a true circular economy. The team has developed technology that identifies waste material at the molecular level—starting with…
Algorithm - Serving System Engineer Department: Algorithm Location: Seoul HQ Employment Type: FullTime About the Role 우리 팀은 NPU 기반 inference의 roofline 자체를 개선하는 방법을 연구합니다. 6개월~2년 뒤에 제품에 적용될 수 있는 도전적인 과제를 스스로 정의하고 탐색하며,…
Senior Solutions Engineer at TensorWave troubleshoots and resolves complex technical escalations for customers running large-scale AI workloads, focusing on Kubernetes, GPU infrastructure, and high-performance networking.
We couldn't check your fit for this role — add a CV to your profile to see it next time.