Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Tiger Analytics is seeking a highly experienced Lead AI Engineer to lead the end-to-end AI Engineering workstream for the Luma platform. This is a hands-on technical leadership role responsible for driving the…
Position Title: GenAI-Team Lead III - AI -GR-39656-70585-JR201049 Job Family: IFT > Engineering/Dev Shift: Job Description: Job Title Team Lead III - AI Requirement Type Full-Time Employee Job Location…
Senior AI/ML Engineer 8-15 Years Experience JOB SUMMARY We are seeking a Senior AI/ML Engineer to provide technical leadership within our Artificial Intelligence and Automation team. The ideal candidate combines deep,…
Company Overview Intuit is the global financial technology platform that powers prosperity for the people and communities we serve. With approximately 100 million customers worldwide using products such as TurboTax,…
Would you like to join a team curious about understanding how foundation models work and to expand their capabilities in scientific domains? We perform and publish novel research and apply our findings to drive product…
Senior Software Engineer builds and optimizes a Kubernetes-native cloud platform for AI inference workloads, focusing on low-latency, high-scale GPU services and performance tuning.
Leads architecture and performance for CoreWeave’s Kubernetes-native AI inference platform, optimizing GPU resource management and cost-per-token under strict SLAs.
Designs and deploys AI infrastructure solutions for enterprise customers, focusing on LLM workloads and GPU-based cloud platforms.
Senior engineer builds and optimizes CoreWeave’s Kubernetes-native AI inference platform, improving latency, throughput, and reliability to meet strict P99 SLAs while mentoring peers.
Senior engineer building and optimizing CoreWeave's Kubernetes-native AI inference platform to meet strict latency and reliability SLAs, using Python/Go, CUDA, and distributed systems.
Senior engineer builds and runs Kubernetes-native benchmarking services to measure latency, throughput, and reliability across CoreWeave’s global AI cloud, publishing results like MLPerf.
Develops and optimizes AI model-serving systems on GPU infrastructure, focusing on latency, reliability, and cost while working with tools like Triton, vLLM, and Kubernetes.
Leads architecture and performance for CoreWeave’s Kubernetes-native AI inference platform, optimizing low-latency, high-throughput systems and GPU resource management.
Write, profile, and optimize CUDA kernels for LLM inference to maximize throughput and minimize latency on NVIDIA GPUs, using DSLs like Triton or Mojo and benchmarking with MLPerf.
Applied AI Engineer, Inference at CoreWeave. Improve real-world performance of AI models via benchmarking, profiling, and optimization.
CoreWeave is The Essential Cloud for AI™. Built for pioneers by pioneers, CoreWeave delivers a platform of technology, tools, and teams that enables innovators to build and scale AI with confidence. Trusted by leading…
Design and deliver AI infrastructure demos and proofs-of-concept for new customers, partnering with sales and engineering to onboard greenfield AI teams onto CoreWeave’s GPU-powered cloud platform.
Build and optimize AI infrastructure performance monitoring tools, ensuring high availability and observability for CoreWeave’s cloud platform.
Company Qualcomm Middle East Information Technology Company LLC Job Area Engineering Group, Engineering Group > Software Engineering General Summary About Us Qualcomm is enabling a world where everyone and…
Welcome to the future of cloud networking and security! Cato Networks is the first company to converge enterprise networking and security into one centralized and global service that is delivered by cloud. It is led by…
We couldn't check your fit for this role — add a CV to your profile to see it next time.