Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
The AI revolution is not powered by models alone, rather it advances when enormous amounts of computation become fast, efficient, and economical enough to turn new ideas into products people can use on a global scale.…
For two decades, NVIDIA has pioneered visual computing, the art and science of computer graphics. With our invention of the GPU - the engine of modern visual computing - the field has expanded to encompass personal…
Manage a team that designs, develops, troubleshoots and debugs software programs for databases, applications, tools, networks etc. Lead the end-to-end NPI lifecycle for current and next-generation high-performance NICs…
RELOCATION ASSISTANCE: Relocation assistance may be available CLEARANCE REQUIRED FOR START: Yes CLEARANCE TYPE: Top Secret TRAVEL: Yes, 10% of the Time Description At Northrop Grumman, our employees have incredible…
About Us Quartermaster is building the world's most comprehensive maritime intelligence platform. Our SmartMast™ system transforms commercial and civilian vessels into a persistent, distributed sensing…
Qpurpose is looking for one Algorithmic Software Engineer with a focus on GPU programming and accelerated computing to join our team. Whether you have just graduated from university or have many years of experience,…
Senior Software Engineer builds and optimizes a Kubernetes-native cloud platform for AI inference workloads, focusing on low-latency, high-scale GPU services and performance tuning.
Leads architecture and performance for CoreWeave’s Kubernetes-native AI inference platform, optimizing GPU resource management and cost-per-token under strict SLAs.
Senior engineer builds and optimizes CoreWeave’s Kubernetes-native AI inference platform, improving latency, throughput, and reliability to meet strict P99 SLAs while mentoring peers.
Senior engineer building and optimizing CoreWeave's Kubernetes-native AI inference platform to meet strict latency and reliability SLAs, using Python/Go, CUDA, and distributed systems.
Senior engineer builds and runs Kubernetes-native benchmarking services to measure latency, throughput, and reliability across CoreWeave’s global AI cloud, publishing results like MLPerf.
Develops and optimizes AI model-serving systems on GPU infrastructure, focusing on latency, reliability, and cost while working with tools like Triton, vLLM, and Kubernetes.
Leads architecture and performance for CoreWeave’s Kubernetes-native AI inference platform, optimizing low-latency, high-throughput systems and GPU resource management.
Write, profile, and optimize CUDA kernels for LLM inference to maximize throughput and minimize latency on NVIDIA GPUs, using DSLs like Triton or Mojo and benchmarking with MLPerf.
Staff Applied Research Engineer on CoreWeave's OpenPipe team, developing self-improving AI agents that learn from experience using reinforcement learning and LLM post-training techniques. Work spans from RL research to production systems, leveraging GPU-rich infrastructure and tools like Megatron and Kubernetes to solve bottlenecks in continuous agent learning.
Research and build continuous learning systems for self-improving AI agents on the OpenPipe team. You'll investigate RLHF, reward modeling, and on-policy distillation to solve production bottlenecks in agent training. Core stack uses PyTorch/JAX for model training with Kubernetes and Megatron for distributed GPU infrastructure.
Build and optimize AI infrastructure performance monitoring tools, ensuring high availability and observability for CoreWeave’s cloud platform.
This role involves optimizing large-scale foundation model training on TPU infrastructure by profiling and improving JAX/XLA workloads and developing high-performance kernels. The engineer will lead technical projects and collaborate across teams to enhance training efficiency, scalability, and throughput.
Welcome to the future of cloud networking and security! Cato Networks is the first company to converge enterprise networking and security into one centralized and global service that is delivered by cloud. It is led by…
Builds and maintains cloud-native infrastructure (AWS, Terraform, CI/CD) to support quantum algorithm development tools and simulations for researchers.
We couldn't check your fit for this role — add a CV to your profile to see it next time.