Research Engineer / Scientist - Storage for LLM
Summary
Design and build distributed caching systems for large language models, optimizing GPU-aware KV caches, replication, and low-latency access to improve LLM performance.
freehire launches on Product Hunt on 26 August. Follow the page and you'll hear the moment it opens.
Follow →Design and build distributed caching systems for large language models, optimizing GPU-aware KV caches, replication, and low-latency access to improve LLM performance.