Tech jobs
Job listings
Research Engineer / Scientist - Storage for LLM
Research Engineer/Scientist focused on building distributed KV cache systems and GPU-aware caching layers for LLM inference. Core tasks include optimizing low-latency access, implementing memory-aware sharding, and integrating cache with token streaming pipelines.
Research Engineer - LLM Training Infrastructure - Seed Infra
Research Engineer focused on optimizing and scaling infrastructure for large language model (LLM) training, addressing performance bottlenecks and designing distributed training strategies for exascale systems.
Research Engineer / Scientist - Storage for LLM
Research Engineer/Scientist to design and optimize distributed KV cache systems for large language model (LLM) token streaming pipelines, focusing on consistency, low-latency access, and memory-efficient sharding.
Senior Research Engineer / Scientist - Storage for LLM
Develops and optimizes distributed KV cache systems for large language model (LLM) token streaming pipelines, focusing on GPU-aware caching, consistency protocols, and performance tuning.