Tech jobs
Job listings
1 job
23 hours ago
Research Engineer / Scientist - Storage for LLM
RemoteNorth America
United States
Research Engineer/Scientist to design and optimize distributed KV cache systems for large language model (LLM) token streaming pipelines, focusing on consistency, low-latency access, and memory-efficient sharding.
batching c++ cache aware scheduling caching algorithms cuda +27 skills