Tech jobs
Job listings
AI Solution Architect - BytePlus
Designs and deploys AI agents, LLM safety systems, and retrieval pipelines to bridge research with production, optimizing performance and integration with model APIs.
LLM Algorithm Engineer Intern - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
A PhD internship focused on improving and optimizing large language models (LLMs) for production use, including multimodal AI agents, RL training, and SFT pipelines.
Research Engineer / Scientist - Storage for LLM
Research Engineer/Scientist focused on building distributed KV cache systems and GPU-aware caching layers for LLM inference. Core tasks include optimizing low-latency access, implementing memory-aware sharding, and integrating cache with token streaming pipelines.
Research Engineer / Scientist - Storage for LLM
Research Engineer/Scientist to design and optimize distributed KV cache systems for large language model (LLM) token streaming pipelines, focusing on consistency, low-latency access, and memory-efficient sharding.
Senior Research Engineer / Scientist - Storage for LLM
Develops and optimizes distributed KV cache systems for large language model (LLM) token streaming pipelines, focusing on GPU-aware caching, consistency protocols, and performance tuning.