Tech jobs
Job listings
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Develops and optimizes high-performance inference engines for large AI models, focusing on GPU/NPU hardware acceleration, parallelism techniques, and performance tuning to reduce latency and improve throughput.
Tech Lead, Machine Learning Engineer - Global E-Commerce (Conversational AI)
Lead a team to design and ship conversational AI features for a global e-commerce platform, including ML observability and cross-team technical alignment.
Software Engineer Graduate (Inference Infrastructure) - 2026 Start (PHD)
Build and optimize GPU-based AI inference infrastructure for large-scale LLM deployments, focusing on cluster management, security, and cost efficiency.
Applied Scientist - LLM Training System as a Service - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
Build and optimize large language model training and inference systems, including reinforcement learning frameworks, using GPU/CUDA for ByteDance’s AI products.