Tech jobs
Job listings
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Develops and optimizes high-performance inference engines for large AI models, focusing on GPU/NPU hardware acceleration, parallelism techniques, and performance tuning to reduce latency and improve throughput.
Machine Learning Backend Engineer Graduate (AML MLdev) - 2027 Start
A graduate-level role focused on optimizing machine learning model performance by applying graph optimization, kernel fusion, and high-performance kernel libraries to improve computational efficiency and inference/training speed for Bytedance’s AI systems.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Works on adapting inference engines to GPU/NPU hardware, benchmarking with vLLM/TensorRT-LLM, and optimizing distributed parallel inference solutions—including cache, memory, and latency improvements.