freehire launches on Product Hunt on 26 August.

Follow →

Large Language Model Inference System Engineer Graduate (Applied Machine Learning) - 2027 Start

Summary

Build and optimize inference systems for large language models as part of a Managed AI Service (MaaS), focusing on distributed KV caching, GPU performance tuning, and multi-node heterogeneous inference.

- Develop inference system for large model MaaS - Implement disaggregated multi role inference - Implement distributed KV Cache system - Improve inference stability - Optimize GPU performance for inference - Optimize large model inference cost - Optimize large model inference performance - Optimize multi node heterogeneous inference Perks/Benefits: - 401k match - Dental insurance - Life insurance - Long-term disability - Medical insurance - Paid Holidays - Paid parental leave - Paid personal time - Paid sick days - Short-term disability - Vision insurance

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available