freehire launches on Product Hunt on 26 August.

Follow →

Research Engineer - LLM/VLM Inference Optimization (Seed Infra)

Summary

Research Engineer focused on optimizing inference for large-scale LLMs and VLMs. The role involves building model toolchains, conducting performance analysis, and designing high-performance inference engines and deployment pipelines to improve system efficiency.

- Build model toolchains and technical ecosystem - Conduct performance analysis and identify bottlenecks - Design high performance inference systems for large scale LLMs and VLMs - Develop model inference engines with performance optimization techniques - Optimize end to end deployment pipelines Perks/Benefits: - 401k matching - Dental insurance - Life insurance - Long-term disability - Medical insurance - Paid Holidays - Paid parental leave - Paid personal time - Paid sick days - Short-term disability - Vision insurance - Wellbeing benefits

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available