Research Engineer - LLM/VLM Inference Optimization (Seed Infra)
Summary
Research Engineer focused on optimizing inference for large-scale LLMs and VLMs. The role involves building model toolchains, conducting performance analysis, and designing high-performance inference engines and deployment pipelines to improve system efficiency.
- Build model toolchains and technical ecosystem
- Conduct performance analysis and identify bottlenecks
- Design high performance inference systems for large scale LLMs and VLMs
- Develop model inference engines with performance optimization techniques
- Optimize end to end deployment pipelines
Perks/Benefits:
- 401k matching
- Dental insurance
- Life insurance
- Long-term disability
- Medical insurance
- Paid Holidays
- Paid parental leave
- Paid personal time
- Paid sick days
- Short-term disability
- Vision insurance
- Wellbeing benefits