freehire launches on Product Hunt on 26 August.

Follow →

Research Engineer - Language Model Pre-Training

Open 60d

Zyphra is an artificial intelligence company based in San Francisco, California.

The Role:

As a Research Engineer - Language Model Pre-Training, you'll shape our language model roadmap through end-to-end pretraining development. You will work extremely closely with our pretraining team, who will integrate your insights into our next-generation models.

You'll Work Across:

  • Large-scale training runs and model parallelization

  • Performance optimization of our pretraining stack

  • Dataset collection, processing, and evaluation

  • Architecture and methodology research, including optimizer ablations

What We're Looking For / Requirements:

  • Strong engineering aptitude for rapidly implementing reliable and robust systems

  • Can rapidly learn new fields and are excited to implement new ideas

  • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

Qualifications / Additional Skills:

  • Deep expertise and intuition for solving machine learning problems and training models

  • Experience with training on large-scale (multi-node) GPU clusters

  • Deep understanding of model training pipelines – including model/data parallelism, distributed optimizers, etc.

  • Strong grasp of proper experimental methodology for running rigorous ablations and other hypothesis testing

  • Understanding of large-scale, highly parallel data processing pipelines

  • High proficiency with PyTorch and Python.

  • Strong ability to dive into large pre-existing codebases and rapidly get up to speed

  • Published machine learning research in well-respected venues is a plus

  • Postgraduate degree in a scientific subject (Computer Science, EE/EECS, Math, Physics)

Why Work at Zyphra:

  • Our research methodology is to make grounded, methodical steps toward ambitious goals. Both deep research and engineering excellence are equally valued

  • We strongly value new and crazy ideas and are very willing to bet big on new ideas

  • We move as quickly as we can; we aim to minimize the bar to impact as low as possible

  • We all enjoy what we do and love discussing AI

Benefits and Perks:

  • Comprehensive medical, dental, vision, and FSA plans

  • Competitive compensation and 401(k) plan

  • Relocation and immigration support on a case-by-case basis

  • In-office snacks and meals provided

  • Unlimited PTO and company holidays

  • In-person team in San Francisco with a collaborative, high-energy environment

What this application asks

ashby

Name, Email, Resume

  • LinkedIn Profile optional
  • Are you willing to work from our office in San Francisco, California full-time, 3+ days a week? yes / no
  • If you are not located near San Francisco, California, are you willing to relocate if offered employment? yes / no
  • Are you legally authorized to work in the United States for Zyphra? yes / no
  • Do you now, or will you in the future, require sponsorship for employment visa status (e.g., H-1B visa, etc.) or a visa extension in order to start or continue to work legally for Zyphra in the United States? yes / no
  • Why do you want to work for Zyphra? written answer
  • Why would you be a good fit for this role? written answer
  • Which one project most demonstrates the skills needed for this role and why? written answer · optional
  • Which one paper around pretraining have you found insightful and why? written answer · optional
  • What  is  one  technique or idea you want to test out at scale? written answer · optional
  • If  you have any open-source work, what commit or project are you most proud of? written answer · optional
  • When is the earliest you would want to start working with us?

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available