AIRLHFPyTorch

Staff Research Engineer, Post-training

Frontier AI Lab (client name withheld under NDA)

San Francisco / Remote Full-time $420-620k + equity

Own post-training for a frontier model family - RLHF, preference data and evaluation - inside a small, senior research org.

About the client

A well-funded frontier lab building general-purpose models. The post-training group is deliberately small and works directly with pre-training and product. Client name is shared with shortlisted candidates after a mutual NDA.

What you'll do

  • Design and run post-training experiments across RLHF, DPO and preference-data pipelines
  • Own evaluation harnesses that decide what ships
  • Partner with pre-training and inference teams on data and serving trade-offs
  • Mentor research engineers and set technical direction for the group

What we're looking for

  • Deep experience training or fine-tuning large language models in PyTorch or JAX
  • Track record of published work, shipped models, or equivalent industrial impact
  • Comfort operating at large GPU scale with distributed training
  • Strong engineering fundamentals - this is a hands-on role

Interview process

  1. 01Intro call with a Progress Lane partner
  2. 02Technical deep-dive with the research lead
  3. 03Take-home or paired research session
  4. 04Onsite panel and offer

Client names are withheld until a mutual NDA is in place. We never approach your current employer and never share your details without your written consent.