Xterra AI logo
Xterra AIPosted 1 month ago

Founding Research Scientist

On-siteSan Francisco, California, United States

Full TimeDoctorate Or Professional DegreeStartup

Job Summary

Design and train reinforcement learning systems using RLHF, RLAIF, and reward modeling approaches for scientific hypothesis generation and evaluation. Develop process reward models and verifiers to provide fine-grained supervision over intermediate reasoning steps. Contribute to alignment and oversight research for complex scientific tasks where ground truth is expensive or ambiguous. Build robust training pipelines, run large-scale experiments, and iterate quickly across the research-to-production lifecycle. Contribute to meaningful benchmarks and evaluation methods for domain-specific reasoning. Own problems end-to-end from idea through experimentation to production.

Required Qualifications

  • Strong fundamentals in machine learning
  • Hands-on experience training large models
  • Demonstrated experience with reinforcement learning
  • Comfort working across the research-engineering spectrum
  • Familiarity with reward modeling
  • Familiarity with RLHF/RLAIF pipelines
  • Familiarity with search and planning methods
  • Familiarity with AI alignment techniques

Desired Qualifications

  • LLMs
  • Publication record
  • A track record of identifying and driving high-impact research directions independently
  • Experience mentoring other researchers and influencing technical strategy
  • Deep expertise in one or more of the core technical areas listed above

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce