Researcher, Synthetic RL
OpenAI - San Francisco, California, United States
Posted Jan 9, 2026
Benefits
- Parental leave
- Not verified
- Non-birth-parent leave
- Not verified
- Family-building benefits
-
- Fertility benefits: Not verified
- Adoption assistance: Not verified
- Surrogacy assistance: Not verified
- Mental health support
- Not verified
- Relocation assistance
- Not verified
- Childcare support
- Not verified
- Learning budget
- Not verified
- Verification
- Not verified
- Salary
- Not verified
- 401(k) match
- Not verified
Was this benefit information wrong? Tell us.
Schedule
- Shift type
- Not verified
- Weekend work
- Not verified
Application
- Cover letter
- Not verified
- Assessment
- Not verified
- Deadline
- Not stated
Where they hire
State eligibility is not yet verified.
About this role
Researcher, Synthetic RL San Francisco, California, United States About the Team The Synthetic RL team develops reinforcement learning methods that leverage synthetic data, environments, and feedback to train and evaluate frontier AI models. The team explores approaches such as self-play, simulators, and other synthetic evaluations to push model capability, generalization, and alignment beyond what is possible with the current prevailing methodology. About the Role As a Research Scientist on the Synthetic RL team, you will develop novel reinforcement learning techniques that use synthetic environments and feedback to improve large-scale models. You'll work closely with other researchers to design experiments, analyze learning dynamics, and translate research insights into training approaches used in production systems. We're looking for researchers who enjoy working on open-ended problems, value fast iteration, and want their work to directly shape how frontier models are trained. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: - Research and develop reinforcement learning algorithms - Design and run experiments to study training dynamics and model behavior at scale - Collaborate with engineers and researchers to integrate successful approaches into model training pipelines You might thrive in this role if you: - Have a strong background in reinforcement learning, machine learning research, or related fields - Have strong engineering and statistical analysis skills - Enjoy exploring new problem spaces where data, objectives, and evaluation are imperfect
Read the full description at jobs.ashbyhq.com. FewerJobs shows a source-linked preview and links to the original posting.
Apply link not verified; last-live date unavailable.
What verified means
Verified means a displayed claim has a recorded source field, a source URL when available, and a timestamp showing when FewerJobs checked or enriched the evidence.
Related jobs
-
Security Coordinator 4 (12675-1. 15471-1. 13771-1)
Northrop Grumman - United States-Utah-Roy
-
Loan Servicing Representative
AXOS Financial INC - Las Vegas, NV
-
Staff Test Conductor
Northrop Grumman - United States-California-Palmdale
-
Off Premise Specialist
Constellation Brands - 2 Locations