AI Job Radar

RLHF Jobs – Page 2

Aktuelle KI-Jobs mit RLHF, passende Lernpfade und Bewerbungsbezug.

How to use RLHF in applications

If a job requires RLHF, the skill should be supported by a project, course or portfolio example. The application check reviews whether the skill is actually evidenced in your CV.

12
Results
6
Companies
95.0
Average score
Remote

2 results on this page. 12 results in total. More results are available via pagination, company pages, skill pages and job detail pages.

San Francisco, CAUSAgreenhouse2026-06-23

Why this is a real AI job: The role explicitly focuses on research and development of post-training techniques for LLMs, including SFT, RLHF, and reward modeling. The job description highlights the application of these techniques to enhance LLM capabilities and solve core AI problems.

Scale works with the industry’s leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling). This role will focus on optimizing data curation an…

Details Open source / apply

San Francisco, CAUSAgreenhouse2026-06-15

Why this is a real AI job: The role is explicitly focused on building and improving systems for training large language models using reinforcement learning. The job description details tasks directly related to ML algorithms, infrastructure, and model finetuning.

About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working to…

Details Open source / apply