KrissDevHub
Technologies
0%
All positions

RLHF Specialist

AI Workforce Remote (Worldwide) Contract

We are seeking RLHF Specialists to compile preference data that aligns models with human intent. You will review multiple model completions and rank them based on truthfulness, helpfulness, and safety.

What you'll do

  • Compare model responses and construct preference pairings (preferred vs. rejected)
  • Provide written justifications explaining ranking choices using detailed guides
  • Conduct multi-turn dialogue evaluations to judge context retention and coherence
  • Assist in calibrating training datasets to minimize model sycophancy and bias

What we're looking for

  • Exceptional critical thinking skills and cognitive empathy
  • Background in philosophy, ethics, cognitive science, or communication
  • Proven ability to adhere to highly specific and abstract guidelines
  • Prior experience in AI safety or alignment work is highly valued

Benefits & Perks

  • Flexible schedule with contract autonomy
  • Fully remote collaboration model
  • Be at the forefront of safe AI alignment methodologies
  • Opportunity for continuous learning and professional development

Ready to apply?

Takes about 5 minutes. We review every application personally.

Apply for this position

Similar roles