hirly

Apply with hirly

Research Scientist - RLHF, RLAIF & Reward Modeling

Weekday AI · India

Upload your resume to see how well you match this job — free, in seconds, no account needed.

Your resume is used only to score it against this job. If you don't create an account, it is deleted within 24 hours.

This role is for one of Weekday’s clients Salary range: Rs 5000000

  • Rs 10000000 (ie INR 50
  • 100 LPA) Min Experience: 3 years Location: Bengaluru, Karnataka, India JobType: full-time We are looking for a highly skilled and research-oriented Research Scientist with 3–6 years of experience in machine learning, reinforcement learning, and large language model (LLM) alignment. The ideal candidate will have strong hands-on experience with Reinforcement Learning from Human Feedback (RLHF), Reinforce…
Apply: Research Scientist - RLHF, RLAIF & Reward Modeling at Weekday AI