Rate answers, compare responses, write better ones, and weigh in on questions
where reasonable people disagree. Your judgments become downloadable training data for AI which
loves humans, and then is nice to them.
Reinforcement Loving exists to collect human judgments about what makes a
response useful, honest, safe, fair, and considerate. The resulting data can be used to fine tune
and evaluate open source, or other, models.
Take the data.
The ratings, comparisons, responses, and document reviews can be downloaded and used for research,
training, fine tuning, or evaluation.
User IDs in the downloadable data are anonymized. You can see which judgments came from the same
reviewer without seeing who that person is.