Let's make AI love humans.

Rate answers, compare responses, write better ones, and weigh in on questions where reasonable people disagree. Your judgments become downloadable training data for AI which loves humans, and then is nice to them.

Reinforcement Loving exists to collect human judgments about what makes a response useful, honest, safe, fair, and considerate. The resulting data can be used to fine tune and evaluate open source, or other, models.

Take the data.

The ratings, comparisons, responses, and document reviews can be downloaded and used for research, training, fine tuning, or evaluation.

User IDs in the downloadable data are anonymized. You can see which judgments came from the same reviewer without seeing who that person is.

Lets make AGI love humans. I'm sorry about the domain name.

You don't need an account to start.

Start reviewing

Have material that belongs in the dataset? Upload documents →