Skip to content

Glossary

RLHF

Reinforcement learning from human feedback: training a model using human preference judgments.

Estimate my data's value