Perspectives on the Social Impacts of Reinforcement Learning with Human Feedback

rlhfsocial-impactethicsai-policy

Abstraction: Social and ethical impacts of RLHF across seven societal dimensions

Key points:

Connections: Openai · Anthropic · Deepmind · Reinforcement Learning From Human Feedback · AI Safety

Source: https://arxiv.org/abs/2303.02891