RSS Amplifier

PhilPapers: Recent additions to PhilArchive

González Barman, Kristian ; Lohse, Simon & de Regt, Henk W.: Reinforcement Learning from Human Feedback in LLMs: Whose Culture, Whose Values, Whose Perspectives?

0
Sign in to vote or save

This page did not load. You can still read it on the original site — the toolbar below keeps your place in the directory.

_Philosophy and Technology_ 38 (2):1-26. 2025We argue for the epistemic and ethical advantages of pluralism in Reinforcement Learning from Human Feedback (RLHF) in the context of Large Language Models (LLMs). Drawing on social epistemology and pluralist philosophy of science, we suggest ways in which RHLF can be made more responsive to human needs and how we can address challenges along the way.…

Read on philarchive.org

Comments

Nothing yet. Say the first thing.

    Sign in to join the conversation.