PhilPapers: Recent additions to PhilArchive
González Barman, Kristian ; Lohse, Simon & de Regt, Henk W.: Reinforcement Learning from Human Feedback in LLMs: Whose Culture, Whose Values, Whose Perspectives?
0Sign in to vote or save
This page did not load. You can still read it on the original site — the toolbar below keeps your place in the directory.
_Philosophy and Technology_ 38 (2):1-26. 2025We argue for the epistemic and ethical advantages of pluralism in Reinforcement Learning from Human Feedback (RLHF) in the context of Large Language Models (LLMs). Drawing on social epistemology and pluralist philosophy of science, we suggest ways in which RHLF can be made more responsive to human needs and how we can address challenges along the way.…

Comments
Nothing yet. Say the first thing.
Sign in to join the conversation.