skinny.

Latest / AI Ethics with Fexingo: Bias, Safety, and Responsible Artificial Intelligence

When AI Learns Bias from Human Feedback

In this episode of AI Ethics with Fexingo, Lucas and Luna explore a subtle but powerful source of algorithmic bias: reinforcement learning from human feedback, or RLHF. As AI systems increasingly rely on human raters to learn what makes a 'good' response, the preferences of those raters can quietly shape the model's behavior. Lucas breaks down a 2024 study from the Allen Institute for AI and the University of Washington that found ChatGPT's outputs were rated as 'more fluent' by non-native English speakers, while native-English raters favored more complex 'overly complex' responses. The…

The skinny

The skinny isn't ready yet — notes appear once the transcript is processed.