The500Feed.Live

Everything going on in AI - updated daily from 500+ sources

← Back to The 500 Feed
📄 ResearchMay 26, 2026

Attribute-Based Diagnosis of LLM Alignment with Hate Speech Annotations

Hate speech annotation is costly, subjective, and prone to annotator disagreement, making large-scale dataset construction challenging. We systematically analyze how well large language models (LLMs) align with human judgments across ten theoretically grounded subjective attributes, such as dehumani...

Read Original Article →

Source

http://arxiv.org/abs/2605.27025v1