Korean Hate Speech
KoreanHigher is better
Judge whether comments from a Korean entertainment-news site are hate speech. It measures how accurately the model detects hate speech, not whether it produces any. Horangi runs a fixed 100-question subset rather than the full dataset, so its numbers are not directly comparable with full-dataset scores. Scores run from 0 to 100%. Higher is better.
Top score78.0%Claude Fable 5
Models tested48
Last updated2026.08.03
Model release date (newest on the right)
Best score so farHigher on the chart is better
RankModelSettingReleasedPrice/1MScore
- 1high-effort2026.06.09$4078.0%
- 2high-effort2025.10.15$476.0%
- 3high-effort2025.05.22$6071.0%
- 4high-effort2025.12.01$0.8370.0%
- 42025.11.19$0.4270.0%
- 62026.04.28$669.0%
- 62026.03.15$3.3069.0%
- 8high-effort2026.02.17$1267.0%
- 82026.02.16$2.8067.0%
- 82026.03.16$0.4967.0%
The "harness" is the agent program the AI used to carry out the task. The same model can score very differently depending on its harness and reasoning effort.