Versatile AI Risk Assessment

AI threat catalogueHarmful ContentProduction

Hate Speech and Discrimination

The AI system produces content that demeans individuals or groups on the basis of protected characteristics such as origin, gender, religion, or disability, or that calls for their exclusion.

As of: July 2026 · Catalogue version 2026.07.17.3 · 5 mitigations · 9 verified sources

Description

Such output arises in three ways: on direct request, through a jailbreak (the deliberate circumvention of the safety controls built into the model), or unintentionally, when the model reproduces prejudice and bias absorbed from its training data. The range runs from stereotyping phrasing and disparaging language to incitement of hatred or violence against an identity group. Any channel in which the system generates free-form text is affected, including chatbots, assistants, and automated decisions.

Possible impact

Operators face reputational damage, legal exposure under anti-discrimination law such as the German General Equal Treatment Act (AGG), and regulatory consequences. Discriminatory output violates the fundamental right to non-discrimination and directly harms the people concerned. In automated processes such as recruitment, disadvantaging results can systematically exclude entire groups of people.

Example

A recruitment chatbot phrases a rejection in a way that demeans female applicants because of their gender, or a customer-service assistant answers a harmless question with a stereotyping statement about an ethnic group.

Recommended mitigations (5)

Every mitigation states its control type, effect, implementation level and the reason for the classification.

Framework mappings

Verified locations in OWASP, NIST AI RMF, MITRE ATLAS, the EU AI Act and further frameworks. The mappings are taxonomic, not evidence of compliance.

OWASP LLM Top 10 LLM01:2025NIST AI RMF Section 2.3 · Section 2.6 · MEASURE 2.11MITRE ATLAS AML.T0048.002EU AI Act Article 55(1)(b)BSI R5BIML BIML-LLM output:11 · BIML78 system:1

Verified references (9)

Every reference states the framework, the exact location and the publishing organisation.

Terms on this page

Glossary terms that occur in this entry. Every link leads to the full explanation.

More entries from the topic group Harmful Content.

Assess this threat in your own system

The live demo contains all 52 threats of this catalogue, including the EU AI Act and GDPR assessment. The free single modules cover AI risk, the EU AI Act and GDPR. No sign-up; the assessment runs locally in your browser.

Cite this entry

For reports, policies or internal documents; the link leads directly to this entry.

“Hate Speech and Discrimination”. Versatile AI Risk Assessment, AI threat catalogue, as of July 2026.
https://www.versatile-ai-risk-assessment.com/en/wissensbasis/threats/hate-speech-discrimination/

← Back to the full catalogue