Home Knowledge Base BOLD

BOLD is the Bias in Open-Ended Language Generation benchmark that evaluates social bias patterns in free-form model outputs across demographic domains - it focuses on bias in generation rather than only classification tasks.

What Is BOLD?

Why BOLD Matters

How It Is Used in Practice

BOLD is a key benchmark for bias assessment in open-ended language generation - domain-level sentiment and regard analysis helps identify representational harms in real conversational and content-generation use cases.

boldboldevaluation

Explore 500+ Semiconductor & AI Topics

From EUV lithography to CUDA optimization — search the full knowledge base or chat with our AI assistant.