Model Safety
Kind: quality-attribute
Layer: Correctness Core
Record: architecture:model-safety
Severity: contextual
Scope: model-backed system, model, application
Canonical: Ontology
The degree to which a model-backed system prevents harmful inputs, outputs and actions through evaluation, guardrails and monitoring.
Listed in Architecture principles, after Explainability and before Prompt Engineering.