Model Safety

Kind: quality-attribute

Layer: Correctness Core

Record: architecture:model-safety

Severity: contextual

Scope: model-backed system, model, application

Canonical: Ontology

The degree to which a model-backed system prevents harmful inputs, outputs and actions through evaluation, guardrails and monitoring.

Listed in Architecture principles, after Explainability and before Prompt Engineering.

Requires

Reinforces

Enables

Conflicts with

In tension with

Tensions

Violated by

Refactored by

Severity

Category

Reinforced by

Linked from