Model Inference

Kind: capability

Layer: Correctness Core

Aliases: Runtime Prediction/Generation

Record: architecture:model-inference

Severity: contextual

Scope: service, model serving

Canonical: Ontology

The ability to run a trained model on validated input at runtime and return a checked prediction or generation.

Listed in Architecture principles, after Model Evaluation and before Retrieval-Augmented Generation (RAG).

Requires

Reinforces

Conflicts with

In tension with

Tensions

Violated by

Refactored by

Severity

Category

Required by

Linked from