Model Inference
Kind: capability
Layer: Correctness Core
Aliases: Runtime Prediction/Generation
Record: architecture:model-inference
Severity: contextual
Scope: service, model serving
Canonical: Ontology
The ability to run a trained model on validated input at runtime and return a checked prediction or generation.
Listed in Architecture principles, after Model Evaluation and before Retrieval-Augmented Generation (RAG).