> ## Documentation Index
> Fetch the complete documentation index at: https://handbook.sutro.sh/llms.txt
> Use this file to discover all available pages before exploring further.

# Deployment

> Choices to make when your models are ready for action.

Deployment pages cover the runtime choices that shape AI system cost, latency, reliability, and operational control.

## Pages in This Section

* [Batch vs. Real-Time Inference](/deployment/batch-vs-real-time-inference): when to run analytical AI workloads as batch jobs instead of real-time APIs.
* [Model Selection](/deployment/model-selection): how to choose a model based on task fit, cost, latency, control, and operational constraints.
