Start here. This is the direct spoken answer to practice first.
Overview
Foundry deployment types change processing geography, capacity, billing, and supported models; the label is not merely a hosting preference.
I start with the required model, processing geography, traffic shape, latency target, and commitment tolerance. Shared pay-per-token deployment fits variable demand, provisioned throughput fits predictable critical load, and batch fits asynchronous volume. Regional or data-zone choices follow residency requirements, while managed compute is considered when the required open or custom model is not available as a managed API.