MODELS
Choose fast inference for user-facing work or lower-cost inference for background, evaluation, and long-running agents. Terminal handles routing and scaling behind the API.

MODELS
Choose fast inference for user-facing work or lower-cost inference for background, evaluation, and long-running agents. Terminal handles routing and scaling behind the API.
