Skip to main content
Pay-per-token. No minimums, no commitments. For current per-model pricing, see telnyx.com/pricing/inference-api.

Service tiers

For standard Inference API requests to Telnyx-hosted models, omitting service_tier uses default. Set service_tier to choose another tier supported by the model. Flex and Priority are available for select models. Check the model’s service_tiers field and use the corresponding tier’s rates on the pricing page. See Service tiers for model availability, request examples, and guidance on choosing a tier.

Billing units