Service tiers
Inference requests use thedefault service tier unless you set service_tier to
priority. The priority tier is available for select models and is billed at
the priority rates shown on the pricing page.
Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
Check out our upcoming events and meetups! View events →
Pay-per-token pricing for the Telnyx Inference API with no minimums or commitments. See current per-model rates and available regions.
default service tier unless you set service_tier to
priority. The priority tier is available for select models and is billed at
the priority rates shown on the pricing page.
| Category | Basis | Notes |
|---|---|---|
| Text generation | Per 1M tokens (input + output) | Input and output priced separately; cached input tokens at a discount |
| Audio transcription | Per second of audio | Varies by model |
| Text-to-speech | Per 1M characters | Varies by voice/model |
| Embeddings | Per 1M tokens | Single rate |
Was this page helpful?