Skip to main content
Open-weight LLMs hosted on Telnyx GPU infrastructure. All models accessible via the Chat Completions API (OpenAI-compatible). Not every model here is available to AI Assistants, and of the assistant models only a subset is verified for voice calls — see Voice AI models.

Chat Models

Reasoning (thinking) behavior differs by surface: reasoning models return their chain-of-thought in a separate reasoning_content field on chat completions (see getting started), but reasoning is always disabled on voice calls and cannot be enabled there — see Reasoning on voice calls.

Embedding Models