Model Configurations
Configure the speech-to-speech or bring-your-own-key models your Talkr agents use.
How model configurations work
Model Configurations define the default AI model setup for your organization. Agents use this configuration unless you set agent-level model overrides in the agent settings.
To configure models, open Models in your Talkr dashboard:
- Hosted:
https://talkr.intelstacks.com/model-configurations - Self-hosted:
http://localhost:3010/model-configurations - Local development:
http://localhost:3000/model-configurations
Talkr is bring-your-own-key (BYOK) only — there is no managed AI model tier. The Models page has two top-level sections:
| Section | When to use it |
|---|---|
| Speech to Speech | Use a realtime speech-to-speech model for the live conversation. You still configure an LLM alongside it for variable extraction and QA. |
| BYOK | Bring your own provider keys and configure LLM, Voice, Transcriber, and Embedding models separately. |
Model settings are organization-scoped. If no agent-level override is set, every agent in the organization uses the saved global configuration.
Speech to Speech
Use Speech to Speech when you want a realtime model to handle the live spoken conversation directly. In this mode, the realtime model handles speech input and speech output, so you do not configure separate Voice and Transcriber services.

The Speech to Speech section has nested tabs:
| Tab | What to configure |
|---|---|
| Realtime Model | The speech-to-speech provider, model, voice, and API key. |
| LLM | A standard LLM used for non-realtime tasks such as variable extraction and QA analysis. |
| Embedding | An embedding model used by features that need embeddings, such as retrieval from knowledge base content. |
An LLM is still required when you use Speech to Speech. The realtime model handles the live voice conversation, but Talkr uses the LLM for analysis tasks that happen outside the live audio stream.
BYOK
Use BYOK when you want to bring your own provider accounts and API keys. This gives you separate control over each model category.
When you use BYOK or external model providers, Talkr sends only the data required for the selected service to operate. Depending on the provider and service type, this may include prompts, conversation history, transcripts, audio, generated text, tool/function definitions, tool inputs or results, and request metadata.
Provider data handling varies. Review each provider's data processing, retention, model training, and regional hosting policies before using sensitive data.

The BYOK section has nested tabs:
| Tab | What to configure |
|---|---|
| LLM | The chat or reasoning model provider, model, optional base URL, and API key. |
| Voice | The text-to-speech provider, voice, model, speed, optional base URL, and API key. |
| Transcriber | The speech-to-text provider, model, language, and API key. |
| Embedding | The embedding provider, model, and API key. |
Provider-specific fields appear only when they apply. For example, OpenAI-compatible LLM providers can expose a Base URL field, ElevenLabs voices can expose a voice ID, and transcribers can expose language options.
Agent-level model overrides
You can override the organization model configuration for an individual agent. This is useful when different agents need different models, voices, transcribers, or providers.
To configure an override:
- Open the agent.
- Go to Settings.
- Open Model Overrides.
- Enable the override for the service you want to customize.
- Configure the provider, model, and keys for that service.
- Save the agent settings.
Agent-level overrides are selective. For example, you can override only the Voice service for one agent while it continues to use the organization-level LLM and Transcriber configuration.