Skip to main content

Model catalog

The model catalog behind every picker in Tale — where it lives under Settings > AI providers, what the capability tags mean, which defaults ship, and how the list stays fresh.

7 min read

Every model picker in Tale — the chat's model menu, an agent's model binding, the defaults the crawler and RAG services use — draws from one catalog: the models declared on your organisation's AI providers. A fresh instance ships with a single provider, OpenRouter, whose one key covers chat, vision, embeddings, transcription, text-to-speech, and image generation. This page is the reference for where that catalog lives in the UI, what the tags on each model mean, and what ships out of the box.

Where the catalog lives

Open Settings > AI providers and click a provider row. The drawer lists everything the provider declares: its base URL and API key, its Default Models, and the Models list itself — searchable, with Show more past the first ten. Add model declares a new entry by hand; Fetch models pulls the list the provider's API reports. Models an admin marks as Hidden from model pickers stay resolvable for existing bindings but stop appearing in menus — that is how superseded versions retire without breaking old agents.

Each model carries one or more capability tags: Chat, Vision, Embedding, Transcription, Text-to-speech, Image generation, Image edit. The tags are load-bearing — they decide which pickers a model shows up in and which platform capability is allowed to call it. A model with no matching tag never appears where that capability is needed.

The shipped defaults

The Default Models card names which model each background capability uses when nothing more specific is bound:

CapabilityShipped default
ChatDeepSeek V4 Flash
VisionQwen3 VL 32B
EmbeddingQwen3 Embedding 8B
Image generationFLUX.2 [pro]
TranscriptionWhisper v1

Text-to-speech for voice mode ships on OpenAI's GPT-4o mini TTS through the same OpenRouter key, and image generation defaults to FLUX.2 [pro].

How the list stays fresh

Models drift faster than docs. Two mechanisms on the AI providers page keep the catalog current: the Model catalog card refreshes model capabilities — pricing, context window, reasoning, vision — from OpenRouter's public catalog daily, and the Weekly auto-sync of provider config toggle merges newly released flagship versions into the org's provider config once a week, hiding superseded ones and leaving any field you customised untouched.

The shipped list below is regenerated from the same source, so it matches what a fresh instance sees:

ProviderModelCapabilitiesContextInput ($/M)Output ($/M)
AI21Jamba Large 1.7chat256K2.008.00
AmazonNova Premierchat, vision1M2.5012.50
AmazonNova 2 Litechat, vision1M0.302.50
AnthropicClaude Fable (latest)chat, vision1M10.0050.00
AnthropicClaude Fable 5chat, vision1M10.0050.00
AnthropicClaude Sonnet 4.6chat, vision1M3.0015.00
AnthropicClaude Haiku 4.5chat200K1.005.00
AnthropicClaude Opus 4.8chat, vision1M5.0025.00
Black Forest LabsFLUX.2 [flex]image-generation, image-edit
Black Forest LabsFLUX.2 [max]image-generation, image-edit
Black Forest LabsFLUX.2 [pro]image-generation, image-edit
CohereCommand Achat256K2.5010.00
CohereCommand Rchat128K0.150.60
DeepSeekDeepSeek V4 Prochat1M0.430.87
DeepSeekDeepSeek V4 Flashchat1M0.090.18
GoogleGemini 3 Prochat, vision1M2.0012.00
GoogleGemini 3 Flashchat, vision1M0.503.00
GoogleGemma 4 31B ITchat, vision262K0.120.35
GoogleGemma 4 26B A4B ITchat, vision262K0.060.33
GoogleNano Banana (Gemini 2.5 Flash Image)image-generation, image-edit33K0.302.50
LiquidLFM2 24Bchat128K0.030.12
MetaLLaMA 4 Maverickchat1M0.150.60
MetaLLaMA 4 Scoutchat10M0.100.30
MicrosoftPhi-4chat16K0.070.14
MiniMaxMiniMax M3chat, vision1M0.301.20
MistralMistral Large 3chat262K0.501.50
MistralMistral Medium 3.5chat, vision262K1.507.50
Moonshot AIKimi K2.6chat, vision262K0.683.41
Moonshot AIKimi K2.7 Codechat, vision262K0.613.07
NVIDIANemotron 3 Ultrachat1M0.502.20
NVIDIANemotron 3 Superchat1M0.090.45
OpenAIGPT-OSS 120Bchat131K0.040.18
OpenAIGPT-4o mini TTStext-to-speech
OpenAIGPT-5.3 Chatchat, vision128K1.7514.00
OpenAIGPT-5.5chat, vision1M5.0030.00
OpenAIGPT-5.5 Prochat, vision1M30.00180.00
OpenAIWhisper v1transcription
PerplexitySonar Prochat, vision200K3.0015.00
PerplexitySonarchat, vision127K1.001.00
QwenQwen3.6 Max Previewchat262K1.046.24
QwenQwen3 Coder 480Bchat1M0.221.80
QwenQwen3 VL 32Bchat, vision262K0.100.42
QwenQwen3.6 Flashchat, vision1M0.191.13
QwenQwen3 Embedding 8Bembedding0.010.00
QwenQwen3.7 Pluschat, vision1M0.321.28
RekaReka Flash 3chat66K0.100.20
XiaomiMiMo V2.5 Prochat1M0.430.87
Z.AIGLM 5.1chat203K0.983.08
Z.AIGLM 5 Turbochat262K1.204.00
Z.AIGLM 5V Turbochat, vision131K1.204.00
xAIGrok 4.20chat, vision2M1.252.50

The full and live catalogue lives at openrouter.ai/models; any model OpenRouter exposes can be added to your instance from the same drawer.

Where this fits

Models are the layer beneath every agent, every chat reply, every voice output, and every image the platform renders. OpenRouter is the default, not a requirement — adding a direct vendor, a local Ollama or vLLM server, or a second gateway is admin work covered in Providers, and the file-based form of the same configuration lives under Configuration → providers. For picking between chat models when more than one could do the job, Arena Mode is the workflow built for exactly that question.

© 2026 Tale by Ruler GmbH — ISO 27001 & SOC 2 certified.

Tale is MIT licensed — free to use, modify, and distribute.