Quincer AI Docs Home Support Start free

AI models & your own keys

Every brand picks the language model behind its assistant. Out of the box you run on the platform default — nothing to configure, usage on the platform meter. Add your own provider key and you can pin any model in that provider's catalog, from a handful of Claude models to OpenRouter's 300+.

1. Pick a model for a brand

Go to Customize & Deploy → AI agent. The AI agent card at the top has two controls:

Picking a concrete model fires a quick live test against the provider — you'll see Testing model... and then Tested OK, or the provider's error if the model can't answer on your key. Save when you're happy.

AI agent
Customize & Deploy — Acme
Provider
OpenRouter
Language model
deepseek v4
DeepSeek V4 Flash
openrouter/deepseek/deepseek-v4-flash-0731
DeepSeek V4 Pro
openrouter/deepseek/deepseek-v4-pro
DeepSeek V3.1
openrouter/deepseek/deepseek-v3.1
✓ Tested OK

The AI agent card — type into Language model to narrow a big catalog; picking a model runs a live test on your key.

i

Platform default is a moving pointer, on purpose. A brand left on Platform default resolves the model at request time, so when we upgrade the default your brand comes along automatically. The dropdown shows what it currently resolves to (“Platform default — currently …”).

2. Platform meter, or your own key

There are two ways to pay for model usage, and the picker reflects which one you're on:

Keys live in Settings → Workspace, in the AI API keys card (“Add your own API keys to unlock additional models”). There's a field per provider — OpenAI, Anthropic, Kindo, X-AI (Grok), Google Gemini, Baseten, OpenRouter and Together.ai — and one Save API keys button. When a provider is currently covered by the shared platform key, the field says so — naming that provider: “Currently using Quincer's shared OpenRouter key. Add your own to bill usage to your account and control the rate limits.”

Every key is checked when you save it: we ask the provider to list its models on that key and report back — Verified with a model count, rejected (the key is saved but nothing will run on it — replace it), or couldn't reach the provider (usually temporary). Saving several keys back-to-back can skip the re-check for a few seconds — save again if you want that one checked too. The same verdicts appear on the Customize & Deploy picker, so a rejected key is never a silent gap in the Provider list.

Workspace
Settings · AI API keys
AI API keys

Add your own API keys to unlock additional models.

OpenRouter API key
sk-or-••••••••••••
✓ Verified — OpenRouter accepted this key and offers 312 models.

Unlocks OpenRouter — one key across hundreds of models, with automatic failover between upstream providers.

Together.ai API key
Key from api.together.ai

Unlocks Together.ai — open-weight models (Llama, Qwen, DeepSeek, Mixtral) on their own inference.

Save API keys

The AI API keys card on Settings → Workspace — each key is verified against the provider the moment you save it.

i

Keys belong to the organization that owns the brand. If you manage a brand owned by a different organization (a client's, say), a key added under your Settings doesn't apply to it — the picker tells you when that's the case.

3. OpenRouter — one key, 300+ models

OpenRouter is a router: one key, and its catalog spans hundreds of models across many vendors — the full list appears in your picker (300+ chat models at the time of writing), which is exactly why the model box is typeable. Three things happen automatically:

4. Together.ai — the open-weight lineup

Together.ai hosts open-weight models — Llama, Qwen, DeepSeek, Mixtral and friends — on its own inference. Add a Together.ai key and its chat catalog joins your picker. Together's listing also includes embeddings, rerankers, image and moderation models; those are filtered out, so everything you can pick can actually hold a conversation.

5. Voice: choose the OpenAI realtime model

Voice has its own model setting, separate from text chat. If your default voice provider is OpenAI Realtime, you can choose which realtime model answers calls — under Settings → Voice, in the Call behavior card:

  1. Set Default voice provider to OpenAI Realtime (gpt-realtime). An OpenAI realtime model field appears below it.
  2. Type any model id — it's free text, so you can adopt a new realtime release the day it ships, without waiting for us. The field commits when you click away.
  3. Leave it blank to stay on the platform default (gpt-realtime). Clearing the field puts you back on the default.

The model you pick is the model that answers — the same value drives the call connection, the live session, and the per-call cost report, so billing always reflects the model that actually spoke. Two cautions, both shown inline in the dashboard:

Voice
Settings · Call behavior
Default voice provider
OpenAI Realtime (gpt-realtime)
OpenAI realtime model
gpt-realtime-mini

Calls currently use gpt-realtime-mini. Leave blank for the platform default (gpt-realtime). Must be a realtime speech-to-speech model.

Check pricing for this model. Voice cost is looked up by model id. If we hold no price for gpt-realtime-mini, calls still work but their cost reports as $0 until a price is added.

The OpenAI realtime model field on Settings → Voice — type any realtime model id; blank means the platform default.

Voice keys, silence timeout, call length, and the per-persona voice picker are covered in Voice chat.

Troubleshooting

What you seeWhat it means
A provider you added a key for isn't in the Provider listThe picker says why, right above the dropdown: the provider rejected the key (replace it in Settings → Workspace), we couldn't reach the provider (usually temporary — reload in a minute), or the key was accepted but offers no usable chat models.
You only see Platform defaultYou haven't added any provider keys yet — the shared platform keys serve the default model but don't open their catalogs for picking. Add a key under Settings → Workspace.
A warning names a model your brand is pinned toThe brand is pinned to a model it can no longer run — usually because the key that unlocked it was removed. The banner offers to clear the pin back to Platform default.
Model test failedThe provider refused that model on your key. Check the key's permissions, or pick another model — the error text is the provider's own.
Voice calls report $0 costYou set a custom OpenAI realtime model we hold no price for. The calls are fine; ask us to add the price so cost reporting is accurate.