Any model you want, and a backup ready before you need it
Pick from live catalogs of Anthropic, OpenAI and 350+ OpenRouter models, with context size, price and tool support shown up front. Stack up to five backups from any provider, and if a model errors or stalls, the next one answers at once. Every attempt, token and cent is on the record.
A closer look at Any model + fallbacks
A model picker that shows you what matters
Search live catalogs from Anthropic and OpenAI (with your key) and OpenRouter. Each model shows its context window, price per million tokens and whether it supports tool calling and thinking. Know a model id that isn't listed yet? Type it in.
- Live catalogs, cached for 30 minutes
- Context size and price per million tokens
- Tool-calling and thinking support flags
- Any model id can be typed in
Outages become a non-event
Give a character up to five backup models, mixing providers. If a model errors, hits a rate limit, refuses, is missing a key or passes the character's timeout, the next one is tried immediately. SDK retries are switched off so nothing waits twice, and every attempt shows in the debugger.
- Up to five cross-provider backups
- Fails over on errors, rate limits, refusals, missing keys and timeouts
- No double waiting: SDK retries are off
- Every attempt visible in the debugger
Know what every conversation costs
Each run records its token usage and cost, so you can see exactly what an agent spends per conversation, per booking and per day, and choose the model that gives the best result for the money.
- Input, output and reasoning tokens per run
- Cost per run and per agent
- Adaptive thinking with per-character effort
- Use your own provider keys or the server defaults
The details
Anthropic
Messages API with adaptive thinking, summarised reasoning and per-agent effort.
OpenAI
Responses API with reasoning summaries and encrypted reasoning passed back each turn.
OpenRouter
350+ models through one key, with unified reasoning kept across tool calls.
Your keys
Add your own provider keys in Settings, encrypted at rest with AES-256-GCM.
Timeouts
A per-character timeout (default 60 s) that triggers the next backup.
Health
Paused providers and failing keys are surfaced before they become a problem.
Who it's for
Always-on SMS
A fast primary model with a different-provider backup keeps replies flowing through outages.
Cost control
A premium model for complex steps, a cheaper model for routing and confirmations.
Model evaluation
A/B test two characters on different models and compare conversions.
Any model + fallbacks: common questions
Can I mix providers in the fallback chain?
Yes. Backups can come from Anthropic, OpenAI and OpenRouter in any order.
Does failover slow replies down?
It's immediate. SDK retries are switched off so a failed model never makes the contact wait twice.
Can I use a model that isn't in the list?
Yes. Type any model id into the picker.
Where do I see what a run cost?
In the debugger for each run and in Analytics across agents.
Works hand in hand with
Characters
Reusable personas with a voice, avatar, model, thinking level and SMS-style message splitting.
Explore OperateTest & Debug
Sandbox conversations, a full debugger for every reply, and an AI diagnose-and-fix loop.
Explore GrowAnalytics
Agents, conversation map, flow drop-off, response times, AI cost, leads and bookings.
ExplorePut it to work on your own inbox
Connect an inbox, describe the job, test it in a sandbox and go live with exactly the autonomy you're comfortable with.