Skip to content
APIs for agentic telephony

Voice AI API pricing, per engine and per second

Wixzel Voice costs $0.0466 to $0.1851 per connected minute, depending on the voice engine, billed per second of actual usage from one prepaid balance. One API key, one balance, every voice engine; no subscription, no seats, and a $5 minimum top-up.

How much does an AI phone call cost per minute?

An AI phone call on Wixzel Voice costs $0.0466 to $0.1851 per connected minute, depending on the voice engine, plus your carrier’s rate for the SIP trunk. The figures include the flat $0.012 per minute orchestration fee and are generated from the price book the meter bills against. The meter charges per second, so a 90-second call is billed for 90 seconds, not two minutes.

Engine1 min call3 min call5 min call
classic$0.0867$0.2601$0.4335
sarvam$0.0623$0.1869$0.3115
gemini-live$0.0466$0.1398$0.2330
deepgram-agent$0.1851$0.5553$0.9255

Engine cost at typical speech rates, plus your carrier’s per-minute rate for the trunk, which Wixzel Voice neither includes nor marks up. A web or app session on the same engine costs the same, with no carrier cost at all.

Pricing

Pick an engine, or compose one

Prices are per connected minute at typical speech rates, including the flat $0.012 orchestration fee. The table is generated from the same price book the meter bills against, so it cannot drift from what you are charged.

EnginePipelinePer minute
classicDeepgram + GPT-4o-mini + ElevenLabs$0.0867
sarvamSarvam, Indian languages end to end$0.0623
gemini-liveGemini Live, native audio$0.0466
deepgram-agentDeepgram Voice Agent$0.1851

Telephony is not included and not marked up. Phone calls run over your own SIP trunk, so you keep your carrier rates and your carrier relationship. Web and app sessions are billed at the same per-minute rates, with no telephony at all.

How credit works, and the minimum top-up

Estimate a month

Estimated monthly cost
$433.50
Per call
$0.4335
Minutes per month
5,000
classic, per minute
$0.0867

At typical speech rates, including the $0.012/min platform fee. Metered per second, token and character, so a quiet call costs less and a talkative one more. Telephony is billed by your carrier and is not included.

Compose your own

Name a model for each stage. The speech-to-text sets the family, the other two stages follow it, and the price is what those three add up to per minute.

Per connected minute
$0.0867

Runs on the Deepgram · OpenRouter · ElevenLabs pipeline, which the speech-to-text selects. At typical speech rates, including the platform fee; metered per second, token and character, and billed for the model that actually ran. Within a family the pipeline picks the variant for the agent’s language. How composing works

Paste into POST /v1/agents
{
  "voice": {
    "stt": {
      "model": "deepgram/nova-3"
    },
    "llm": {
      "model": "openrouter/gpt-4o-mini"
    },
    "tts": {
      "model": "elevenlabs/eleven_turbo_v2_5"
    }
  }
}
Capacity

How many calls you can run at once

Concurrency is the number that decides whether this fits, and it is not a plan tier. Every account starts at five simultaneous calls; if you need more, that is a conversation rather than an upgrade button.

5

Concurrent calls per account

The default every account starts on. Enough to build, test and run a small line. Raised on request. It is a deliberate change on our side, not a plan tier.

402

Refused, never degraded

Past your limit, the next call is refused with a clear error and the calls already running are untouched. Nobody gets choppy audio because somebody else got busy.

0

Queue

There is no waiting room. A refusal is immediate and explicit, so your own retry logic decides what happens next rather than a hidden queue deciding for you.

Concurrency is not billed. You pay for connected seconds whether one line is busy or forty, so a higher limit costs nothing until it is used. The limit exists so we can size the platform honestly rather than oversell it.

Tell us what you need: how many lines busy at peak, roughly how many minutes a month, and which carrier you are on. Those three answers are usually enough for us to commit to a number in one reply.

Billing

Every charge is traceable to a second of audio

Credit is reserved before a call is placed and debited as it runs. Every micro that leaves the balance is explained by a usage row, so a bill is answerable without asking support.

  1. Reserved at admission

    Credit for the first stretch of the call is held before a port is opened or a provider socket is created. An account that cannot fund it is refused with 402, so a call that never happens costs nothing.

  2. Debited as it runs

    Speech seconds, tokens and characters are metered against the balance every few seconds while the call is live. Run low and the agent warns your caller; run out and it says goodbye rather than dropping the line.

  3. Settled at hangup

    The hold is released and only what was used stays debited, one row per component, reconciled against the call’s real duration.

$ curl .../v1/billing/balance

{
  "object": "balance",
  "balance_display": "$24.75",
  "held_micros": 250000,   # held by live calls
  "available_micros": 24500000
}

Amounts are integer micro-USD, so a charge of $0.000021 has a representation that agrees with the ledger. Every response carries a display string beside it.

No surprise invoices

Prepaid. You cannot be billed for more than you added.

Prices frozen per call

A call is priced at admission and finishes on those rates, so a price change never moves a meter mid-conversation.

Under-billed on our failures

If we crash mid-call you are charged for the seconds we can prove, not the amount we reserved.

Refused before it costs you

A call that cannot fund its first 15 seconds is refused before anything is allocated. At about a minute of runway the agent warns your caller and then hangs up cleanly.

Self-host license

Or run it on your own server

With a self-host license you install Wixzel Voice on servers you control and pay your voice providers directly, with no orchestration fee on calls. It costs $499, paid once, and every future release is included.

Pricing questions

What does Wixzel Voice cost per minute?
From $0.0466 (gemini-live) to $0.1851 (deepgram-agent) per connected minute at typical speech rates, including the flat $0.012 per minute orchestration fee. Every figure is generated from the price book the meter bills against and rounded up, never down. Telephony runs over your own SIP trunk at your carrier's rate and is not included or marked up.
Is there a subscription, a seat price or a minimum monthly spend?
No. Credit is prepaid from a $5 minimum top-up and spent per second of actual usage. An account with no calls spends nothing.
How are web and app sessions billed?
Exactly like a phone call on the same engine: the same per-minute price, from the same balance, per second. A session counts toward your concurrent-call limit, appears in your call log with channel: "web" and ends with the same callCompleted webhook. There is no telephony cost, because there is no phone line.
How am I billed?
Prepaid. You add credit, calls spend it per second of speech, per token and per character, and you cannot be billed for more than you put in. The minimum top-up is $5.
What happens when credit runs out?
A call that cannot fund its first 15 seconds is refused with a 402 before anything is allocated. Mid-call, at about a minute of runway, the agent warns your caller, then speaks a closing line and hangs up. It does not drop the line mid-sentence.
Can I retry a request safely?
Send an Idempotency-Key. It is required on POST /v1/calls, POST /v1/agents/{id}/test-call and POST /v1/billing/topups, and a replay with the same key and body returns the original response instead of dialling twice. Keys are kept for 24 hours. POST /v1/campaigns/{id}/start takes no key, and starting a stopped campaign dials every lead on it again.
Is there a free trial?
There is no time-limited trial and no free tier. Credit is prepaid from $5, a short test call costs a few cents, and the Credits and Refunds policy explains what can be refunded. You can talk to each engine from the landing page before adding any credit.
How many calls can run at once?
Five per account by default, phone calls and web sessions counted together. Past the limit the next call is refused with a clear error and the calls already running are untouched. The cap is raised on request; email the address on the contact page with the volume you expect.
Is there a way to pay once instead of per minute?
Yes. The self-host license is $499, paid once, to run Wixzel Voice on your own servers with your own voice provider accounts and no orchestration fee on calls. The installer, every future release and the source code are included. You then pay your providers and your carrier directly.

Start with a curl

Sign up and create a key. Point a SIP trunk at us for phone calls, or mint a realtime session to put the agent in your app.