If you’re pricing a real-time AI avatar API, the sticker price on the pricing page is the wrong number to anchor on. What you actually pay is cost per minute of conversation, multiplied by how many minutes your users spend talking to the avatar — and at any real scale, that’s where the bill lives.
For Spatius specifically, the current public rates are straightforward: $0.009/min on Starter, Builder, and Growth; $0.007/min on monthly Scale; and $0.0056/min on annual Scale. This guide breaks down what drives real-time avatar cost, why those Spatius rates are lower than cloud-streamed alternatives, and how the most affordable option on the market gets there.
If you’re comparing named vendors rather than only chasing the lowest rate, pair this with our market-wide pricing comparison, Spatius vs Anam.ai, and best platforms like Synthesia.
Why real-time AI avatars are expensive in the first place
Most real-time avatar platforms render the avatar in the cloud and stream the resulting video to the user. That architecture has two cost drivers baked in:
- GPU rendering in the cloud, per session, in real time. Rendering video frames continuously is GPU-intensive, and you’re paying for that compute for every minute every user is connected.
- Bandwidth. Streaming video to the device needs a sustained 1–2 MB/s. That’s egress cost for the provider and, ultimately, a line item in your rate.
Add it up and cloud-streamed avatar platforms often carry a materially higher live-minute cost than on-device rendering. The point is not that every platform has the same rate; it is that cloud GPU rendering and video egress make each conversation minute more expensive. For the architectural background, see on-device AI avatar vs cloud streaming.
How on-device rendering changes the math
Spatius uses a different architecture, and the cost follows directly from it. Instead of rendering video in the cloud, the cloud Motion Server does only lightweight driving inference and sends down compact Motion data (driving parameters) — about 10–20 KB/s. The client SDK renders the avatar locally.
Two cost drivers improve at once:
- Bandwidth drops from ~1–2 MB/s to ~10–20 KB/s — roughly two orders of magnitude less data to move.
- GPU cost is minimized. The heavy rendering moves off the cloud and onto the user’s device, where it runs as rendering only, with no inference. There’s still lightweight driving inference in the cloud, so cloud GPU work is dramatically reduced compared with rendering full video frames server-side. As the Spatius docs put it, the approach is “replacing heavy cloud rendering with a lightweight stream and rendering on edge.”
That’s why an interactive avatar can run on hardware with no dedicated GPU at all — the device only renders. And it’s why Spatius can publish rates as low as $0.007/min on monthly Scale and $0.0056/min on annual Scale.
The numbers: per-minute and per-hour
On the Spatius Scale plan, the monthly effective rate is $0.007/min — about $0.42 per hour of conversation. Annual Scale lowers the published public rate to $0.0056/min, or about $0.34 per hour. (Worth flagging precisely: those hourly figures are Scale-plan rates. Starter, Builder, Growth, Free, and Enterprise have different economics.)
The gap versus cloud-streamed alternatives is easiest to feel with a fixed budget. Spend $5,000 on conversation minutes:
| Effective rate | Hours of conversation for $5,000 | |
|---|---|---|
| Spatius monthly Scale | $0.007/min | ~11,905 hours |
| Spatius annual Scale | $0.0056/min | ~14,881 hours |
| Cloud-streamed alternatives | Higher variable live-minute cost from cloud GPU rendering and video egress | Fewer hours on the same budget |
Same budget, more talk time. For a high-frequency use case — a language tutor, a customer-service agent, a companion app where sessions run long — that ratio is the difference between viable unit economics and a runaway cloud bill.
Full Spatius pricing breakdown
Spatius prices in credits, where 10 credits ≈ 1 minute of conversation. Monthly plans (as of July 9, 2026):
| Plan | Price | Credits/mo | ≈ Minutes | Concurrency | Max session |
|---|---|---|---|---|---|
| Free | $0 | 1,000 | ~100 min | 2 | 10 min |
| Starter | $19/mo | 20,000 | ~2,000 min | 4 | 30 min |
| Builder | $49/mo | 55,000 | ~5,500 min | 8 | 60 min |
| Growth | $149/mo | 180,000 | ~18,000 min | 22 | 120 min |
| Scale | $299/mo | 400,000 | ~40,000 min | 40 | Unlimited |
| Enterprise | Custom | Unlimited | Unlimited | Custom | Custom |
Annual billing saves 20% and lowers the effective public rate further — Starter to $15.75/mo equivalent ($0.0072/min), Builder to $39.08/mo equivalent ($0.0072/min), Growth to $119.08/mo equivalent ($0.0072/min), and Scale to $239.08/mo equivalent ($0.0056/min).
A few things to know so the budgeting is accurate:
- Credits don’t roll over on subscription plans — they reset each cycle. Credits you purchase or that are gifted to your team are permanent.
- Avatar generation (building a custom avatar from a photo) is a separate quota and doesn’t consume your conversation credits. It’s currently in Beta, granted manually by the team, and failed generations are automatically refunded.
- There’s a genuinely usable permanent free tier (1,000 credits/~100 min/month) for prototyping. See the full pricing page.
”Most affordable” should never mean “no AI included” by surprise
One honest caveat that applies to every platform in this category: a real-time avatar API is the avatar + driving/rendering layer. The AI agent — speech recognition (ASR), the LLM, and text-to-speech (TTS) — is a separate stack. With Spatius, you bring your own AI (or wire up providers you already use); Spatius does not provide ASR/LLM/TTS. That keeps the avatar layer affordable and swappable, but it means your total cost includes whatever you spend on those AI services. When you compare “most affordable API” claims across vendors, check what’s bundled — you’re not always comparing the same scope. More on the three-layer split in our interactive avatar complete guide.
How the per-minute cost compares to other platforms
The cost advantage shows up clearly in head-to-head comparisons, because it’s structural, not promotional:
- Spatius vs Synthesia — on-device rendering at roughly 99% lower cost per minute than Synthesia’s video-generation plans.
- Spatius vs Tavus — ~98% lower cost per minute than cloud video streaming.
- Spatius vs LiveAvatar — ~95% lower cost per minute.
- Spatius vs Anam.ai — on-device rendering at a fraction of cloud cost.
For the full landscape, see best on-device AI avatar platforms in 2026.
The takeaway
The most affordable real-time AI avatar API achieves its low cost not through a temporary discount, but through where the work happens. Cloud-streamed video carries GPU-rendering and bandwidth cost into every minute of every session. Move rendering to the device, stream only Motion data, and the per-minute rate drops by roughly an order of magnitude — which on a fixed budget turns hundreds of hours of conversation into thousands.
Start on the free tier, check the math on the pricing page, or just talk to a live avatar in the Playground and watch your network usage while you do.
Recommended reading
- On-Device AI Avatar vs Cloud Streaming: Architecture, Bandwidth, and Cost
- Best On-Device AI Avatar Platforms in 2026 (Ranked & Compared)
- Comparing AI Avatar Platforms for Speed