[ 01 / Pricing ]
Simple,
pay-as-you-go.
No subscriptions, no credit packs, no seat fees. Pay only for what you generate.
[ 01 / Rendering ]
$0.117 / min · $0.0019 / sec
Realtime passthrough (you provide audio) or async POST /v1/generate, prorated by output duration.
- GPU avatar rendering
- WebRTC video stream
- Single reference image, one-shot
- Use your own STT / LLM / TTS
- React SDK & REST API
- No fixed realtime session-length cap
[ 02 / Contracted capacity ]
Reserved capacity, higher throughput, and contracted support. Pricing depends on workload shape and volume.
- Everything in Rendering
- Dedicated GPU capacity
- Higher request limits and burst quota
- Configurable retention windows
- Contracted support and response targets
Offline / async video
$7 / hr of generated video ($0.117 / min · $0.0019 / sec). Upload a face image + audio, get a finished video back via POST /v1/generate. Same rate as realtime, billed by output duration. See API docs.
[ 03 / Production & enterprise ]
When you need more than self-serve
Self-serve at $7/hr is designed for evaluation and standard usage. Production and enterprise are for teams that need reserved capacity, a specific deployment topology, or contracted support. We don't publish a fixed price for those, every workload is sized differently. Reach out and we'll quote yours.
$7 / hour
pay-as-you-go · no commitment
- Shared GPU pool
- 30 RPM basic self-serve limit
- Email support
- Pay-per-second billing
Custom
monthly or annual · high-volume throughput and concurrency
- High-volume request limits and burst quota
- Configurable retention windows
- Direct support over Slack
- Contracted response targets
Custom
annual · reserved GPU capacity
- Dedicated GPU pool · contracted capacity
- SSO via your IdP · operational logs
- DPA · sub-processor list · 99.9% SLA
- Shared Slack/Teams · named technical owner
- BYO storage · stricter retention controls
Email eric@northmodellabs.com with rough monthly volume, realtime workload shape, and any deployment constraints. We'll follow up with sizing and next steps.
◆ Built for Enterprise
Enterprise controls and support.
SCIM, SSO, dedicated GPU capacity, passthrough by default, and a real audit trail. Enterprise contracts only.
SCIM / SAML SSO
Enterprise IAM via your IdP. Provision and de-provision users programmatically; no per-seat fees.
Passthrough by default
Atlas receives audio and returns video. Your prompts, customer data, and knowledge base never leave your stack.
Dedicated capacity + SLA
Reserved GPU capacity, contracted throughput, and uptime terms scoped in Enterprise agreements.
Audit log + security review
DPA, sub-processor list, and security questionnaire support for procurement.
How billing works
The same $7 hourly rendering rate applies to realtime and async output. Usage is measured by active session time or generated video duration.
01
Realtime
Billed by active session time and prorated to the second.
02
Async video
Billed by generated output duration at the same rendering rate.
03
Enterprise
Reserved capacity and contracted support are scoped and quoted separately.
[ Atlas Realtime Avatar ]
Why Atlas is different
A compact avatar rendering model designed for efficient realtime and async inference. See the architecture and operating model behind Atlas.
See the benchmarkStart building.
Add a payment method, generate an API key, and start building with realtime avatars.