North Model LabsNorth
Model Labs
ShowcaseResearchAPIPricingEnterpriseSolutionsFAQTeam
Dashboard

Models

  • Atlas Realtime Avatar
  • Showcase
  • Benchmark
  • Compare

Developers

  • Documentation
  • Examples
  • Pricing
  • Dashboard ↗

Solutions

  • Customer support
  • Sales / SDR
  • AI tutors
  • All solutions

Enterprise

  • Enterprise overview
  • Safety & identity
  • Compare
  • Book deployment review ↗
  • eric@northmodellabs.com

Company

  • Team
  • Partnerships
  • FAQ
  • eric@northmodellabs.com

Connect

  • GitHub ↗
  • Discord ↗
  • eric@northmodellabs.com
North Model Labs© 2026 North Model Labs
PrivacyTermseric@northmodellabs.comBook 30 min ↗

[ 01 / Pricing ]

Simple,
pay-as-you-go.

No subscriptions, no credit packs, no seat fees. Pay only for what you generate.

Get started
Self-serve

[ 01 / Rendering ]

$7/ hour

$0.117 / min · $0.0019 / sec

Realtime passthrough (you provide audio) or async POST /v1/generate, prorated by output duration.

  • GPU avatar rendering
  • WebRTC video stream
  • Single reference image, one-shot
  • Use your own STT / LLM / TTS
  • React SDK & REST API
  • No fixed realtime session-length cap
Start rendering
Production & enterprise

[ 02 / Contracted capacity ]

Custom

Reserved capacity, higher throughput, and contracted support. Pricing depends on workload shape and volume.

  • Everything in Rendering
  • Dedicated GPU capacity
  • Higher request limits and burst quota
  • Configurable retention windows
  • Contracted support and response targets
Book 30 min ↗or email eric@northmodellabs.com

Offline / async video

$7 / hr of generated video ($0.117 / min · $0.0019 / sec). Upload a face image + audio, get a finished video back via POST /v1/generate. Same rate as realtime, billed by output duration. See API docs.

[ 03 / Production & enterprise ]

When you need more than self-serve

Self-serve at $7/hr is designed for evaluation and standard usage. Production and enterprise are for teams that need reserved capacity, a specific deployment topology, or contracted support. We don't publish a fixed price for those, every workload is sized differently. Reach out and we'll quote yours.

Self-serve

$7 / hour

pay-as-you-go · no commitment

  • Shared GPU pool
  • 30 RPM basic self-serve limit
  • Email support
  • Pay-per-second billing
Start now ↗
Production

Custom

monthly or annual · high-volume throughput and concurrency

  • High-volume request limits and burst quota
  • Configurable retention windows
  • Direct support over Slack
  • Contracted response targets
Talk to us ↗
Enterprise

Custom

annual · reserved GPU capacity

  • Dedicated GPU pool · contracted capacity
  • SSO via your IdP · operational logs
  • DPA · sub-processor list · 99.9% SLA
  • Shared Slack/Teams · named technical owner
  • BYO storage · stricter retention controls
Enterprise overview →

Email eric@northmodellabs.com with rough monthly volume, realtime workload shape, and any deployment constraints. We'll follow up with sizing and next steps.

◆ Built for Enterprise

Enterprise controls and support.

SCIM, SSO, dedicated GPU capacity, passthrough by default, and a real audit trail. Enterprise contracts only.

SCIM / SAML SSO

Enterprise IAM via your IdP. Provision and de-provision users programmatically; no per-seat fees.

Passthrough by default

Atlas receives audio and returns video. Your prompts, customer data, and knowledge base never leave your stack.

Dedicated capacity + SLA

Reserved GPU capacity, contracted throughput, and uptime terms scoped in Enterprise agreements.

Audit log + security review

DPA, sub-processor list, and security questionnaire support for procurement.

Enterprise overviewBook a deployment reviewAcceptable use & abuse policy

How billing works

The same $7 hourly rendering rate applies to realtime and async output. Usage is measured by active session time or generated video duration.

01

Realtime

Billed by active session time and prorated to the second.

02

Async video

Billed by generated output duration at the same rendering rate.

03

Enterprise

Reserved capacity and contracted support are scoped and quoted separately.

[ Atlas Realtime Avatar ]

Why Atlas is different

A compact avatar rendering model designed for efficient realtime and async inference. See the architecture and operating model behind Atlas.

See the benchmark

Start building.

Add a payment method, generate an API key, and start building with realtime avatars.

DashboardAPI docs