Pricing
Know what you pay before you ship.
Published catalog rates, adjusted for demand and capacity. Providers earn a transparent share of each completed job.
- Per-token billing
- Catalog rates published
- Live dashboard
Per-tokenInput & output billed separately
PublishedRates before you ship
TransparentSame economics for devs and providers
LiveUsage in Scalattice Cloud
Live rates
Rates by model.
Loading live rates…
Developers
Per-token inference.
Billed per million input and output tokens. Rates vary by model and demand. Compare usage in Scalattice Cloud.
Published
Rates you can quote before you ship
Per token
Input and output billed separately
Demand
Rates adjust for capacity
# Per-million tokens (live catalog)
# Loading published rates…
Catalog families
Qwen
Llama
DeepSeek
Gemma
Mistral
gpt-oss
Providers
Earn per job served.
You set availability. Scalattice routes paying inference to your hardware. Payouts tracked in the dashboard with no connection fees.
Share
Revenue split
The majority of developer token spend on each completed inference job.
Payouts
On demand
Request a payout any time your available balance is above the minimum threshold.
Control
Per-machine schedule
Set availability windows for each GPU from Scalattice Cloud.
# Provider dashboard (example)
jobs completed 847
tokens served 12.4M
earnings $284.50
available $42.50
# request payout any time (min $10)
Enterprise
Volume and committed use.
Contact us for committed capacity, custom regions, and invoicing. Enterprise features are arranged with our team.