Kinesis
Pricing & Services

Pay for what runs. Cap at the reserved rate.

True-Util™ is the one pricing model behind every Kinesis service — whether the compute is ours, sourced from a partner, or yours. Pay only for the cycles your workloads actually consume, capped at the reserved rate. Idle capacity isn't wasted — it's monetized.

Traditional cloud

$0.00

True-Util™

$0.00

Idle Capacity

$0.00

Available to monetize on the Kinesis grid

00:0008:0016:0024:00
Reserved rate (traditional)True-Util™ (actual usage)Idle capacity · monetizable
The True-Util™ Model

One pricing model. Two ways to buy.

True-Util™ works the same way regardless of how much computing power you need. Pick the buying mode that matches your workload — the metering, caps, and telemetry are identical.

TRUE-UTIL™ SERVERLESS

Metered usage, capped at the Dedicated rate

Serverless compute on the Kinesis grid. You pay for the compute, memory, storage, and bandwidth your workloads actually use, and never more than you would pay for the equivalent Dedicated machine. Spiky, variable, or hard-to-forecast workloads save the most.

  • Best for inference, dev/staging, agencies, MVPs
  • Runs across a range of CPU and GPU classes
  • No upfront commitments
TRUE-UTIL™ DEDICATED

The whole machine, billed by the hour

Single-tenant compute that is yours for as long as you run it. Full control over the box, predictable billing, and the same Kinesis orchestration and telemetry as Serverless. For workloads where steady utilization is a given.

  • Best for steady training, production HPC, regulated workloads
  • Choice across providers, which reduces lock-in
  • Full control over configuration, performance, privacy
Pricing

Rates

Dedicated server rates, effective July 2026.

A100$1.35

Per GPU / Per Hour

1x GPU, 28 CPUs, 120GB RAM, 750GB Storage

Available in 1x, 2x, 4x GPU configurations.

H100$2.50

Per GPU / Per Hour

1x GPU, 28 CPUs, 180GB RAM, 750GB Storage

Available in 1x, 2x, 4x GPU configurations.

A100 NVLink$12.00

Per Node / Per Hour

8x GPU, 252 CPUs, 1920GB RAM, 6500GB Storage

H100 NVLink$22.00

Per Node / Per Hour

8x GPU, 252 CPUs, 1440GB RAM, 6500GB Storage

H200 SXM$34.00

Per Node / Per Hour

8x GPU, 176 CPUs, 1800GB RAM, 48000GB Storage

B200 SXM$52.00

Per Node / Per Hour

8x GPU, 252 CPUs, 2048GB RAM, 40000GB Storage

Compute Optimized CPU$0.035

Per vCPU / Per Hour

1 vCPU, 2GB RAM, 50GB NVMe

Available as serverless True-Util™. Available as 2x, 4x, 8x, 16x, 32x and 64x configurations.

General Purpose CPU$0.045

Per vCPU / Per Hour

1 vCPU, 4GB RAM, 50GB NVMe

Available as serverless True-Util™. Available as 2x, 4x, 8x, 16x, 32x and 64x configurations.

Memory Optimized CPU$0.055

Per vCPU / Per Hour

1 vCPU, 8GB RAM, 50GB NVMe

Available as serverless True-Util™. Available as 2x, 4x, 8x, 16x, 32x and 64x configurations.

Prices shown are for representative configurations. Actual specifications may vary by available stock.

Where True-Util™ saves the most

The same workloads that cost the most on traditional clouds save the most on True-Util™.

AI startups & LLM inference
The pain

H100s sit idle between prompts. The bill is the same whether you served 100 requests or 100,000.

The Kinesis win

True-Util™ Shared meters inference time only. No queries, no cost. Bursty traffic caps at the Reserved rate.

Early-stage SaaS & MVPs
The pain

Overprovisioning for traffic that hasn’t shown up. Or worse — under-provisioning and falling over the first time it does.

The Kinesis win

Pay pennies at low traffic. Costs cap at the Reserved rate during spikes. Headroom without prepayment.

Dev, Staging & CI/CD
The pain

Staging servers run 24/7 to be ready, but burn nights and weekends.

The Kinesis win

True-Util™ drops the bill as activity drops. Same reservation, lower cost when the team’s asleep.

Enterprises with idle capacity
The pain

Reserved AWS instances, on-prem servers, donated lab GPUs — capacity already paid for, sitting underused.

The Kinesis win

Run the Kinesis grid on your hardware at 20% of Shared. Same orchestration, FinOps visibility, 80% less spend on what you already own.

Try it on a real app

$100 in free credit. No credit card required. Deploy your first container in under five minutes — bring a GitHub repo, a Dockerfile, or just describe what you want