Pricing
GPU rental
| GPU | VRAM | Per GPU-hour |
|---|---|---|
RTX 4090 (rtx4090) | 24 GB | $0.40 |
RTX 5090 (rtx5090) | 32 GB | $0.70 |
- Billed per minute while a deployment is
ready. Image pulls, start-up and failed deployments are free. - No deployment fee, no egress fee, no minimum term. Stop with
DELETE /platform/deployments/{id}and billing stops. - To start a rental, your balance must cover one hour of its price.
- The price is fixed when a GPU is assigned to your deployment, so later list-price changes don't affect a running rental.
Inference
Per token, input and output priced separately, with a minimum charge of $0.01 per request. Prices vary by model.
Live, machine-readable prices
The figures above can change. The authoritative source, which billing itself uses, is public and needs no key:
curl -s https://api.green-compute.com/platform/pricing
{
"currency": "USD",
"gpu_rental": {
"billing": "per minute while running, per GPU",
"deployment_fee_usd": 0,
"gpus": [
{"gpu_model": "rtx4090", "vram_gb": 24, "cents_per_gpu_hour": 40, "usd_per_gpu_hour": 0.4},
{"gpu_model": "rtx5090", "vram_gb": 32, "cents_per_gpu_hour": 70, "usd_per_gpu_hour": 0.7}
]
},
"inference": {
"billing": "per token; minimum charge per request",
"minimum_charge_usd": 0.01,
"models": [{"model": "…", "usd_per_million_input_tokens": 0.2, "usd_per_million_output_tokens": 0.6}]
}
}
Paying
Top up on the Billing page with a card or TAO. Your balance is drawn down per minute by rentals and per request by inference.
Green Compute