Pricing

GPU rental

GPUVRAMPer GPU-hour
RTX 4090 (rtx4090)24 GB$0.40
RTX 5090 (rtx5090)32 GB$0.70
  • Billed per minute while a deployment is ready. Image pulls, start-up and failed deployments are free.
  • No deployment fee, no egress fee, no minimum term. Stop with DELETE /platform/deployments/{id} and billing stops.
  • To start a rental, your balance must cover one hour of its price.
  • The price is fixed when a GPU is assigned to your deployment, so later list-price changes don't affect a running rental.

Inference

Per token, input and output priced separately, with a minimum charge of $0.01 per request. Prices vary by model.

Live, machine-readable prices

The figures above can change. The authoritative source, which billing itself uses, is public and needs no key:

curl -s https://api.green-compute.com/platform/pricing
{
  "currency": "USD",
  "gpu_rental": {
    "billing": "per minute while running, per GPU",
    "deployment_fee_usd": 0,
    "gpus": [
      {"gpu_model": "rtx4090", "vram_gb": 24, "cents_per_gpu_hour": 40, "usd_per_gpu_hour": 0.4},
      {"gpu_model": "rtx5090", "vram_gb": 32, "cents_per_gpu_hour": 70, "usd_per_gpu_hour": 0.7}
    ]
  },
  "inference": {
    "billing": "per token; minimum charge per request",
    "minimum_charge_usd": 0.01,
    "models": [{"model": "…", "usd_per_million_input_tokens": 0.2, "usd_per_million_output_tokens": 0.6}]
  }
}

Paying

Top up on the Billing page with a card or TAO. Your balance is drawn down per minute by rentals and per request by inference.