# Pricing

## GPU rental

| GPU | VRAM | Per GPU-hour |
|---|---|---|
| RTX 4090 (`rtx4090`) | 24 GB | **$0.40** |
| RTX 5090 (`rtx5090`) | 32 GB | **$0.70** |

- **Billed per minute** while a deployment is `ready`. Image pulls, start-up and failed deployments are free.
- **No deployment fee, no egress fee, no minimum term.** Stop with `DELETE /platform/deployments/{id}` and billing stops.
- To start a rental, your balance must cover **one hour** of its price.
- The price is fixed when a GPU is assigned to your deployment, so later list-price changes don't affect a running rental.

## Inference

Per token, input and output priced separately, with a minimum charge of $0.01 per request. Prices vary by model.

## Live, machine-readable prices

The figures above can change. The authoritative source, which billing itself uses, is public and needs no key:

```bash
curl -s https://api.green-compute.com/platform/pricing
```

```json
{
  "currency": "USD",
  "gpu_rental": {
    "billing": "per minute while running, per GPU",
    "deployment_fee_usd": 0,
    "gpus": [
      {"gpu_model": "rtx4090", "vram_gb": 24, "cents_per_gpu_hour": 40, "usd_per_gpu_hour": 0.4},
      {"gpu_model": "rtx5090", "vram_gb": 32, "cents_per_gpu_hour": 70, "usd_per_gpu_hour": 0.7}
    ]
  },
  "inference": {
    "billing": "per token; minimum charge per request",
    "minimum_charge_usd": 0.01,
    "models": [{"model": "…", "usd_per_million_input_tokens": 0.2, "usd_per_million_output_tokens": 0.6}]
  }
}
```

## Paying

Top up on the [Billing](https://www.green-compute.com/billing) page with a card or TAO. Your balance is drawn down per minute by rentals and per request by inference.
