Models API
Models are billed per usage:
Pricing varies by model. Check the Model Catalog in the dashboard for current per-model rates.
GPU Instances
Instances are billed per hour while running. Approximate rates:Prices are approximate and may vary based on availability. Check the deployment page for current pricing when you configure an instance.
Cloud Storage
No egress fees. You only pay for provisioned volume size.
Billing Rules
Cost Optimization Tips
Terminate Idle Instances
Instances are billed while running, even if idle. Terminate instances you are not actively using.
Use the Right GPU
Do not pay for an H100 when an RTX 4090 handles your workload. Start small and scale up only if needed.
Enable Auto-Recharge
Avoid losing work from unexpected termination. Set a threshold that gives you enough buffer time.
Monitor Burn Rate
Check your billing dashboard regularly to track spend and adjust usage patterns.