Inference and agents, in the KingdomPay per token · Saudi RiyalDCP-Agent for Saudi business · agents.dcp.saAgents can rent a GPU · npx -y github:dhnpmp-tech/dcp-mcpEarn Riyal from your GPUPDPL · Saudi data residencyInference and agents, in the KingdomPay per token · Saudi RiyalDCP-Agent for Saudi business · agents.dcp.saAgents can rent a GPU · npx -y github:dhnpmp-tech/dcp-mcpEarn Riyal from your GPUPDPL · Saudi data residency
DCP
Sign inStart free →
DCP
Explore
01Overviewنظرة عامةSovereign Arabic AI runtime02InferenceالاستدلالOpenAI-compatible · live catalog03Fine-Tuningالضبط الدقيقLoRA contracts · proof-gated serving04BatchالدُفعاتJSONL validation · execution gated05DeploymentsالنشرEndpoint proof · traffic gated06BenchmarksالقياساتMeasured rows · quality claims gated07PricingالأسعارPer-million-token · SAR08GPU Podsحاويات GPURent a whole GPU on demand09AgentsالوكلاءZero-human onboarding · MCP10DocsالتوثيقOpenAI-compatible API11EarnاكسبEarn Riyal from your GPU12SupportالدعمTalk to the team
In-Kingdom · PDPL© 2026 · Riyadh

Pay for what you use. Refunded when you stop.

No procurement, no quota, no flat monthly GPU. Inference is billed per million tokens; pods are billed per GPU-second, cost-plus from the live market, and the unused time is refunded the instant you stop. Everything is priced in Saudi Riyal, shown before you commit.

Inference APIPer million tokens5halala / 1M fromapi.dcp.sa/v1 · OpenAI-compatible
GPU PodsPer GPU-second2.5SAR / hr fromRefunded on stop · root + SSH + Jupyter
New accountsStarter credit100SAR · no cardFund later in Saudi Riyal
SubscriptionsMonthly plans375SAR / mo fromDiscounted token allowance
this page rented a pod when you arrived00:000.0000 SARsimulated · real pods bill exactly like this
Rate by model class — halala per 1 million tokensHalala = 1/100 SAR
Embeddingembeddings5
Tinychat model15
Smallchat model30
Mediumchat model150
LargeDeepSeek V4 Pro400

Each chat-completion response also returns per-call usage pricing in USD and SAR.

Live API catalog — from /v1/modelschecking live catalog
Loading live model pricing...
On-demand GPU types — indicative SAR / hour, fromBilled per second · refunded on stop
NVIDIA H200141 GB23.05
NVIDIA H10080 GB17.27
NVIDIA A10080 GB7.30
NVIDIA L40S48 GB5.20
NVIDIA RTX 509032 GB5.20
NVIDIA RTX 409024 GB3.62
NVIDIA RTX 309024 GB2.50

GPU pod prices are cost-plus from the live market and refresh every few minutes; the rate at launch is the rate you pay for that pod.

Starter

375 SAR / mo

15% off
  • Discounted token allowance
  • Per-second pod billing
  • Email support
Discounted allowanceStart →
Growth

1,500 SAR / mo

22% off
  • Larger discounted allowance
  • Priority pod scheduling
  • Workspace sharing
Discounted allowanceStart →
Scale

5,625 SAR / mo

30% off
  • Max discounted allowance
  • Reserved capacity option
  • Dedicated CSM
Discounted allowanceStart →
Enterprise

Custom

VPC · DPA · MSA

Run it in your own VPC, with a DPA, MSA, and data-flow appendix. Dedicated capacity and a CSM.

Sovereignty preservedTalk to sales →

Pay-as-you-go remains the default — subscriptions are optional and do not lock you in. Unused subscription tokens do not roll over.

How is GPU rental priced on DCP?

GPU rental is billed prepaid per GPU-second in Saudi Riyal, cost-plus from the live market. Indicative on-demand hourly rates: NVIDIA RTX 3090 from 2.5 SAR/hr, RTX 4090 from 3.62 SAR/hr, RTX 5090 from 5.2 SAR/hr, L40S (48 GB) from 5.2 SAR/hr, A100 (80 GB) from 7.3 SAR/hr, H100 (80 GB) from 17.27 SAR/hr, and H200 (141 GB) from 23.05 SAR/hr. You are billed only for the seconds a verified GPU is actually serving you, and the unused time is refunded the instant you stop.

How is inference billed?

Inference is billed per million tokens in Saudi Riyal. Rates are by model class — from about 5 halala per 1M tokens for embedding models, 15 halala for tiny, 30 for small, 150 for medium, up to around 400 halala for large. Each chat-completion response carries per-call usage pricing in both USD and SAR. New renter accounts start with 100 SAR of credit and no card is required to begin.

Is there a subscription plan, or is it all pay-as-you-go?

Both. Pay-as-you-go is the default — per million tokens for inference and per GPU-second for pods, with a prorated refund when you stop early. Optional monthly subscriptions give a discounted token allowance for teams with steady usage: Starter at 375 SAR/mo, Growth at 1,500 SAR/mo, and Scale at 5,625 SAR/mo. Unused subscription tokens do not roll over.

What happens if my balance runs out mid-session?

A chargeable call returns HTTP 402 with a machine-readable body ({ code: "insufficient_balance", required_sar, balance_sar, topup_url, retryable: true }). No pod or charge is created, so you can top up and retry safely. Running pods are stopped (not killed silently) so you can resume after topping up.

Are prices fixed, or can they change?

GPU pod prices are cost-plus from the live market and refresh every few minutes, so the per-second rate you see at launch is the rate you pay for that pod. Per-token inference rates are stable per model class. Prices are always shown in SAR before you commit; nothing is billed opaquely.

Do I pay for data egress or cross-border transfer?

No. Inference, pods, and storage run in-Kingdom on Saudi-owned hardware, so there is no egress fee. Cross-border frontier models are available only by explicit per-tenant opt-in and never incur hidden transfer charges — any cross-border cost is disclosed before you enable it.

Start with 100 SAR. No card.