# NEAR AI Cloud pricing

NEAR AI Cloud charges per token, with no minimums and no subscriptions.
Input and output tokens are priced separately per 1,000,000 tokens, and the
rate depends on which model you call. Every model below runs inside a
hardware-secured enclave, and the attestation on each response is included
in the price.

Canonical page: https://near.ai/pricing
Generated: 2026-10-06T22:44:30.725Z

## How pricing works

- You pay per token. Input and output are priced separately, quoted per 1M tokens.
- The rate depends on which model you call. There is no platform fee on top of the model rate.
- No minimums, no subscriptions, and no seat licences on pay-as-you-go.
- Confidential execution and the hardware-signed attestation on every response are included in the rate, not billed as an add-on.
- These are open-weight models, so the rate tracks a competitive market rather than a single vendor's list price.
- Enterprise agreements add reserved GPU capacity, private models, volume rates and SLAs.

## Price per model

All rates in US dollars per 1,000,000 tokens.

| Model | Model ID for the API | Context | Privacy tier | Input USD/1M | Output USD/1M |
| --- | --- | --- | --- | --- | --- |
| GLM 5.3 Flash | z-ai/glm-5.3-flash | 1048K | Confidential TEE | $0.15 | $0.5 |
| Qwen3-VL-30B-A3B-Instruct | Qwen/Qwen3-VL-30B-A3B-Instruct | 16K | Confidential TEE | $0.15 | $0.55 |
| Qwen 3.6 35B A3B FP8 | Qwen/Qwen3.6-35B-A3B-FP8 | 262K | Confidential TEE | $0.17 | $1.1 |
| Qwen 3.8 27B | Qwen/Qwen3.8-27B | 262K | Confidential TEE | $0.44 | $3.3 |
| qwen3.5-397b-a17b | qwen/qwen3.5-397b-a17b | 128K | 3P Confidential TEE | $0.5 | $3.3 |
| Kimi K2.6 | moonshotai/kimi-k2.6 | 262K | 3P Confidential TEE | $0.81 | $3.85 |
| deepseek-v3.2 | deepseek/deepseek-v3.2 | 128K | 3P Confidential TEE | $1.1 | $1.1 |
| Kimi K3 | moonshotai/kimi-k3 | 1048K | 3P Confidential TEE | $3.3 | $16.5 |

## Enterprise

Enterprise pricing is quoted rather than listed. It covers reserved GPU
capacity, custom and private models, volume rates with SLAs, on-premise or
VPC deployment, and solutions engineering.

## Pricing questions

### How does NEAR AI Cloud pricing work?

You pay per token, with input and output tokens priced separately and quoted per 1M tokens. The rate depends on which model you call. There are no minimums, subscriptions or seat licences on pay-as-you-go, so a request costs the model's input rate for the prompt plus its output rate for the completion.

### Does confidential compute or attestation cost extra?

No. Every model listed on this page runs inside a hardware-secured trusted execution environment with attestation. Both are included in the per-token rate rather than billed as an add-on.

### Is there a minimum spend or subscription?

No. Pay-as-you-go has no minimums and no subscriptions. You are billed for the tokens you actually use.

### What is the difference between Confidential TEE and 3P Confidential TEE?

Confidential TEE models run inside NEAR AI's own trusted execution environments. 3P Confidential TEE models run inside an attested third-party provider's enclave. Both are hardware-isolated and attested, but the verification tools available differ by tier, so check the documentation for the model you use. The table on this page labels which is which.

### Can I control what each person on my team spends?

Yes. Create an API key for each person and set a spend cap on the key as you create it. Usage is reported per key, so you can see what each one costs. There are no per-seat charges on pay-as-you-go, so you pay for the tokens each key actually uses rather than for the number of people who have access.

### How is enterprise pricing different?

Enterprise agreements are quoted rather than listed. They cover reserved GPU capacity, custom and private models, volume pricing with SLAs, on-premise or VPC deployment, and solutions engineering.

### Where do the prices on this page come from?

They are read directly from the public NEAR AI Cloud catalog API at request time. This page is generated on the server, so the rates shown here are the current published rates rather than a hand-maintained copy.

### Can we pay by invoice instead of prepaying for credits?

Approved customers can be invoiced at the end of the month rather than prepaying for credits. Qualification covers business verification, a billing and compliance review, and expected usage. Contact the team to apply.

### How do I start paying for NEAR AI Cloud?

Create an API key in the NEAR AI Cloud console at cloud.near.ai. The API is OpenAI-compatible, so you can point an existing client at the NEAR AI base URL without changing your code.
