openai/gpt-5.6-luna
1050K context|$0.2/M input tokens|$1.2/M output tokens|$0.02/M cache read
OpenAI GPT-5.6 model optimized for cost-sensitive workloads
Input / M tokens
$0.2
Output / M tokens
$1.2
Cache read
$0.02
Context
1050K
Sample code and API for GPT-5.6 Luna
Use the Responses API (/v1/responses) for OpenAI models.
You can also use NEAR AI Cloud with OpenAI's client API:
Privacy & Execution
This is an Incognito model: it runs on a partner provider rather than inside a NEAR AI GPU TEE. Requests are routed through a shared NEAR AI API key, so your identity is not disclosed to the provider.
Your prompt content, however, is transmitted to and processed by OpenAI in accordance with their terms of service and privacy policy. Data is encrypted in transit, but execution happens outside of trusted hardware — there is no hardware-level privacy, attestation, or cryptographic proof of execution for this model.