gpt-6-astra- Input
- $2.000
- Output
- $10.000
- Cache Write
- -
- Cache Read
- $0.200
Choose a (model + Delivery Lane) pair. We coordinate eligible delivery paths behind a single API for coding agents, AI applications, and automation workflows.
Delivery Lanes
Choose a Delivery Lane, then compare its published price with the official model baseline.
Economy Lane: A lower-cost general-purpose lane that favors savings over delivery consistency. Standard context. Standard processing.
| Model ID | Input | Output | Cache Write | Cache Read | Savings |
|---|---|---|---|---|---|
gpt-6-astra | $2.000 | $10.000 | - | $0.200 | 80.0% OFF |
gpt-5.6-sol | $0.440 | $2.200 | $0.550 | $0.044 | 89.0% OFF |
gpt-5.6-terra | $0.170 | $1.020 | $0.213 | $0.017 | 91.5% OFF |
gpt-5.6-luna | $0.056 | $0.336 | $0.070 | $0.006 | 72.0% OFF |
gpt-5.5 | $0.450 | $2.700 | - | $0.045 | 91.0% OFF |
gpt-5.4 | $0.233 | $1.395 | - | $0.023 | 90.7% OFF |
gpt-5.4-mini | $0.071 | $0.428 | - | $0.007 | 90.5% OFF |
gpt-6-astragpt-5.6-solgpt-5.6-terragpt-5.6-lunagpt-5.5gpt-5.4gpt-5.4-miniUsage observability
Review requested model, served model, Delivery Lane, token usage, cache activity, latency, throughput, and charge for each request.
Built for coding work
Long-running coding sessions magnify latency, cache behavior, and manual recovery. Keep those delivery concerns below the workflow.
How delivery works
Your workload chooses the public model and Delivery Lane. The policy coordinates eligible delivery paths while the implementation stays behind the API.
Start with a workload
Use the documented Base URL and a dedicated workspace API key, choose a Published model, Delivery Lane, and endpoint, then verify the first request.
Send the first coding task with a Published model, Delivery Lane, and endpoint, then review the completed request in Usage.
Send the first application request with a Published model, Delivery Lane, and endpoint, then review the completed request in Usage.
Run the first automation workflow with a Published model, Delivery Lane, and endpoint, then review the completed request in Usage.
Use Pricing to choose a Published request path, follow Quickstart for the first request, and inspect protocol scope in API Reference.
Trust and FAQ
Review content handling, request facts, pricing, availability, and failure behavior from the public records that define them.
Routed API prompts and responses are not written to 1api's application databases or long-term logs. A private continuation cache may temporarily retain the minimum required conversation content for response continuation and one bounded recovery attempt. Read Trust for the complete content boundary and Privacy for provider-side processing context.
Token counts, request status, timing, latency, and charge details stay auditable. Review the customer-facing record expectations in Trust.
Compare Lane prices with the official baseline. Where a model's public price evidence includes cache pricing, cache read or cache write is shown as a separate pricing line. Cached-token usage is shown in the request record when it applies. Review the selected model's public price evidence in Pricing.
Where the documented mechanism applies, another candidate can be attempted before a response is committed. The mechanism is not a guarantee that a fallback will occur. The workload should still handle an unsuccessful request. Read Docs for the request workflow.
Select a Published model, Delivery Lane, and endpoint before sending a request. Use Pricing for publication coverage and API Reference for protocol scope.
Read Trust and Privacy for the complete record and policy details.
Create an API key, choose a Delivery Lane, and follow Quickstart for your first request.