API · Pay per token
Your bill follows the input and output tokens your product uses.
Explore APINew Standard One: text or image in, fast decision out
Pay per token. Or get a plan for your product or coding.
Use inference on demand, with no capacity commitment. Pay only for the tokens you use.
Explore APIReserve steady base capacity and scale for peaks. Keep your product’s data on our platform, under your control.
Explore Product PlanSwitch between models in your coding tool on one plan with no usage cap. $149 a week or $549 a month, per developer.
Explore Coding PlanAPI and Product Plan both power AI features in your product. Choose between token-based billing and a monthly subscription.
Your bill follows the input and output tokens your product uses.
Explore APINo monthly token cap. Use free extra capacity when available, or enable auto scaling for traffic peaks.
Explore Product PlanEight open models from six labs, plus our own Standard One 3B and 8B, all through one API.
Explore API modelsKimi K3, GLM-5.3, and DeepSeek V4.1 Flash are also in Product Plan and Coding Plan.
Ask a question about a message or an image. One 3B and One 8B answer with one of your own options and say how sure they are.
You set the bar. Your code acts on answers above the threshold you choose, and a person gets the rest.
Open weights, or our API. Download both sizes from Hugging Face under Apache 2.0, or call the API from $0.019 per 1M input tokens. Output tokens are free.
We process your inputs and outputs to serve your requests. Retention depends on the service and features you use.
Operational metadata. Content-free metadata is kept for billing, reliability, and service operations.
Product Plan data platform. Optionally store your product’s data on Standard Thinking. Storage stays off until you turn it on, and you control access and how it is used.
Your data is for your use only, and storage is optional.