New Standard One: text or image in, fast decision out

Open model inference for
products and coding

Pay per token. Or get a plan for your product or coding.

Building a product? Choose how to pay.

API and Product Plan both power AI features in your product. Choose between token-based billing and a monthly subscription.

API · Pay per token

Usage-based billingPay for tokens used

Your bill follows the input and output tokens your product uses.

Explore API

Product Plan · A monthly base

Guaranteed baseOptional extra capacity

No monthly token cap. Use free extra capacity when available, or enable auto scaling for traffic peaks.

Explore Product Plan

Explore the model lineup.

Eight open models from six labs, plus our own Standard One 3B and 8B, all through one API.

Explore API models

Kimi K3, GLM-5.3, and DeepSeek V4.1 Flash are also in Product Plan and Coding Plan.

Standard One turns text and images into decisions.

Ask a question about a message or an image. One 3B and One 8B answer with one of your own options and say how sure they are.

You set the bar. Your code acts on answers above the threshold you choose, and a person gets the rest.

Open weights, or our API. Download both sizes from Hugging Face under Apache 2.0, or call the API from $0.019 per 1M input tokens. Output tokens are free.

Meet Standard One
standard-one-3b

Your data stays under your control.

We process your inputs and outputs to serve your requests. Retention depends on the service and features you use.

Operational metadata. Content-free metadata is kept for billing, reliability, and service operations.

Product Plan data platform. Optionally store your product’s data on Standard Thinking. Storage stays off until you turn it on, and you control access and how it is used.

Read the data processing policy
Your application sends requests to Standard Thinking inference, with retention governed by the applicable service terms. Data you designate is stored on the data platform for your use only, and storage is optional. Standard Thinking processes it to provide your requested services, not for its own independent purposes. Your application sends requests to Standard Thinking inference, with retention governed by the applicable service terms. Data you designate is stored on the data platform for your use only, and storage is optional. Standard Thinking processes it to provide your requested services, not for its own independent purposes.

Your data is for your use only, and storage is optional.