Skip to content

Product updates and changes.

Follow changes to model access, pricing, supported features, and integration requirements.

· Standard One

Standard One arrives in two open-weight sizes.

Standard One 3B and Standard One 8B read a message, an image or both, and answer each question with one of your options and a probability. Both are open weights under Apache 2.0 on Hugging Face. Rates are $0.019 and $0.049 per 1M text input tokens, 50% off list, and output is not billed. The accuracy, latency and calibration results in the write-up are measured on the served endpoint.

Meet Standard One

· API

GLM-5.3-Flash joins the API model list.

GLM-5.3-Flash takes text, image and video input with a 1.31M-token context, at $0.1275 per 1M input tokens and $0.425 per 1M output tokens. GLM-5.2, Kimi-K2.6 and Kimi-K2.7-Code are no longer listed.

Compare API rates

· Resources

Two guides for deciding with Standard One.

The Resources library adds a guide to setting the confidence threshold at which an LLM’s answer ships without review, and a comparison of when a fine-tuned classifier beats an LLM on accuracy, speed and cost. Standard One’s figures come from its model cards, and every third-party figure is quoted from its source with the date it was read.

Set a confidence threshold

· Resources

A comparison of how coding plans count usage.

The Resources library adds a comparison of six coding subscriptions: what one unit of allowance counts at each vendor, whether the formula behind it is published, and when each 5-hour, weekly or monthly window refills. Every third-party figure is quoted from the vendor’s own page with the date it was read.

Compare coding plan units

· Resources

Two references for comparing reserved capacity.

The Resources library adds a comparison of how nine providers sell reserved capacity, covering the unit each one measures, the billing clock, the minimum term, and whether the price is published, and a glossary of the capacity and plan terms those offerings use. Every third-party figure is quoted from the vendor’s own page with the date it was read.

Compare reserved capacity

· Resources

Three resources for sizing a Product Plan.

The Resources library adds a Product Plan capacity and break-even calculator, a guide to buying AI inference for a product on a monthly plan, and a guide to what an aggregate output rate means for a product. Every third-party figure is quoted from the vendor’s own page with the date it was read. Further comparisons and guides follow over the coming weeks.

Browse the resources

· Site

Product Plan and Coding Plan have new addresses.

The two plan pages now live at /product-plan/ and /coding-plan/, and the pricing page’s sections and contact topics use the same names. The previous addresses redirect permanently, including links with query strings.

Open Product Plan

· API

Model cards show input modalities and rates.

Each card in the model catalog now names the input types the model accepts and its Standard API rates, so a model can be compared without leaving the page.

See the model catalog

· Legal

Terms and data processing policy revised.

Zero Data Retention is stated to apply only where expressly identified for a service or configuration, optional storage through the Product Plan data platform is described as customer-directed, and Section 5.3 of the Terms now sets the conditions under which Service-Specific Terms may permit model training: the covered service and content, the purposes, the retention period, and the controls must be named before that use applies.

Read the data processing policy

· Pricing

Standard API rates show each developer’s list price.

The rate table shows each model developer’s published list price beside our rate, with the date the list prices were checked. DeepSeek-V4.1-Flash joined the catalog and the retired DeepSeek-V4-Flash-0731 was removed.

Compare API rates

· API documentation

API contract and deployment terms clarified.

The site now describes the API contract and clarifies that regional processing controls apply when included in an order.

Read API documentation · Review data processing terms