Product updates and changes.
Follow changes to model access, pricing, supported features, and integration requirements.
· Standard One
Standard One arrives in two open-weight sizes.
Standard One 3B and Standard One 8B read a message, an image or both, and answer each question with one of your options and a probability. Both are open weights under Apache 2.0 on Hugging Face. Rates are $0.019 and $0.049 per 1M text input tokens, 50% off list, and output is not billed. The accuracy, latency and calibration results in the write-up are measured on the served endpoint.
· API
GLM-5.3-Flash joins the API model list.
GLM-5.3-Flash takes text, image and video input with a 1.31M-token context, at $0.1275 per 1M input tokens and $0.425 per 1M output tokens. GLM-5.2, Kimi-K2.6 and Kimi-K2.7-Code are no longer listed.
· Resources
Two guides for deciding with Standard One.
The Resources library adds a guide to setting the confidence threshold at which an LLM’s answer ships without review, and a comparison of when a fine-tuned classifier beats an LLM on accuracy, speed and cost. Standard One’s figures come from its model cards, and every third-party figure is quoted from its source with the date it was read.
· Resources
A comparison of how coding plans count usage.
The Resources library adds a comparison of six coding subscriptions: what one unit of allowance counts at each vendor, whether the formula behind it is published, and when each 5-hour, weekly or monthly window refills. Every third-party figure is quoted from the vendor’s own page with the date it was read.
· Resources
Two references for comparing reserved capacity.
The Resources library adds a comparison of how nine providers sell reserved capacity, covering the unit each one measures, the billing clock, the minimum term, and whether the price is published, and a glossary of the capacity and plan terms those offerings use. Every third-party figure is quoted from the vendor’s own page with the date it was read.
· Resources
Three resources for sizing a Product Plan.
The Resources library adds a Product Plan capacity and break-even calculator, a guide to buying AI inference for a product on a monthly plan, and a guide to what an aggregate output rate means for a product. Every third-party figure is quoted from the vendor’s own page with the date it was read. Further comparisons and guides follow over the coming weeks.
· Site
Product Plan and Coding Plan have new addresses.
The two plan pages now live at /product-plan/ and /coding-plan/, and the pricing page’s sections and contact topics use the same names. The previous addresses redirect permanently, including links with query strings.
· API
Model cards show input modalities and rates.
Each card in the model catalog now names the input types the model accepts and its Standard API rates, so a model can be compared without leaving the page.
· Legal
Terms and data processing policy revised.
Zero Data Retention is stated to apply only where expressly identified for a service or configuration, optional storage through the Product Plan data platform is described as customer-directed, and Section 5.3 of the Terms now sets the conditions under which Service-Specific Terms may permit model training: the covered service and content, the purposes, the retention period, and the controls must be named before that use applies.
· Pricing
Standard API rates show each developer’s list price.
The rate table shows each model developer’s published list price beside our rate, with the date the list prices were checked. DeepSeek-V4.1-Flash joined the catalog and the retired DeepSeek-V4-Flash-0731 was removed.
· API documentation
API contract and deployment terms clarified.
The site now describes the API contract and clarifies that regional processing controls apply when included in an order.