Three Chinese labs shipped low-latency reasoning tiers in the same window: Zhipu's GLM-5.3-FlashX, Xiaomi's Mimo v2.6 (Flash and Pro), and StepFun's Step-5 Preview. Together with Alibaba's Qwen Omni Turbo, that's five new chat models now live on NovAI. Here is how they price out and where each one fits.
| Model | Lab | Input / 1M | Output / 1M | Context |
|---|---|---|---|---|
| mimo-v2.6-flash | Xiaomi | $0.10 | $0.28 | 128K |
| glm-5.3-flashx | Zhipu AI | $0.12 | $0.40 | 128K |
| qwen-omni-turbo | Alibaba | $0.30 | $0.60 | 32K |
| mimo-v2.6-pro | Xiaomi | $0.43 | $0.87 | 128K |
| step-5-preview | StepFun | $0.40 | $1.20 | 128K |
For reference inside each family: GLM-5.3-FlashX sits an order of magnitude below Zhipu's flagship GLM-5.3 ($1.25/$4.00), and Mimo v2.6 Flash undercuts Mimo v2.6 Pro by roughly 4x on input - the classic flash/pro ladder, now available from both labs.
from openai import OpenAI
client = OpenAI(base_url="https://aiapi-pro.com/v1", api_key="YOUR_NOVAI_API_KEY")
resp = client.chat.completions.create(
model="glm-5.3-flashx", # or mimo-v2.6-flash / step-5-preview / qwen-omni-turbo
messages=[{"role": "user", "content": "Summarize this contract in 5 bullets: ..."}],
)
print(resp.choices[0].message.content)
No SDK changes, no new keys: all five models appear in GET /v1/models and bill per token at the prices above, with zero platform fee.
$2 free credit on signup. Zero platform fee, provider list prices, OpenAI-compatible API.
Sign Up Free