When to use it
Try GLM when a plan needs to account for several things at once, or when you want to compare options in detail.
Z.ai model in April
Several reasoning settings for questions that involve multiple steps or competing priorities.
You choose the model and its reasoning setting. Its answer also depends on your question and the April information you let it use.
Try GLM when a plan needs to account for several things at once, or when you want to compare options in detail.
GLM uses the relevant April information you have enabled. Choose Off for a direct response, or set reasoning to Minimal, Low, Medium or High.
GLM is in the middle of April's provider-cost range. The route and provider price can change, so April measures actual usage for each request.
Higher reasoning settings take longer. They do not make an answer a substitute for medical, legal or financial advice.
For simpler questions, try Qwen or Luna with reasoning turned off.
These provider prices were checked on 9 September 2026 and may change. April's costs also depend on taxes, routing, caching and request length. You are not billed per token. Your model allowance is shown in the app. Source: Z.ai model pricing.
The model receives your message and the relevant April information you have enabled. Answers can be incomplete, out of date or wrong. They are not medical advice.
Temporary Chat does not use saved memory or previous chats, and it does not appear in your chat history. An encrypted copy of the response may be kept for up to 24 hours to recover interrupted delivery.
OpenAI analyses attachments first. Your selected model receives a description of the relevant content, rather than the original file.
Check important answers before acting on them. Do not rely on a model to assess an emergency, diagnose a condition, make medication decisions or replace professional advice.
Model and provider names are trademarks of their respective owners. Inclusion does not imply endorsement.