Model List
MaaS provides multiple models for you to use. You can easily integrate model services into your business by following tutorials or API documentation.
Text Generation
Tutorial: Text Generation; API: Sending a Chat Request (Chat/Post); price: MaaS Text Generation Models.
| Model Parameter (Model ID) | Supported Capabilities | Length Limit | Default Rate Limit |
|---|---|---|---|
| glm-5.3 | Deep thinking (cannot be disabled) Function calling Prefix caching | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 128K |
|
| glm-5.2 | Deep thinking (configurable) Function calling | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 64K |
|
| glm-5.1 | Deep thinking (configurable) Function calling | Context length: 198K Max input length: 192K Max output length: 128K Max CoT length: 96K |
|
| deepseek-v4.1-flash | Deep thinking (configurable) Function calling Prefix caching | Context length: 1M Max input length: 1M Max output length: 384K Max CoT length: 96K |
|
| deepseek-v4-flash | Deep thinking (configurable) Function calling Prefix completion | Context length: 1M Max input length: 1M Max output length: 384K Max CoT length: 96K |
|
| deepseek-v4-pro | Deep thinking (configurable) Function calling Prefix completion | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 96K |
|
Deep Thinking
Tutorial: Deep Thinking; API: Sending a Chat Request (Chat/Post); price: MaaS Text Generation Models.
| Model Parameter (Model ID) | Supported Capabilities | Length Limit | Default Rate Limit |
|---|---|---|---|
| glm-5.3 | Deep thinking (cannot be disabled) Function calling Prefix caching | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 128K |
|
| glm-5.2 | Deep thinking (configurable) Function calling | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 64K |
|
| glm-5.1 | Deep thinking (configurable) Function calling | Context length: 198K Max input length: 192K Max output length: 128K Max CoT length: 96K |
|
| deepseek-v4.1-flash | Deep thinking (configurable) Function calling Prefix caching | Context length: 1M Max input length: 1M Max output length: 384K Max CoT length: 96K |
|
| deepseek-v4-flash | Deep thinking (configurable) Function calling Prefix completion | Context length: 1M Max input length: 1M Max output length: 384K Max CoT length: 96K |
|
| deepseek-v4-pro | Deep thinking (configurable) Function calling Prefix completion | Context length: 1M Max input length: 1M Max output length: 128K Max CoT length: 96K |
|
What is your overall rating for this page?
Thank you very much for your feedback. We will continue working to improve the documentation.See the reply and handling status in My Cloud VOC.
For any further questions, feel free to contact us through the chatbot.
Chatbot