【Model Launch】DeepSeek-V4.1-Flash
- Model ID:
deepseek-ai/deepseek-v4.1-flash - Model description: DeepSeek-V4.1-Flash is designed for coding assistance, terminal and computer-use agents, long-horizon complex tasks, and long-context analysis. It also natively supports image understanding, making it suitable for applications that need to balance response speed, task performance, and cost efficiency.
【Model Launch】DeepSeek-V4-Flash-0731
- Model ID:
deepseek-ai/deepseek-v4-flash-0731 - Model description: DeepSeek-V4-Flash-0731 is designed for general chat, coding assistance, and complex task processing scenarios. It is suitable for applications that need to balance response speed and reasoning quality.
【Model Launch】GLM-5.2 PD Disaggregated Version
- Model ID:
glm-5.2 - Model description: The GLM-5.2 PD disaggregated version uses a Prefill / Decode disaggregated deployment architecture. It is suitable for long-context, high-throughput, or stability-sensitive inference workloads.

