Skip to main content
DeepSeek
DeepSeek-V4.1-Flash is now available

【Model Launch】DeepSeek-V4.1-Flash

  • Model ID: deepseek-ai/deepseek-v4.1-flash
  • Model description: DeepSeek-V4.1-Flash is designed for coding assistance, terminal and computer-use agents, long-horizon complex tasks, and long-context analysis. It also natively supports image understanding, making it suitable for applications that need to balance response speed, task performance, and cost efficiency.
DeepSeek
DeepSeek-V4-Flash-0731 is now available

【Model Launch】DeepSeek-V4-Flash-0731

  • Model ID: deepseek-ai/deepseek-v4-flash-0731
  • Model description: DeepSeek-V4-Flash-0731 is designed for general chat, coding assistance, and complex task processing scenarios. It is suitable for applications that need to balance response speed and reasoning quality.
GLM
GLM-5.2 PD disaggregated version is now available

【Model Launch】GLM-5.2 PD Disaggregated Version

  • Model ID: glm-5.2
  • Model description: The GLM-5.2 PD disaggregated version uses a Prefill / Decode disaggregated deployment architecture. It is suitable for long-context, high-throughput, or stability-sensitive inference workloads.