DataLearner 标志

DeepSeek-V4-FlashvsGLM 5.1

在 6 个共同 benchmark 中,GLM 5.1 整体领先:DeepSeek-V4-Flash 领先 2 项,GLM 5.1 领先 4 项,持平 0 项,平均分差 -2.55。

DeepSeek-AI
DeepSeek-V4-Flash

DeepSeek-AI · 2026-04-24 · 推理大模型

智谱AI
GLM 5.1

智谱AI · 2026-03-27 · 推理大模型

DeepSeek-V4-Flash2 项(33%)(67%)4 项GLM 5.1

评测分数

按能力类目分组,每组内按分差大小排列;共 6 项。

写作与创意表达

GLM 5.1 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
Creative Writing1,55948 / 110Normal (No Tools)1,59243 / 110Normal (No Tools)-33.10

多轮记忆与持续上下文

GLM 5.1 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
Context Arena26.47117 / 126Normal (No Tools)30.29116 / 126Normal (No Tools)-3.82

常识推理

DeepSeek-V4-Flash 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
SimpleBench61.1032 / 96Normal (No Tools)55.1045 / 96Normal (No Tools)+6

自主开发与终端任务

GLM 5.1 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
Terminal Bench Hard34.1082 / 244Normal (With Tools)35.6070 / 244Normal (With Tools)-1.50

跨应用与工具编排

DeepSeek-V4-Flash 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
PinchBench v281.748 / 45Reported best (effort unspecified)59.9533 / 45Reported best (effort unspecified)+21.79

跨能力综合测试

GLM 5.1 领先 1/1
评测项DeepSeek-V4-FlashGLM 5.1分差
LiveBench65.4852 / 117Normal (No Tools)70.1837 / 117Normal (No Tools)-4.70

规格对比

字段DeepSeek-V4-FlashGLM 5.1
发布机构DeepSeek-AI智谱AI
发布时间2026-04-242026-03-27
模型类型推理大模型推理大模型
架构MoE 架构MoE 架构
参数规模2840亿7540亿
上下文长度1M200K
最大输出384K125K

API 调用价格

价格优先使用 DataLearner 配置的 API 记录;缺失项不做推测。

价格项DeepSeek-V4-FlashGLM 5.1
文本输入$0.14 / 1M tokens$1.4 / 1M tokens
文本输出$0.28 / 1M tokens$4.4 / 1M tokens
缓存读取$0.0028 / 1M tokens$0.26 / 1M tokens

小结

  • DeepSeek-V4-Flash在以下类目领先:常识推理 (1/1)、跨应用与工具编排 (1/1)
  • GLM 5.1在以下类目领先:写作与创意表达 (1/1)、多轮记忆与持续上下文 (1/1)、自主开发与终端任务 (1/1)、跨能力综合测试 (1/1)

6 个共同 benchmark 上,GLM 5.1 平均高出 2.55 分。

单项差距最大的 benchmark:Creative Writing — DeepSeek-V4-Flash 1,559,GLM 5.1 1,592(分差 -33.10)。

本页正文由结构化模型、价格与 benchmark 数据生成,不使用实时 LLM 撰写。