← 模型库

deepseek-ai/deepseek-ai/DeepSeek-V3.1-Terminus

100B 以上持续维护
官方来源 ↗
发布日期 未标注·最后更新 08-09·数据来源 官方仓库 / 官网·已核对官方链接
这是你发布的模型吗?认领后可标注官方身份并长期维护主页信息。
入选场景榜📄 长文本 57.4
任务类型基座模型
参数规模671B (MoE, 激活 37B)激活 37B
上下文长度163840
开源协议暂无可靠数据
商用情况暂无可靠数据
推荐显存暂无可靠数据
主要语言暂无可靠数据
本地部署未标注

模型介绍

中文速览

发布方 deepseek-ai
上下文长度 163840

资料来源:DeepInfra 模型库,以下正文为官方原始说明。

本条目的官方说明为英文,以下内容保留原文,关键章节标题已中文标注。

deepseek-ai/DeepSeek-V3.1-Terminus

DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model’s original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model’s performance in coding and search agents. It is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes. It extends the DeepSeek-V3 base with a two-phase long-context training process. Users can control the reasoning behaviour with the reasoning enabled boolean. Learn more in our docs The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows.

  • 上下文长度:163,840 tokens
  • 最大输出:163,840 tokens
  • 标签:chat、prompt_cache、reasoning_effort、reasoning
  • 定价(每百万 token,USD):输入 $0.2700 / 输出 $0.9500
  • 提供方:deepseek-ai
  • 来源:DeepInfra

适合场景

暂无可靠数据。

已知限制

  • MoE 架构:显存需按 总参数 准备,激活参数只影响推理速度,不减少常驻显存。
  • 本站未登记确切发布日期,版本时间线可能不完整。
  • 本页评测与硬件数据为区间估算,未做统一环境实测,不能替代官方评测报告。

评测

暂无可靠数据。本站只收录标注了「评测来源、模型版本、是否官方数据、测试日期」的成绩, 不混合不同设置下的分数,也不把单一 Benchmark 解释为整体能力。