查看所有模型

在 Tabbit 中使用 Kimi K2.5

在 Tabbit 中使用

在 Tabbit 中使用 Kimi K2.5

Kimi K2.5 · 模型速览

先看资源与证据,再决定怎么用

这里集中放置已审核的任务指南、公开测评和证据边界。客户端是否开放模型,仍以当前账户为准。

官方来源
任务指南2
测评来源4
来源已核对2
编辑精选5

Tabbit 模型与权限需以当前账户核实。

按任务找指南

把资料整理成结构化结果,制作视觉原型,或开始一个编码任务。

全部提示词与工作流
编程 · Agent 工作流来源已核对

Kimi K2.5 的视觉编码与 Agent Swarm 任务提示

Kimi 官方将 K2.5 的视觉输入、代码任务和 Agent Swarm 分解连接起来,并要求用完成检查和证据回收约束并行执行。

准备输入
task goal、source or reference material、runtime constraints、acceptance criteria
运行环境
Kimi K2.5 client or API; confirm the live model ID, tools, permissions, and version before execution.
查看步骤
reasoning · Agent 工作流未核验

Kimi K2.5 Thinking/Instant 与视觉工具配置

Kimi 官方仓库区分 Thinking 与 Instant 的参数和视觉工具配置,提醒部署时同时核对 chat template 与 reasoning_content。

准备输入
task goal、source or reference material、runtime constraints、acceptance criteria
运行环境
Kimi K2.5 client or API; confirm the live model ID, tools, permissions, and version before execution.
查看步骤

阅读证据与选型边界

公开结果来自不同版本、档位和测试方法;未知值保持未知。

全部测评与来源
Kimi Tech Blog / Visual Agentic Intelligence厂商自报

Kimi K2.5 官方发布:多模态、Agent Swarm 与编码基准

Kimi 官方发布把 K2.5 定位为视觉、编码和 Agent Swarm 模型,并公开 Thinking、工具、上下文和部分 benchmark 条件。

证据类型
厂商自报
边界
支持理解官方能力定位和 harness 边界;不支持把发布分数外推为第三方部署或生产成功率。
Fireworks AI独立测量

Fireworks 对 Kimi K2.5 官方 API 与部署栈的质量对照

Fireworks 用 Kimi 官方 API 做部署复测,指出 chat template、EOS、reasoning_content、采样和负载错误都会改变 tool-call 质量。

证据类型
独立测量
边界
支持把 provider/harness 配置纳入 K2.5 复测;不支持将其代表所有云端或本地部署。
BenchLM独立测量

BenchLM 对 Kimi K2.5 的公开基准账本与任务分层

BenchLM 的 Kimi K2.5 动态账本汇总 Coding、Agentic、Reasoning 和 Multimodal 来源,但其总分和排名是自定义聚合。

证据类型
独立测量
边界
支持按原始 benchmark 分层比较;不支持把 BenchLM 总分当作统一受控排名或当前价格。

Moonshot AI

在 Tabbit 中使用 Kimi K2.5

汇总 Kimi K2.5 的提示词指南、测评与社区反馈,帮助你判断适合的任务并直接开始使用。