可进入对比
档案摘要 / Decision snapshot
先判断能不能用于选型。
有真实任务证据
来自模型卡
边界优先
30 秒结论
可进入选型对比,但仍以案例证据为准。
来自公开模型资料库,并已补厂商/官方/模型家族证据;当前站内暂无可核验 A 类案例。
这页有来源、有风险说明、有公开可核验 A 类案例。
数据可信度 / 证据强度
强
A 类案例
6
AA 评分
26.32
官方来源
3 个补充入口
案例来源
6 条可核验
来自公开资料与模型数据库 + Artificial Analysis + 3 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。
Evidence Distribution
案例证据分布
Kimi K2 当前关联 6 条 A 类案例;这里按任务、来源和复核时间观察证据结构。
Evidence Snapshot
证据快照覆盖
Kimi K2 的 6 条 A 类案例中,已有 6 条完成本地证据快照;0 条仍在快照队列。
可用 manifest 复核
等待落盘
需补抓
需人工复核
证据完整度
5 / 5
Feishu Bitable
3 个入口
6 条
已绑定
已标注
能力边界
用标签表达信号,不用标签替代证据。
通用/待核验
暂无数据
待核验
平台/待核验
中
| 厂商 | Kimi / Moonshot AI |
|---|---|
| 厂商页 | Kimi / Moonshot AI |
| 发布时间 | 2025-07-11 |
| 模型 ID | kimi-k2 |
| 上下文 | 暂无数据 |
| 输出 | 暂无数据 |
| 模态 | 文本 / 多模态待核验 |
| 推理 | 否或官方未披露 |
| 价格 | 官方未披露 / 暂无数据 |
| 可用平台 | Kimi / Moonshot AI |
| AA 评分 | 26.32 |
| 基础模型发布时间 | Artificial Analysis 未披露 |
| 资料来源 | 公开资料、厂商信息和案例库 |
| A 类案例 | 6 |
谱系位置
它在厂商路线中的位置
此页将 Kimi K2 放入 Kimi / Moonshot AI 的当前 Atlas 路线中。若同厂前后代资料不足,先保留为路线追踪入口,不脑补谱系关系。
2025-07-11
Kimi K2 资料状态
来自公开资料与模型数据库 + Artificial Analysis + 3 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。
公开档案
模型状态
Kimi K2 当前标记为已有真实案例。
案例库
公开案例复核
已有 6 条可核验 A 类案例,达到完整补齐线
案例时间线
最近进入档案的 A 类案例
用采集时间呈现证据进入 Atlas 的顺序,方便复核来源新鲜度。
2026-06-27T06:01:09Z
sinameraji built KimiFlare, a terminal coding agent powered by Kimi K2 on Cloudflare Workers AI
sinameraji / developer_tools
2026-06-27T06:01:09Z
LLM-Red-Team built Kimi CC to run Claude Code on Kimi K2 0711 Preview
LLM-Red-Team / developer_tools
2026-06-27T05:52:39Z
Token DD Agent uses Kimi K2 as the default reasoning model for crypto token due diligence
untanao / crypto_due_diligence_agent
2026-06-27T05:52:39Z
Seismo-Lingo uses Kimi-K2 for natural-language geophysical data analysis
yberkayozkan / geophysical_data_analysis_assistant
2026-06-27T03:30:00Z
Workshop Labs 使用 LoRA 微调 Kimi K2 Thinking 的实践
Workshop Labs (workshop-labs-pbc) / model_finetuning
适合场景
- - 作为完整模型索引和厂商路线追踪入口
- - 通用模型对比
不适合
- - 缺少真实案例时不要包装成推荐
- - 缺字段时不要脑补价格、性能或上下文
真实案例
可核验 A 类案例精选
当前展示 6 条代表案例;排序优先看模型卡精选、证据可信度、快照状态和公开仓库/技术任务信号。案例总数仍为 6。
Representative Evidence
代表案例排序
当前第一代表案例是“LLM-Red-Team built Kimi CC to run Claude Code on Kimi K2 0711 Preview”。它会优先出现在本页,是因为该案例同时进入模型卡精选、具备 100/100 的证据可信度,并带有 已快照证据。
模型卡精选优先
A+ 完整链路
archive/evidence/kimi-k2-llm-red-team-kimi-cc-claude-code/manifest.json
Coding / repo / agent
LLM-Red-Team 使用 Kimi K2 处理软件工程任务执行
LLM-Red-Team · Kimi K2
A+ 完整链路 · 代码仓库证据
LLM-Red-Team 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:01:09Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Build a public command-line integration that lets developers drive Claude Code through Moonshot's Kimi service, explicitly using the kimi-k2-0711-preview model and the Kimi Open Platform API key flow.
公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:Public GitHub repository for Kimi CC with install script, multilingual README files, documented Kimi Open Platform API-key setup, and a workflow where running the claude command is backed by Kimi K2 0711 Preview for low…
模型作用:Kimi K2 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Kimi K2 0711 Preview is the model backend that supplies the coding-agent reasoning and generation used through Claude Code; the project wraps authentication and endpoint configuration so the Claude Code interface can de…
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is a public GitHub README and repository explicitly naming kimi-k2-0711-preview. It is a community developer tool rather than an official Moonshot customer story; no independent usage metrics are published.
原始记录:LLM-Red-Team built Kimi CC to run Claude Code on Kimi K2 0711 Preview
sinameraji 使用 Kimi K2 处理软件工程任务执行
sinameraji · Kimi K2
A+ 完整链路 · 代码仓库证据
sinameraji 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:01:09Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Create a terminal-native coding agent that uses the Kimi K2 family through Cloudflare Workers AI, with optional AI Gateway routing for observability, caching, per-turn cost logging, and account-owned request telemetry.
公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:Public GitHub repository for KimiFlare, described as a terminal coding agent powered by Kimi K2.7 on Cloudflare Workers AI, with documented interactive TUI, plan/edit/auto modes, image input support, shell and file tool…
模型作用:Kimi K2 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Kimi K2 provides the core coding-agent intelligence: reading and editing code, planning multi-step changes, using terminal/file/browser tools, handling screenshots or image prompts, and producing code or command decisio…
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is a public GitHub README and repository. The README targets the Kimi K2.7 variant exposed by Cloudflare Workers AI, so this is bound to the Kimi K2 family rather than only the initial K2 release. NPM package p…
原始记录:sinameraji built KimiFlare, a terminal coding agent powered by Kimi K2 on Cloudflare Workers AI
untanao 使用 Kimi K2 处理软件工程任务执行
untanao · Kimi K2
A+ 完整链路 · 代码仓库证据
untanao 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T05:52:39Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Open-source web agent where a user pastes an EVM token contract address and the system gathers Etherscan metadata, DexScreener market data, GoPlus security signals, honeypot simulation results, and web evidence to asses…
公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The application streams tool-call progress to the UI and produces a structured Markdown due-diligence report with a 0-100 risk score and verdict for the submitted token.
模型作用:Kimi K2 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README states Kimi K2 works out of the box and is the default reasoning model via Fireworks (`accounts/fireworks/models/kimi-k2p6`). Kimi K2 drives the agent loop, selects among available tools, interprets retrieved…
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is from the public GitHub README and code artifact; the project is model-agnostic and supports other LLMs, but its documented default model is Kimi K2 via Fireworks.
原始记录:Token DD Agent uses Kimi K2 as the default reasoning model for crypto token due diligence
yberkayozkan 使用 Kimi K2 处理软件工程任务执行
yberkayozkan · Kimi K2
A+ 完整链路 · 代码仓库证据
yberkayozkan 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T05:52:39Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Open-source LangChain/Streamlit assistant for geophysical workflows, including reading and analyzing SEGY seismic data, processing LAS well logs, answering natural-language questions, and generating seismic sections, we…
公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The repository documents an interactive chat interface that can answer seismic and well-data questions such as showing inline/crossline sections, generating amplitude histograms, listing wells, and plotting logs.
模型作用:Kimi K2 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README states the assistant is powered by the Kimi-K2 model through OpenRouter (`moonshotai/kimi-k2:free`). Kimi-K2 provides natural-language understanding, intelligent tool selection and parameter extraction, and c…
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is from the public GitHub README and code artifact; runtime requires an OpenRouter API key and user-provided geophysical files, but the model binding is explicitly documented as Kimi-K2.
原始记录:Seismo-Lingo uses Kimi-K2 for natural-language geophysical data analysis
Stray Labs (straylabs-ai) 使用 Kimi K2.5 处理代码审查和测试生成
Stray Labs (straylabs-ai) · Kimi K2.5
A+ 完整链路 · 代码仓库证据
Stray Labs (straylabs-ai) 公开的代码审查与测试案例,来源为 公开代码库,复核于 2026-06-27T03:30:00Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:使用 Kimi K2.5 模型构建自主 Web 应用渗透测试代理 Deadend CLI。该工具采用反馈驱动迭代架构(ADaPT),在标准工具失败时自动生成自定义 Python 漏洞利用代码,观察响应并迭代优化攻击策略。支持完全本地执行,无数据外泄。
公开产物:Deadend CLI 在 XBOW 104 题验证套件上使用 Kimi K2.5 达到约 80% 的成功率,总 API 成本约 $122。在 SQL 注入(83%)、XSS(91%)、业务逻辑(86%)等类别表现突出,GraphQL 和 SSRF 达到 100%。在盲注 SQL 注入等其他代理得分为 0% 的挑战上成功突破。
模型作用:Kimi K2.5 提供了强大的代码生成和推理能力,使代理能够自主生成自定义漏洞利用代码、分析 Web 应用响应并迭代优化攻击策略。模型的工具调用能力支持与 Playwright、Docker 等沙箱工具的集成。
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:使用 Kimi K2.5 变体(较新版本)。基准测试结果来自 2026 年 1 月的 XBOW 验证套件。
原始记录:Kimi K2.5 驱动的自主渗透测试代理
Workshop Labs (workshop-labs-pbc) 使用 Kimi K2 Thinking 处理真实任务执行
Workshop Labs (workshop-labs-pbc) · Kimi K2 Thinking
A+ 完整链路 · 代码仓库证据
Workshop Labs (workshop-labs-pbc) 公开的真实任务执行案例,来源为 公开代码库,复核于 2026-06-27T03:30:00Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。
任务:Workshop Labs 团队在 HuggingFace 上对 Kimi K2 Thinking 进行 LoRA 微调的完整实践。需要解决的核心挑战:Kimi K2 使用量化专家权重(NF4),HuggingFace 原生不支持训练量化专家。团队通过 monkey-patch 方式绕过限制,使用 8×H200 GPU 在自定义 Yoda 风格数据集上训练 40 步,验证损失下降且模型行为发生预期变化。
公开产物:成功在 HuggingFace 框架上对 Kimi K2 Thinking 进行 LoRA 微调。训练在 8×H200 GPU 上完成,40 步训练后损失下降。产出包含完整的训练代码(train.py)、数据生成脚本(make_yoda_dataset.py)、训练日志和损失曲线可视化。技术博客详细记录了遇到的 bug 和解决方案。
模型作用:Kimi K2 Thinking 的开放权重使社区能够进行微调实验。模型的 MoE 架构(384 专家,每 token 选 8 个 + 1 个共享专家)对微调提出了独特挑战,团队需要 monkey-patch HuggingFace 代码以支持量化专家训练。
A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。
风险边界:这是微调/训练场景而非推理使用。需要 8×H200 GPU。博客文章 URL 可能有访问限制。
原始记录:Workshop Labs 使用 LoRA 微调 Kimi K2 Thinking 的实践
数据缺口
只基于现有字段判断缺口;缺失项不会被猜测填充。
- - 补充方向:核验 价格 的官方/API/案例来源。
- - 补充方向:核验 上下文 的官方/API/案例来源。
- - 补充方向:核验 输出 的官方/API/案例来源。
下一步验证建议
- - 继续保留 A 类案例的原始证据、产物页和版本快照。
- - 补官方发布/API/System Card 来源;若官方未披露,继续标注“官方未披露”。
- - 把 benchmark、教程、发布文保留为背景资料,不提升为真实案例。