DATA CUT 2026-07-27 116 活跃模型 680 A 类案例

模型档案

GPT-4o (Aug)

GPT-4o (Aug) 已有可核验 A 类案例,可优先进入选型对比。

已有真实案例

档案摘要 / Decision snapshot

先判断能不能用于选型。

证据完整度 100%
公开状态 已有真实案例

可进入对比

案例覆盖 5 条 A 类

有真实任务证据

主要适用 作为完整模型索引和厂商路线追踪入口

来自模型卡

谨慎场景 缺少真实案例时不要包装成推荐

边界优先

30 秒结论

可进入选型对比,但仍以案例证据为准。

来自公开模型资料库,并已补厂商/官方/模型家族证据;当前站内暂无可核验 A 类案例。

这页有来源、有风险说明、有公开可核验 A 类案例。

数据可信度 / 证据强度

已有真实案例

A 类案例

5

AA 评分

18.64

官方来源

2 个补充入口

案例来源

5 条可核验

来自公开资料与模型数据库 + Artificial Analysis + 2 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

Evidence Distribution

案例证据分布

GPT-4o (Aug) 当前关联 5 条 A 类案例;这里按任务、来源和复核时间观察证据结构。

A CASES 5 5 selected

Evidence Snapshot

证据快照覆盖

GPT-4o (Aug) 的 5 条 A 类案例中,已有 4 条完成本地证据快照;1 条仍在快照队列。

SNAPSHOT 80%
已快照 4

可用 manifest 复核

待快照 1

等待落盘

部分快照 0

需补抓

需关注 0

需人工复核

证据完整度

5 / 5

模型卡 OK

Feishu Bitable

官方/厂商来源 OK

2 个入口

A 类案例 OK

5 条

公开产物 OK

已绑定

风险备注 OK

已标注

能力边界

用标签表达信号,不用标签替代证据。

查看证据方法
Coding Agent

通用/待核验

长上下文

暂无数据

研究任务

待核验

开放生态

平台/待核验

生命周期风险

厂商OpenAI
厂商页OpenAI
发布时间2024-08-06
模型 IDgpt-4o-aug
上下文暂无数据
输出暂无数据
模态文本 / 多模态待核验
推理否或官方未披露
价格官方未披露 / 暂无数据
可用平台OpenAI
AA 评分18.64
基础模型发布时间Oct 1, 2023
资料来源公开资料、厂商信息和案例库
A 类案例5

谱系位置

它在厂商路线中的位置

此页将 GPT-4o (Aug) 放入 OpenAI 的当前 Atlas 路线中。若同厂前后代资料不足,先保留为路线追踪入口,不脑补谱系关系。

2024-08-06

GPT-4o (Aug) 资料状态

来自公开资料与模型数据库 + Artificial Analysis + 2 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

公开档案

模型状态

GPT-4o (Aug) 当前标记为已有真实案例。

案例库

公开案例复核

已有 5 条可核验 A 类案例,达到完整补齐线

案例时间线

最近进入档案的 A 类案例

用采集时间呈现证据进入 Atlas 的顺序,方便复核来源新鲜度。

适合场景

  • - 作为完整模型索引和厂商路线追踪入口
  • - 通用模型对比

不适合

  • - 缺少真实案例时不要包装成推荐
  • - 缺字段时不要脑补价格、性能或上下文

真实案例

可核验 A 类案例精选

当前展示 5 条代表案例;排序优先看模型卡精选、证据可信度、快照状态和公开仓库/技术任务信号。案例总数仍为 5。

Representative Evidence

代表案例排序

当前第一代表案例是“AI-Researcher includes GPT-4o 2024-08-06 as the completion model for autonomous scientific research workflows”。它会优先出现在本页,是因为该案例同时进入模型卡精选、具备 100/100 的证据可信度,并带有 已快照证据。

打开第一代表案例
排序池 5 条

模型卡精选优先

证据可信度 100/100

A+ 完整链路

快照状态 已快照

archive/evidence/gpt-4o-aug-ai-researcher-autonomous-scientific-discovery/manifest.json

技术信号 命中

Coding / repo / agent

HKUDS / AI-Researcher 使用 GPT-4o (Aug) 处理软件工程任务执行

HKUDS / AI-Researcher · GPT-4o (Aug)

A
厂商:OpenAI 模型:GPT-4o (Aug) 来源平台:github 最后复核:2026-06-27T06:53:58Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

HKUDS / AI-Researcher 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:53:58Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:AI-Researcher is an autonomous scientific discovery system that helps researchers move from reference papers or research prompts through idea generation, experiment planning/execution, and paper-writing workflows.

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The project page, documentation, and repository are public; the research_agent constant file sets COMPLETION_MODEL from the environment with gpt-4o-2024-08-06 as the fallback value used by the agent code path.

模型作用:GPT-4o (Aug) 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:GPT-4o 2024-08-06 acts as the completion model for the research agent's planning, idea generation, tool orchestration, and research-writing steps when the default or OpenAI configuration is used.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Project documentation also shows other configurable providers; treat this as an open-source agent integration/default in code, not proof that all public demos currently route to GPT-4o 2024-08-06.

原始记录:AI-Researcher includes GPT-4o 2024-08-06 as the completion model for autonomous scientific research workflows

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

git-bob / Robert Haase 使用 GPT-4o (Aug) 处理软件工程任务执行

git-bob / Robert Haase · GPT-4o (Aug)

A
厂商:OpenAI 模型:GPT-4o (Aug) 来源平台:github 最后复核:2026-06-27T06:53:58Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

git-bob / Robert Haase 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:53:58Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:git-bob is an AI agent that runs in GitHub CI/GitLab runners to understand repository issues and pull requests, comment on issues, review PRs, split issues, and generate code or text changes for software-development wor…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public repository documents the working tool, package, CI workflow usage, and online issue-discussion artifacts; its README lists openai:gpt-4o-2024-08-06 among tested/configurable LLMs and states that tested GPT-4o…

模型作用:GPT-4o (Aug) 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:GPT-4o 2024-08-06 supplies the language understanding and code/text generation used by git-bob to interpret GitHub issues or pull requests and produce implementation, review, or comment outputs inside CI.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Open-source tool evidence; per-run model choice is user-configured via GIT_BOB_LLM_NAME, so claim support/tested use rather than exclusive default production routing.

原始记录:git-bob uses GPT-4o 2024-08-06 to solve GitHub issues and pull requests

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Simon Schubert / Kai 9000 使用 GPT-4o (Aug) 处理软件工程任务执行

Simon Schubert / Kai 9000 · GPT-4o (Aug)

A
厂商:OpenAI 模型:GPT-4o (Aug) 来源平台:github 最后复核:2026-06-27T06:53:58Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Simon Schubert / Kai 9000 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:53:58Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Kai 9000 is an open-source AI assistant with persistent memory across Android, iOS, Windows, macOS, Linux, and Web; it supports chat, generated interactive screens, reminders, email/memory background checks, and tool/MC…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The repository and public web app are reachable; the model catalog explicitly maps gpt-4o-2024-08-06 to GPT-4o with a 128k context window and the README lists OpenAI as a supported service for the assistant product.

模型作用:GPT-4o (Aug) 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:GPT-4o 2024-08-06 is available as an OpenAI model option that can provide the assistant's conversation, reasoning, screen-generation, and memory/task-handling responses when selected by the user/provider configuration.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence confirms product support and exact model catalog entry; it does not prove every Kai deployment defaults to GPT-4o 2024-08-06.

原始记录:Kai 9000 exposes GPT-4o 2024-08-06 in a cross-platform persistent-memory AI assistant

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

uni-medical / MedSegAgent authors 使用 GPT-4o (Aug) 处理软件工程任务执行

uni-medical / MedSegAgent authors · GPT-4o (Aug)

A
厂商:OpenAI 模型:GPT-4o (Aug) 来源平台:github 最后复核:2026-06-27T06:53:58Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

uni-medical / MedSegAgent authors 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:53:58Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:MedSegAgent is a multi-agent system for instructive medical image segmentation: it parses free-form clinical segmentation requests, filters candidate datasets from modality to anatomy to label, and selects/runs speciali…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The repository contains the accepted JBHI project code, dataset metadata, evaluation script, and OAI_CONFIG_LIST example that explicitly configures model gpt-4o-2024-08-06 for the OpenAI tag; eval_example.sh also invoke…

模型作用:GPT-4o (Aug) 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:GPT-4o 2024-08-06 is used as the LLM component for natural-language request understanding and coarse-to-fine segmentation model selection before medical segmentation execution.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Healthcare research artifact rather than a deployed clinical product; do not imply clinical approval or live patient use.

原始记录:MedSegAgent uses GPT-4o 2024-08-06 for instructive medical image segmentation model selection

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Distyl 使用 GPT-4o (Aug) 处理软件工程任务执行

Distyl · GPT-4o (Aug)

A
厂商:OpenAI 模型:GPT-4o (Aug) 来源平台:official_web 最后复核:2026-06-27T06:31:07Z 证据快照:待快照 · 0 / 2 个证据目标 归档说明:已列入归档队列,等待抓取 2 个证据目标。
证据可信度 99/100

A+ 完整链路 · 官方/客户故事

原始证据1 个公开产物复核通过官方/客户故事

Distyl 公开的代码代理与软件工程案例,来源为 官方页面,复核于 2026-06-27T06:31:07Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Distyl applied a fine-tuned GPT-4o model to text-to-SQL workflows including query reformulation, intent classification, chain-of-thought/self-correction, and SQL generation for enterprise data tasks.

公开产物:公开材料提供原始证据链接和可访问产物,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:OpenAI reports Distyl's fine-tuned GPT-4o reached 71.83% execution accuracy and ranked first on the BIRD-SQL leaderboard at the time of the post.

模型作用:GPT-4o (Aug) 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Fine-tuned GPT-4o supplied the domain-specific natural-language-to-SQL reasoning and generation capability used to improve execution accuracy across SQL tasks.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The public evidence is an OpenAI partner success paragraph rather than a Distyl implementation deep dive. The page explicitly ties GPT-4o fine-tuning to selecting gpt-4o-2024-08-06 as the base model.

原始记录:Distyl uses fine-tuned GPT-4o for enterprise text-to-SQL generation

已有真实案例 代码代理与软件工程官方页面A 类可核验real_case auto_approved 进入模型卡精选

数据缺口

只基于现有字段判断缺口;缺失项不会被猜测填充。

价格:待补公开资料 上下文:待补公开资料 输出:待补公开资料 官方/家族来源:已有 发布时间:已有 可用平台:已有 A 类案例:已有
  • - 补充方向:核验 价格 的官方/API/案例来源。
  • - 补充方向:核验 上下文 的官方/API/案例来源。
  • - 补充方向:核验 输出 的官方/API/案例来源。

下一步验证建议

  • - 继续保留 A 类案例的原始证据、产物页和版本快照。
  • - 补官方发布/API/System Card 来源;若官方未披露,继续标注“官方未披露”。
  • - 把 benchmark、教程、发布文保留为背景资料,不提升为真实案例。