DATA CUT 2026-07-27 116 活跃模型 680 A 类案例

模型档案

o1-preview

o1-preview 已有可核验 A 类案例,可优先进入选型对比。

已有真实案例

档案摘要 / Decision snapshot

先判断能不能用于选型。

证据完整度 100%
公开状态 已有真实案例

可进入对比

案例覆盖 7 条 A 类

有真实任务证据

主要适用 作为完整模型索引和厂商路线追踪入口

来自模型卡

谨慎场景 缺少真实案例时不要包装成推荐

边界优先

30 秒结论

可进入选型对比,但仍以案例证据为准。

来自公开模型资料库,并已补厂商/官方/模型家族证据;当前站内暂无可核验 A 类案例。

这页有来源、有风险说明、有公开可核验 A 类案例。

数据可信度 / 证据强度

已有真实案例

A 类案例

7

AA 评分

23.74

官方来源

3 个补充入口

案例来源

7 条可核验

来自公开资料与模型数据库 + Artificial Analysis + 3 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

Evidence Distribution

案例证据分布

o1-preview 当前关联 7 条 A 类案例;这里按任务、来源和复核时间观察证据结构。

A CASES 7 7 selected

Evidence Snapshot

证据快照覆盖

o1-preview 的 7 条 A 类案例中,已有 5 条完成本地证据快照;2 条仍在快照队列。

SNAPSHOT 71%
已快照 5

可用 manifest 复核

待快照 2

等待落盘

部分快照 0

需补抓

需关注 0

需人工复核

证据完整度

5 / 5

模型卡 OK

Feishu Bitable

官方/厂商来源 OK

3 个入口

A 类案例 OK

7 条

公开产物 OK

已绑定

风险备注 OK

已标注

能力边界

用标签表达信号,不用标签替代证据。

查看证据方法
Coding Agent

通用/待核验

长上下文

暂无数据

研究任务

待核验

开放生态

平台/待核验

生命周期风险

厂商OpenAI
厂商页OpenAI
发布时间2024-09-12
模型 IDo1-preview
上下文暂无数据
输出暂无数据
模态文本 / 多模态待核验
推理是 / 推理模型或 thinking 模式
价格官方未披露 / 暂无数据
可用平台OpenAI
AA 评分23.74
基础模型发布时间Oct 1, 2023
资料来源公开资料、厂商信息和案例库
A 类案例7

谱系位置

它在厂商路线中的位置

此页将 o1-preview 放入 OpenAI 的当前 Atlas 路线中。若同厂前后代资料不足,先保留为路线追踪入口,不脑补谱系关系。

2024-09-12

o1-preview 资料状态

来自公开资料与模型数据库 + Artificial Analysis + 3 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

公开档案

模型状态

o1-preview 当前标记为已有真实案例。

案例库

公开案例复核

已有 7 条可核验 A 类案例,达到完整补齐线

案例时间线

最近进入档案的 A 类案例

用采集时间呈现证据进入 Atlas 的顺序,方便复核来源新鲜度。

适合场景

  • - 作为完整模型索引和厂商路线追踪入口
  • - 推理/agentic workflow 候选

不适合

  • - 缺少真实案例时不要包装成推荐
  • - 缺字段时不要脑补价格、性能或上下文

真实案例

可核验 A 类案例精选

当前展示 7 条代表案例;排序优先看模型卡精选、证据可信度、快照状态和公开仓库/技术任务信号。案例总数仍为 7。

Representative Evidence

代表案例排序

当前第一代表案例是“kernel-memory-dump generated an Electron FFmpeg video cropper with a single ChatGPT o1-preview prompt”。它会优先出现在本页,是因为该案例同时进入模型卡精选、具备 100/100 的证据可信度,并带有 已快照证据。

打开第一代表案例
排序池 7 条

模型卡精选优先

证据可信度 100/100

A+ 完整链路

快照状态 已快照

archive/evidence/o1-preview-openai-kernel-memory-dump-ffmpeg-electron-video-cropper-202409/manifest.json

技术信号 命中

Coding / repo / agent

kernel-memory-dump 使用 o1-preview 处理软件工程任务执行

kernel-memory-dump · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:github 最后复核:2026-06-27T09:54:34Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

kernel-memory-dump 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T09:54:34Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:The GitHub repository documents a desktop video-cropping application built with Electron and FFmpeg, and states it was 100% generated with a single ChatGPT o1-preview prompt. The app lets users load an MP4 file, visuall…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public repository contains the generated application and README usage instructions, including features for video preview, drag-based crop selection, progress tracking, crop scaling to the original resolution, instal…

模型作用:o1-preview 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:o1-preview is explicitly credited in the project title as the model used to generate the complete Electron/FFmpeg video-cropper application from one prompt, covering both the application implementation and documented us…

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is from the creator's public GitHub repository rather than an independent customer story; the repository explicitly names ChatGPT o1-preview and exposes the artifact, but there is no separate production deploym…

原始记录:kernel-memory-dump generated an Electron FFmpeg video cropper with a single ChatGPT o1-preview prompt

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

ZackTheDoer 使用 o1-preview 处理可玩交互原型构建

ZackTheDoer · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:github 最后复核:2026-06-27T09:54:34Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

ZackTheDoer 公开的游戏与交互原型案例,来源为 公开代码库,复核于 2026-06-27T09:54:34Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕可玩交互原型构建的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:The repository records a prompt asking for a complete Space Invaders game with a neon aesthetic and heavy particle effects using Three.js, all in a single HTML file ready to copy and paste, and identifies the project as…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public artifact is a GitHub repository containing the generated single-file browser game and README instructions for opening index.html, moving the spaceship with arrow keys, shooting with spacebar, destroying alien…

模型作用:o1-preview 在该案例中承担可玩交互原型构建相关的生成、分析、编排或实现角色。 原始资料写作:o1-preview is explicitly named by the repository and README as the model used to produce the complete game from one prompt, including the embedded JavaScript/HTML implementation and gameplay instructions.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is a public creator repository and not an enterprise deployment; it is still a concrete, reproducible artifact with exact o1-preview attribution and a clear generated software task.

原始记录:ZackTheDoer used o1-preview to create a single-file Three.js Space Invaders game

已有真实案例 游戏与交互原型公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Integuru-AI 使用 o1-preview 处理软件工程任务执行

Integuru-AI · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:GitHub 最后复核:2026-06-27T09:41:39Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Integuru-AI 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T09:41:39Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:Integuru 是公开开源的集成生成 Agent:用户录制 HAR 后输入目标动作(例如下载 utility bills),系统分析并逆向平台内部 API 请求,然后生成可执行 Python 代码完成该动作。README 明确说明如果用户 OpenAI 账户可用,Integuru 会自动切换到 o1-preview 做 code generation。

公开产物:公开产物是 Integuru GitHub 仓库及其命令行工具;输出结果是命中目标平台内部端点的 runnable Python code,用来执行用户描述的集成动作。

模型作用:o1-preview 被用于代码生成阶段,把已分析出的请求图和用户目标动作转化为可运行的 API 调用代码,承担复杂推理与代码合成角色。

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:证据来自 README;o1-preview 的使用条件是用户账户中可用,图生成阶段推荐 gpt-4o,代码生成阶段才自动切换到 o1-preview。

原始记录:Integuru 自动切换到 o1-preview 生成调用内部 API 的集成代码

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Sakana AI 使用 o1-preview 处理研究分析和报告生成

Sakana AI · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:GitHub 最后复核:2026-06-27T09:41:39Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Sakana AI 公开的研究与报告生成案例,来源为 公开代码库,复核于 2026-06-27T09:41:39Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:AI Scientist-v2 是 Sakana AI 开源的自动科研系统,README 的示例运行命令明确将 --model_writeup 设置为 o1-preview-2024-09-12,用于在实验阶段后生成研究论文 writeup。

公开产物:产物是公开可运行的 AI Scientist-v2 仓库;运行流程会在 experiments 目录生成实验日志、tree visualization,并在 writeup 阶段产出 timestamp_ideaname.pdf 论文文件。

模型作用:o1-preview 绑定在 writeup 阶段,负责把自动实验结果组织成科研论文草稿/报告,是最终公开科研产物生成链路中的核心写作推理模型。

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:证据来自项目 README 的示例命令,明确到 dated model o1-preview-2024-09-12;未声称所有默认运行都使用该模型。

原始记录:Sakana AI 在 AI Scientist-v2 中使用 o1-preview 生成科研论文 writeup

已有真实案例 研究与报告生成公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Agent Laboratory project / Samuel S… 使用 o1-preview 处理软件工程任务执行

Agent Laboratory project / Samuel Schmidgall · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:github 最后复核:2026-06-27T06:28:47Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Agent Laboratory project / Samuel Schmidgall 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T06:28:47Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Agent Laboratory is an end-to-end autonomous research assistant that collects and analyzes papers, plans experiments, prepares data, runs code, and generates a comprehensive report.

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The artifact is a public research-agent codebase that orchestrates arXiv, Hugging Face, Python, and LaTeX tools to produce experiment outputs and reports.

模型作用:o1-preview 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README lists o1-preview as a supported OpenAI backend selectable with the `--llm-backend` flag, so o1-preview can drive the agents' literature analysis, planning, experimentation, and report-generation steps.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence shows concrete product integration and selectable model support; it does not provide a named third-party customer deployment.

原始记录:Agent Laboratory offers o1-preview-backed autonomous research workflows

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Harvey 使用 o1-preview 处理软件工程任务执行

Harvey · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:official_customer_story 最后复核:2026-06-27T09:42:50Z 证据快照:待快照 · 0 / 2 个证据目标 归档说明:已列入归档队列,等待抓取 2 个证据目标。
证据可信度 99/100

A+ 完整链路 · 官方/客户故事

原始证据1 个公开产物复核通过官方/客户故事

Harvey 公开的代码代理与软件工程案例,来源为 官方页面,复核于 2026-06-27T09:42:50Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Harvey runs its professional-services legal AI platform on Azure AI infrastructure, including Azure OpenAI o1-preview, to help lawyers summarize and compare documents, perform due diligence, reference case law, analyze …

公开产物:公开材料提供原始证据链接和可访问产物,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The Microsoft customer story says Harvey is deployed across hundreds of law firms and legal teams, used by tens of thousands of lawyers, and that one corporate lawyer reported saving 10 hours of work per week; a Europea…

模型作用:o1-preview 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The story states Harvey's engineers primarily use Azure OpenAI Service o1-preview and o1-mini for reasoning and problem-solving tasks with increased focus, augmented with proprietary data for specific legal tasks and cu…

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence explicitly names Azure OpenAI o1-preview, but Harvey also uses o1-mini and GPT-series models, so the public story does not isolate every product result to o1-preview alone.

原始记录:Harvey uses Azure OpenAI o1-preview in its legal AI platform for document analysis and due diligence

已有真实案例 代码代理与软件工程官方页面A 类可核验real_case auto_approved 进入模型卡精选

Cognition 使用 o1-preview 处理软件工程任务执行

Cognition · o1-preview

A
厂商:OpenAI 模型:o1-preview 来源平台:engineering_blog 最后复核:2026-06-27T09:42:50Z 证据快照:待快照 · 0 / 2 个证据目标 归档说明:已列入归档队列,等待抓取 2 个证据目标。
证据可信度 97/100

A 高可信 · 社区公开记录

原始证据1 个公开产物复核通过社区公开记录

Cognition 公开的代码代理与软件工程案例,来源为 博客记录,复核于 2026-06-27T09:42:50Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Cognition evaluated OpenAI o1-preview in Devin, its autonomous software engineering agent, including tasks that require installing libraries, browsing for context, editing and running code, diagnosing dependency errors,…

公开产物:公开材料提供原始证据链接和可访问产物,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:Cognition reported that replacing key Devin-Base subsystems with the o1 series produced significant gains on its internal cognition-golden evaluation suite; in a concrete sentiment-analysis task, Devin with o1-preview c…

模型作用:o1-preview 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Cognition attributes the improvement to o1-preview's ability to reflect, analyze, backtrack, diagnose root causes instead of symptoms, and research online like a human engineer when resolving complex upstream causes in …

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The blog is an evaluation and product-engineering writeup rather than a third-party customer deployment; it explicitly names o1-preview and includes concrete Devin task artifacts, but some aggregate benchmark gains are …

原始记录:Cognition tested o1-preview inside Devin for root-cause debugging and autonomous coding-agent tasks

已有真实案例 代码代理与软件工程博客记录A 类可核验real_case auto_approved 进入模型卡精选

数据缺口

只基于现有字段判断缺口;缺失项不会被猜测填充。

价格:待补公开资料 上下文:待补公开资料 输出:待补公开资料 官方/家族来源:已有 发布时间:已有 可用平台:已有 A 类案例:已有
  • - 补充方向:核验 价格 的官方/API/案例来源。
  • - 补充方向:核验 上下文 的官方/API/案例来源。
  • - 补充方向:核验 输出 的官方/API/案例来源。

下一步验证建议

  • - 继续保留 A 类案例的原始证据、产物页和版本快照。
  • - 补官方发布/API/System Card 来源;若官方未披露,继续标注“官方未披露”。
  • - 把 benchmark、教程、发布文保留为背景资料,不提升为真实案例。