DATA CUT 2026-07-27 116 活跃模型 680 A 类案例

模型档案

Claude 4.5 Sonnet

Claude 4.5 Sonnet 已有可核验 A 类案例,可优先进入选型对比。

已有真实案例

档案摘要 / Decision snapshot

先判断能不能用于选型。

证据完整度 100%
公开状态 已有真实案例

可进入对比

案例覆盖 6 条 A 类

有真实任务证据

主要适用 作为完整模型索引和厂商路线追踪入口

来自模型卡

谨慎场景 缺少真实案例时不要包装成推荐

边界优先

30 秒结论

可进入选型对比,但仍以案例证据为准。

来自公开模型资料库,并已补厂商/官方/模型家族证据;当前站内暂无可核验 A 类案例。

这页有来源、有风险说明、有公开可核验 A 类案例。

数据可信度 / 证据强度

已有真实案例

A 类案例

6

AA 评分

43.03

官方来源

2 个补充入口

案例来源

6 条可核验

来自公开资料与模型数据库 + Artificial Analysis + 2 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

Evidence Distribution

案例证据分布

Claude 4.5 Sonnet 当前关联 6 条 A 类案例;这里按任务、来源和复核时间观察证据结构。

A CASES 6 6 selected

Evidence Snapshot

证据快照覆盖

Claude 4.5 Sonnet 的 6 条 A 类案例中,已有 6 条完成本地证据快照;0 条仍在快照队列。

SNAPSHOT 100%
已快照 6

可用 manifest 复核

待快照 0

等待落盘

部分快照 0

需补抓

需关注 0

需人工复核

证据完整度

5 / 5

模型卡 OK

Feishu Bitable

官方/厂商来源 OK

2 个入口

A 类案例 OK

6 条

公开产物 OK

已绑定

风险备注 OK

已标注

能力边界

用标签表达信号,不用标签替代证据。

查看证据方法
Coding Agent

通用/待核验

长上下文

暂无数据

研究任务

待核验

开放生态

平台/待核验

生命周期风险

厂商Anthropic / Claude
厂商页Anthropic / Claude
发布时间2025-09-29
模型 IDclaude-4-5-sonnet
上下文暂无数据
输出暂无数据
模态文本 / 多模态待核验
推理是 / 推理模型或 thinking 模式
价格官方未披露 / 暂无数据
可用平台Anthropic / Claude
AA 评分43.03
基础模型发布时间Jul 1, 2025
资料来源公开资料、厂商信息和案例库
A 类案例6

谱系位置

它在厂商路线中的位置

此页将 Claude 4.5 Sonnet 放入 Anthropic / Claude 的当前 Atlas 路线中。若同厂前后代资料不足,先保留为路线追踪入口,不脑补谱系关系。

2025-09-29

Claude 4.5 Sonnet 资料状态

来自公开资料与模型数据库 + Artificial Analysis + 2 个厂商/官方/模型家族证据入口;AA 不是唯一事实来源。

公开档案

模型状态

Claude 4.5 Sonnet 当前标记为已有真实案例。

案例库

公开案例复核

已有 6 条可核验 A 类案例,达到完整补齐线

案例时间线

最近进入档案的 A 类案例

用采集时间呈现证据进入 Atlas 的顺序,方便复核来源新鲜度。

适合场景

  • - 作为完整模型索引和厂商路线追踪入口
  • - 推理/agentic workflow 候选

不适合

  • - 缺少真实案例时不要包装成推荐
  • - 缺字段时不要脑补价格、性能或上下文

真实案例

可核验 A 类案例精选

当前展示 6 条代表案例;排序优先看模型卡精选、证据可信度、快照状态和公开仓库/技术任务信号。案例总数仍为 6。

Representative Evidence

代表案例排序

当前第一代表案例是“Abe Diaz built an AI PM resume analyzer that can use Claude Sonnet 4.5 for six-pillar hiring evaluations”。它会优先出现在本页,是因为该案例同时进入模型卡精选、具备 100/100 的证据可信度,并带有 已快照证据。

打开第一代表案例
排序池 6 条

模型卡精选优先

证据可信度 100/100

A+ 完整链路

快照状态 已快照

archive/evidence/claude-4-5-sonnet-abe238-aipm-resume-analyzer/manifest.json

技术信号 命中

Coding / repo / agent

Abe Diaz / abe238 使用 Claude 4.5 Sonnet 处理软件工程任务执行

Abe Diaz / abe238 · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:github 最后复核:2026-06-27T10:13:27Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Abe Diaz / abe238 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T10:13:27Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Build an open-source Applied AI Product Manager resume analyzer for job seekers and hiring teams, using a six-pillar AI PM evaluation framework to score PDF, DOC, and DOCX resumes and produce structured hiring-readiness…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public project publishes a runnable analyzer and product site; its README states it generates per-candidate Markdown, HTML, and JSON reports with 0-10 scoring by pillar, evidence, level assessment, and model-selecta…

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Claude Sonnet 4.5 is explicitly listed as an Anthropic option and default Anthropic model for the analyzer; the CLI model table binds claude-sonnet-4-5-20250929 to complex resume-analysis work and the README documents r…

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The tool is multi-provider and defaults to OpenAI unless configured; accepted because the README and CLI code explicitly bind the Anthropic pathway to claude-sonnet-4-5-20250929 and provide a public artifact/product for…

原始记录:Abe Diaz built an AI PM resume analyzer that can use Claude Sonnet 4.5 for six-pillar hiring evaluations

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

abe238 / AI PM Resume Analyzer 使用 Claude 4.5 Sonnet 处理软件工程任务执行

abe238 / AI PM Resume Analyzer · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:GitHub 最后复核:2026-06-27T10:12:11Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

abe238 / AI PM Resume Analyzer 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T10:12:11Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Analyze Applied AI PM resumes from PDF/DOC/DOCX files against a 2026 six-pillar framework covering technical skills, product thinking, AI/ML knowledge, communication, strategic thinking, and execution; users can run the…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The tool generates structured resume evaluations with 0-10 pillar scores, evidence quotes, total score out of 60, decision labels such as Strong Screen, Markdown/HTML/JSON reports, and optional deep-analysis consensus o…

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Claude Sonnet 4.5 is implemented as the default Anthropic model in the analyzer and is called to read resume text, optionally evaluate visual design from a resume image, apply the six-pillar rubric, return JSON scoring,…

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Public open-source repository and README verify the task, artifact, and exact Anthropic model id; this is an independent tool rather than an Anthropic customer story, and users must supply their own Anthropic API key wi…

原始记录:AI PM Resume Analyzer uses Claude Sonnet 4.5 to score resumes against a six-pillar AI PM hiring framework

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Mohamed Sorour 使用 Claude 4.5 Sonnet 处理软件工程任务执行

Mohamed Sorour · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:github 最后复核:2026-06-27T08:54:40Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Mohamed Sorour 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T08:54:40Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Create a natural-language AWS operations assistant that monitors and optimizes cloud resources through Amazon Bedrock AgentCore Gateway, exposing CloudWatch, Logs, and EBS APIs as tools instead of requiring manual dashb…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public repository contains a runnable agent, screenshots, setup instructions, local CLI flow, semantic tool discovery, and documented access to 137 AWS tools for monitoring metrics, analyzing logs, and optimizing EB…

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README names Claude Sonnet 4.5 as the LLM responsible for natural-language understanding inside the AgentCore Gateway plus Strands Agents architecture.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The README describes the project as production-ready while AgentCore Runtime deployment is still marked in progress; accepted as a public implementation artifact with clear model binding and task evidence.

原始记录:Mohamed Sorour built an AWS Resource Optimizer Agent with Claude Sonnet 4.5 and Bedrock AgentCore

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

asimniazi63 使用 Claude 4.5 Sonnet 处理软件工程任务执行

asimniazi63 · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:github 最后复核:2026-06-27T08:54:40Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

asimniazi63 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T08:54:40Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Build an autonomous Enhanced Due Diligence research system that gathers intelligence on people and entities, analyzes risks, maps relationships, and generates comprehensive investigation outputs through a LangGraph-base…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The repository publishes the DeepAgents codebase and architecture for a production-ready due-diligence research agent, including orchestrator, Claude service, search service, result aggregation, error handling, and repo…

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README explicitly lists Claude Sonnet 4.5 in the multi-model architecture and assigns it to strategic analysis, query generation, and reflection steps in the due-diligence workflow.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The system also uses GPT-4o, so the case is multi-model; the Claude Sonnet 4.5 contribution is separately identified in the README and tied to specific workflow responsibilities.

原始记录:DeepAgents uses Claude Sonnet 4.5 for autonomous enhanced due-diligence investigations

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Mark2009 使用 Claude 4.5 Sonnet 处理软件工程任务执行

Mark2009 · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:github 最后复核:2026-06-27T08:54:40Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Mark2009 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T08:54:40Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Port the jsonrepair JavaScript library to C++ and add repair capabilities for malformed JSON documents, producing a usable C++ implementation with examples.

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The public repository contains the C++ jsonrepaircpp library, feature list, examples showing invalid JSON repaired into valid JSON, and attribution that the port was produced by Claude Sonnet 4.5 with additional repair …

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:The README states that the C++ port was created by Claude Sonnet 4.5, binding the model to the software translation and implementation work.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:Evidence is the project README and public repository; no independent production deployment is claimed, but the artifact is a reachable code library rather than a tutorial, benchmark, or collection page.

原始记录:Mark2009 used Claude Sonnet 4.5 to port jsonrepair from JavaScript to C++

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

Max Harar 使用 Claude 4.5 Sonnet 处理软件工程任务执行

Max Harar · Claude 4.5 Sonnet

A
厂商:Anthropic / Claude 模型:Claude 4.5 Sonnet 来源平台:github 最后复核:2026-06-27T08:54:40Z 证据快照:已快照 · 2 / 2 个证据目标 归档说明:已完成快照:2 / 2 个证据目标,可用本地 manifest 复核证据指纹。
证据可信度 100/100

A+ 完整链路 · 代码仓库证据

原始证据1 个公开产物复核通过代码仓库证据

Max Harar 公开的代码代理与软件工程案例,来源为 公开代码库,复核于 2026-06-27T08:54:40Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。

任务:这是一个围绕软件工程任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:Build an interactive airline customer-service agent against Sierra's open tau-bench airline domain, using the policy document as the system prompt, seeded airline data as state, and typed tools for tasks such as booking…

公开产物:公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The repository publishes a working single-page live demo at maxharar.com/sierra and documents five scenario chips plus a Pass^1 measurement for claude-sonnet-4-5 on the selected airline tasks.

模型作用:Claude 4.5 Sonnet 在该案例中承担软件工程任务执行相关的生成、分析、编排或实现角色。 原始资料写作:Claude Sonnet 4.5 is listed as the reasoning model in the stack and is described as driving the agent's policy-aware decisions and tool use over the airline-domain workflow.

A 类理由:有具体使用者、具体任务、公开原始证据和可访问产物,可公开核验。

风险边界:当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The domain is based on an open benchmark environment, but the submitted artifact is a concrete public demo/repository rather than a leaderboard-only entry; repo and live demo were reachable during collection.

原始记录:Max Harar built a live Sierra tau-bench airline customer-service agent with Claude Sonnet 4.5

已有真实案例 代码代理与软件工程公开代码库A 类可核验real_case auto_approved 进入模型卡精选

数据缺口

只基于现有字段判断缺口;缺失项不会被猜测填充。

价格:待补公开资料 上下文:待补公开资料 输出:待补公开资料 官方/家族来源:已有 发布时间:已有 可用平台:已有 A 类案例:已有
  • - 补充方向:核验 价格 的官方/API/案例来源。
  • - 补充方向:核验 上下文 的官方/API/案例来源。
  • - 补充方向:核验 输出 的官方/API/案例来源。

下一步验证建议

  • - 继续保留 A 类案例的原始证据、产物页和版本快照。
  • - 补官方发布/API/System Card 来源;若官方未披露,继续标注“官方未披露”。
  • - 把 benchmark、教程、发布文保留为背景资料,不提升为真实案例。