DATA CUT 2026-07-27 116 活跃模型 680 A 类案例

CASE EVIDENCE / A RECORD

the-crypt-keeper (Can AI Code proje… 使用 DeepSeek LLM 67B Base 处理真实任务执行

the-crypt-keeper (Can AI Code project) 公开的真实任务执行案例,来源为 公开代码库,复核于 2026-06-27T00:35:00Z。

原始记录:Can AI Code: DeepSeek LLM 67B Base Code Generation Evaluation

A

Chinese Brief

中文案例导读

the-crypt-keeper (Can AI Code project) 公开的真实任务执行案例,来源为 公开代码库,复核于 2026-06-27T00:35:00Z。 Model Atlas 将它标记为 A 类证据,因为它同时具备具体使用者、具体任务、公开原始证据和可访问产物。 Model Atlas 不把 benchmark、教程、发布说明或集合页包装成真实案例。

厂商DeepSeek
模型DeepSeek LLM 67B Base
任务类型真实任务执行
审核状态auto_approved

任务

真实任务背景

这是一个围绕真实任务执行的真实任务,公开材料可以回溯到具体使用者和具体产物。 原始资料写作:The Can AI Code project (600+ stars), a self-evaluating interview framework for AI coders, evaluated deepseek-ai/deepseek-llm-67b-base for code generation capabilities across multiple programming languages and difficult…

真实任务执行公开代码库A 类可核验real_case
公开产物

公开材料提供公开代码、README 或项目配置,可用于核验任务结果、项目形态和模型绑定关系。 原始资料写作:The evaluation was completed (issue #119 closed), producing code generation benchmark results for deepseek-llm-67b-base. The model was evaluated alongside other LLMs including DeepSeek Coder variants, CodeLlama, and Mis…

模型作用

DeepSeek LLM 67B Base 在该案例中承担真实任务执行相关的生成、分析、编排或实现角色。 原始资料写作:DeepSeek LLM 67B Base was the base model under evaluation for its code generation capabilities, tested through the project's automated self-evaluation interview system that scores models on coding tasks.

风险边界

当前判断基于公开材料;若产物下线、仓库变更或模型参与比例仅来自作者自述,需要在引用前重新复核。 原始资料写作:The evaluation results are stored in ndjson files in the results directory. The model was the subject of evaluation rather than being deployed in production.