来源标注日期:2025-05-28(未提供具体时刻)。 deepseek-reasoner Model Upgraded to DeepSeek-R1-0528: Enhanced Reasoning Capabilities Significant benchmark improvements (Pass@1) AIME 2025: 70.0 → 87.5 (+17.5) GPQA: 71.5 → 81.0 (+9.5) LCB_v6: 63.5 → 73.3 (+9.8) Aider: 57.0 → 71.6
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-01T22:17:20.000Z。最后修改:2025-05-22T23:46:12.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-01T20:47:02.000Z。最后修改:2025-05-22T23:46:03.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-02T13:34:17.000Z。最后修改:2025-05-22T23:44:50.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-04T08:02:25.000Z。最后修改:2025-05-10T15:08:32.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-23T11:30:25.000Z。最后修改:2025-04-29T18:14:44.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-28T07:16:35.000Z。最后修改:2025-04-29T15:59:55.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-classification。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-28T07:15:58.000Z。最后修改:2025-04-29T15:47:24.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-classification。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-02T17:35:27.000Z。最后修改:2025-04-09T10:41:47.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-02T17:03:27.000Z。最后修改:2025-04-09T10:41:20.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-03T15:51:13.000Z。最后修改:2025-04-05T22:36:07.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-03T15:06:08.000Z。最后修改:2025-04-05T22:35:07.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。
模型动态Meta Llama 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2025-04-04T10:03:43.000Z。最后修改:2025-04-05T22:34:47.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。
来源标注日期:2025-03-24(未提供具体时刻)。 deepseek-chat Model Upgraded to DeepSeek-V3-0324: Enhanced Reasoning Capabilities Significant improvements in benchmark performance: MMLU-Pro: 75.9 → 81.2 (+5.3) GPQA: 59.1 → 68.4 (+9.3) AIME: 39.6 → 59.4 (+19.8) LiveCodeBench: 39.2
来源标注日期:2025-01-20(未提供具体时刻)。 deepseek-reasoner is our new model DeepSeek-R1. You can invoke DeepSeek-V3 by specifying model='deepseek-reasoner'. For details, please refer to: DeepSeek-R1 Release For guides, please refer to: Thinking Mode
来源标注日期:2024-12-26(未提供具体时刻)。 The deepseek-chat model has been upgraded to DeepSeek-V3. The API remains unchanged. You can invoke DeepSeek-V3 by specifying model='deepseek-chat'. For details, please refer to: introducing DeepSeek-V3
来源标注日期:2024-12-10(未提供具体时刻)。 The deepseek-chat model has been upgraded to DeepSeek-V2.5-1210, with improvements across various capabilities. Relevant benchmarking results include: Mathematical: Performance on the MATH-500 benchmark has improved from 74.8% to 82
来源标注日期:2024-09-05(未提供具体时刻)。 The DeepSeek V2 Chat and DeepSeek Coder V2 models have been merged and upgraded into the new model, DeepSeek V2.5. For backward compatibility, API users can access the new model through either deepseek-coder or deepseek-chat. The ne
来源标注日期:2024-08-02(未提供具体时刻)。 The DeepSeek API has innovatively adopted hard disk caching, reducing prices by another order of magnitude. For more details on the update, please refer to the documentation Context Caching is Available 2024/08/02.
来源标注日期:2024-07-25(未提供具体时刻)。 Update API /chat/completions JSON Mode Function Calling Chat Prefix Completion(Beta) 8K max_tokens(Beta) New API /completions FIM Completion(Beta) For more details, please check the documentation New API Features 2024/07/25
来源标注日期:2024-07-24(未提供具体时刻)。 The deepseek-coder model has been upgraded to DeepSeek-Coder-V2-0724.
来源标注日期:2024-06-28(未提供具体时刻)。 The deepseek-chat model has been upgraded to DeepSeek-V2-0628. Model's reasoning capabilities have improved, as shown in relevant benchmarks: Coding: HumanEval Pass@1 79.88% -> 84.76% Mathematics: MATH ACC@1 55.02% -> 71.02% Reasoni
来源标注日期:2024-06-14(未提供具体时刻)。 The deepseek-coder model has been upgraded to DeepSeek-Coder-V2-0614, significantly enhancing its coding capabilities. It has reached the level of GPT-4-Turbo-0409 in code generation, code understanding, code debugging, and code com
来源标注日期:2024-05-17(未提供具体时刻)。 The deepseek-chat model has been upgraded to DeepSeek-V2-0517. The model has seen a significant improvement in following instructions, with the IFEval Benchmark Prompt-Level accuracy jumping from 63.9% to 77.6%. Additionally, on API