接入 API · 个人 AI 解读连接自己的模型解读资讯,浏览新闻无需配置。

DeepSeek · 深度求索 最新消息

追踪 DeepSeek 模型发布、推理能力、开放 API 与官方更新。

55 条资讯按来源发布或更新时间排序
算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11049

test-llama-archs : generate dummy test vocab (#29084) Assisted-by: pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b11049/llama-

算力与芯片Intel OpenVINO 发布官方来源

Intel OpenVINO 发布: 2026.4.0

### Summary of major features and improvements * #### More GenAI coverage and framework integrations to minimize code changes * New models supported: * On CPU: Gemma-3n * On CPU, GPU: Kokoro-82M, Qwen3-VL-4B with EAGLE-3, Qwen3-ASR, Muse Glimmer 30B, Qwen3.8 2

Agent 与开发工具LiteLLM 发布官方来源

LiteLLM 发布: v1.102.0-rc.1

## Verify Docker Image Signature All LiteLLM Docker images are signed with [cosign](https://docs.sigstore.dev/cosign/overview/). Every release is signed with the same key introduced in [commit `0112e53`](https://github.com/BerriAI/litellm/commit/0112e53046018d

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.29.0

# v0.29.0 ## Highlights This release features 594 commits from 277 contributors (91 new)! * **Model Runner V2 is now the default for all models** (#53183), completing the rollout that began with pooling models (#48290). MRV2 also gained CUDA graph memory profi

算力与芯片ONNX Runtime 发布官方来源

ONNX Runtime 发布: ONNX Runtime v1.30.0

ONNX Runtime 1.30.0 expands generative AI inference, improves CPU and GPU performance, adds Go bindings, and strengthens runtime reliability. These notes cover changes since ONNX Runtime 1.29.1. ## Highlights - Expanded CUDA inference support with variable-len

Agent 与开发工具Transformers 发布官方来源

Transformers 发布: Release 5.17.0

# Release v5.17.0 ## New Model additions ### HYV4 Hy4-Preview is a 780B-parameter mixture-of-experts language model that activates 49B parameters per token. Each MoE layer holds 256 routed experts plus one always-active shared expert and routes every token to

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4.1-Flash · 仓库更新

模型仓库动态。仓库创建:2026-09-10T02:17:58.000Z。最后修改:2026-09-10T08:18:10.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。

模型动态DeepSeek 官方更新日志官方来源

DeepSeek-V4.1-Flash Release​

来源标注日期:2026-09-10(未提供具体时刻)。 Today, we officially release the DeepSeek-V4.1-Flash model. It is the smallest model in our new architecture family, with native multimodal visual understanding. The new architecture is designed for a higher capability ceiling, fast

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4-Flash-Vision-Exp · 仓库更新

模型仓库动态。仓库创建:2026-08-31T06:16:18.000Z。最后修改:2026-09-01T09:22:10.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.28.0

# v0.28.0 ## Highlights This release features 584 commits from 270 contributors (76 new)! * **Kimi-K3 performance push**: a major optimization effort for Kimi-K3 across the stack — Decode Context Parallel (DCP) support (#50484), fused FlashKDA decode and prefi

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.33.0

## What's Changed ### Claude Desktop Developers can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider. ### Improved caching * Fixed a hang where agent clients that cancel long prefills * Prefill restore points

算力与芯片SGLang 发布官方来源

SGLang 发布: v0.5.18

# Highlights *710 PRs from 212 contributors.* **New models in this release** (see the [cookbook](https://docs.sglang.io/cookbook) for all supported models): | Model | Type | PRs | Cookbook | |---|---|---|---| | Muse Glimmer | Autoregressive (Multimodal) | [#34

模型动态DeepSeek 官方更新日志官方来源

DeepSeek-V4-Flash-Vision-Exp Release​

来源标注日期:2026-08-21(未提供具体时刻)。 Today, the new multimodal vision understanding model DeepSeek-V4-Flash-Vision-Exp is now available on the DeepSeek API platform. This is an experimental model that can be accessed by setting model='deepseek-v4-flash-vision-exp'. Ter

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4-Pro-0813 · 仓库更新

模型仓库动态。仓库创建:2026-08-13T03:05:06.000Z。最后修改:2026-08-13T16:28:28.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。

模型动态DeepSeek 官方更新日志官方来源

DeepSeek-V4-Pro Update​

来源标注日期:2026-08-13(未提供具体时刻)。 The GA release of DeepSeek-V4-Pro has been rolled out on the APP, Web, and API. The API calling method remains unchanged — simply set the model name to deepseek-v4-pro to use the latest version. Significantly enhanced Agent capabili

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.27.0

# vLLM v0.27.0 Release Notes ## Highlights This release features 561 commits from 242 contributors (64 new)! * **Kimi K3 support** with a full stack landing in one release: core model files and kernels (#50089, #50000), Python (#50093) and Rust (#50104) fronte

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4-Flash-0731 · 仓库更新

模型仓库动态。仓库创建:2026-07-31T07:30:24.000Z。最后修改:2026-08-01T03:07:41.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。

模型动态DeepSeek 官方更新日志官方来源

DeepSeek-V4-Flash Update​

来源标注日期:2026-07-31(未提供具体时刻)。 The official release of the DeepSeek-V4-Flash API is now in public beta. The API calling method remains unchanged — simply set the model name to deepseek-v4-flash to use the latest version. Significantly enhanced agent capabilities,

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.26.0

# vLLM v0.26.0 Release Notes ## Highlights This release features 411 commits from 212 contributors (61 new)! * **New Inkling model family** with a full support stack: base modeling (#48799), piecewise CUDA graph support (#48822), Hopper FA4 relative attention

算力与芯片SGLang 发布官方来源

SGLang 发布: v0.5.16

# Highlights *574 PRs from 169 contributors.* **DSpark: confidence-driven speculative decoding**: A new speculative algorithm. It drafts semi-autoregressively in blocks, then sizes each verify window from the draft's own confidence instead of a fixed draft len

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.25.0

# vLLM v0.25.0 Release Notes ## Highlights This release features 558 commits from 232 contributors (64 new)! * **Model Runner V2 is now the default for all dense models** (#44443). Building on quantized-model support from the previous release, MRv2 is now the

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4-Flash-DSpark · 仓库更新

模型仓库动态。仓库创建:2026-06-27T02:27:36.000Z。最后修改:2026-07-04T03:15:12.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/DeepSeek-V4-Pro-DSpark · 仓库更新

模型仓库动态。仓库创建:2026-06-27T03:02:56.000Z。最后修改:2026-07-04T03:14:46.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.24.0

# vLLM v0.24.0 Release Notes ## Highlights This release features 571 commits from 256 contributors (77 new)! * **MiniMax-M3**: Added support for the new **MiniMax-M3** model (#45381), with a fast follow-on of BF16/FP8 indexer via MSA (#45892), MXFP4 support (#

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/eagle3_gemma4_12b_ttt7 · 仓库更新

模型仓库动态。仓库创建:2026-06-28T12:37:17.000Z。最后修改:2026-06-28T12:38:12.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/eagle3_qwen3_14b_ttt7 · 仓库更新

模型仓库动态。仓库创建:2026-06-28T12:36:20.000Z。最后修改:2026-06-28T12:37:15.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/eagle3_qwen3_8b_ttt7 · 仓库更新

模型仓库动态。仓库创建:2026-06-28T12:35:29.000Z。最后修改:2026-06-28T12:36:18.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。

模型动态DeepSeek 模型仓库(镜像)社区 / 第三方

deepseek-ai/eagle3_qwen3_4b_ttt7 · 仓库更新

模型仓库动态。仓库创建:2026-06-28T12:34:59.000Z。最后修改:2026-06-28T12:35:27.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。

把 AI 雷达放到桌面

在支持安装的浏览器中,可以将本站作为应用打开。

安装入口取决于浏览器;应用和网站使用同一份最新内容。

查看完整安装指南