接入 API · 个人 AI 解读连接自己的模型解读资讯,浏览新闻无需配置。

本地 AI 部署

关注本地模型部署、推理引擎、量化和工作站运行工具。

120 条资讯按来源发布或更新时间排序
Agent 与开发工具Claude Code 发布官方来源

Claude Code 发布: v2.1.269

## What's changed - Added `claude plugin eval`: run a plugin's eval suite against Claude Code and get scored, reproducible results (JSON + HTML report); see `claude plugin eval --help` - Added `/output-style [name]` to list and switch output styles, including

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.29.0

# v0.29.0 ## Highlights This release features 594 commits from 277 contributors (91 new)! * **Model Runner V2 is now the default for all models** (#53183), completing the rollout that began with pooling models (#48290). MRV2 also gained CUDA graph memory profi

算力与芯片ONNX Runtime 发布官方来源

ONNX Runtime 发布: ONNX Runtime v1.30.0

ONNX Runtime 1.30.0 expands generative AI inference, improves CPU and GPU performance, adds Go bindings, and strengthens runtime reliability. These notes cover changes since ONNX Runtime 1.29.1. ## Highlights - Expanded CUDA inference support with variable-len

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.34.0

## Use Ollama models in ChatGPT Desktop Ollama models can now be used directly in ChatGPT Desktop, so you can keep your existing workflow while running open models. Setup is available from the Ollama app on MacOS. This release also improves structured output p

算力与芯片ONNX Runtime 发布官方来源

ONNX Runtime 发布: ONNX Runtime v1.29.1

This is a patch release on top of [v1.29.0](https://github.com/microsoft/onnxruntime/releases/tag/v1.29.0), containing GroupQueryAttention capability and KV-cache layout improvements, plugin Execution Provider performance tooling updates, and targeted graph an

模型动态DeepSeek 官方更新日志官方来源

DeepSeek-V4.1-Flash Release​

来源标注日期:2026-09-10(未提供具体时刻)。 Today, we officially release the DeepSeek-V4.1-Flash model. It is the smallest model in our new architecture family, with native multimodal visual understanding. The new architecture is designed for a higher capability ceiling, fast

算力与芯片SGLang 发布官方来源

SGLang 发布: v0.5.19

# Highlights *786 PRs from 214 contributors.* **New models in this release** (see the [cookbook](https://docs.sglang.io/cookbook) for all supported models): | Model | Type | PRs | Cookbook | |---|---|---|---| | Qwen3.8 (2.4T-A95B) | Autoregressive | [#35758](h

Agent 与开发工具CrewAI 发布官方来源

CrewAI 发布: 1.15.19

## What's Changed ### Features - Add Clipper integrations client - Add `now()` to the CEL expression environment - Record how a crew run ended for every user - Report machine size as a coarse band, not a core count - Add injectable client for CrewAI platform t

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.33.3

## What's Changed * gemma4 now supports images and audio on MLX engine * Report cached prompt tokens * Honor GGUF model defined default parameters * MLX, MLX-C, llama.cpp update ## New Contributors * @marcelpetrick made their first contribution in https://gith

算力与芯片ONNX Runtime 发布官方来源

ONNX Runtime 发布: ONNX Runtime v1.28.2

This is a patch release on top of [v1.28.1](https://github.com/microsoft/onnxruntime/releases/tag/v1.28.1), containing a targeted fix for Compile API model serialization. ## Highlights ### Bug Fixes - Fixed Compile API callback serialization to prevent duplica

AI 应用Open WebUI 发布官方来源

Open WebUI 发布: v0.11.3

### Added - ♿ **Accessibility mode reaches the menus.** Accessibility mode now marks the menu entry you are pointing at and the model already chosen with a stronger background, across the dropdown menus, their submenus, and the model picker together with its f

AI 应用Open WebUI 发布官方来源

Open WebUI 发布: v0.11.2

### Added - 🖼️ **Richer previews for terminal files.** Word documents and slide decks produced in the terminal are now previewed as the finished document rather than an approximation, and every document preview gains a page strip down the side with numbered t

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.33.2

## What's Changed * Ollama's app now follows the system appearance again, restoring dark mode support * Fixed the macOS app to properly hand off to an already-running instance instead of starting a second one * The Claude Desktop proxy no longer interrupts in-

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.33.1

## What's Changed * MLX: Qwen3.8 Flash Next support * cmake: make external compat patches idempotent * MLX and llama.cpp update * mlxrunner: add structured output support * mlxrunner: avoid Metal GPU timeouts when loading models from slow storage ## New Contri

Agent 与开发工具Transformers 发布官方来源

Transformers 发布: Release: v5.16.0

# Release v5.16.0 ## New Model additions ### Qwen4-Exp Qwen4-Exp builds on Qwen3.5's hybrid text and multimodal architecture with three key components: GatedResidual (GR), Qwen Sparse Attention (QSA), and Per-Layer Embedding (PLE). GR is a Qwen-developed resid

算力与芯片Intel OpenVINO 发布官方来源

Intel OpenVINO 发布: 2026.3.1

### Summary of major features and improvements * Functionally enabled Muse-Glimmer-30B model: * Model card and IR available: [OpenVINO/Muse-Glimmer-30B-int4-ov · Hugging Face](https://huggingface.co/OpenVINO/Muse-Glimmer-30B-int4-ov) * Notebooks to try: * [Mus

算力与芯片vLLM 发布官方来源

vLLM 发布: v0.28.0

# v0.28.0 ## Highlights This release features 584 commits from 270 contributors (76 new)! * **Kimi-K3 performance push**: a major optimization effort for Kimi-K3 across the stack — Decode Context Parallel (DCP) support (#50484), fused FlashKDA decode and prefi

Agent 与开发工具Ollama 发布官方来源

Ollama 发布: v0.33.0

## What's Changed ### Claude Desktop Developers can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider. ### Improved caching * Fixed a hang where agent clients that cancel long prefills * Prefill restore points

把 AI 雷达放到桌面

在支持安装的浏览器中,可以将本站作为应用打开。

安装入口取决于浏览器;应用和网站使用同一份最新内容。

查看完整安装指南