Rebased onto upstream triton main with the merged warp-pipeline PR #10840 (always-on sched.barrier after each stage's memory ops, SchedGroupMask::non_mem_non_sideeffect). Drops the tutorial's local fence_loads/keep_order commits (superseded by #10840); keeps t
前沿研究Microsoft Research官方来源 Cryptographic code supports vital protections in modern computing systems. Learn how a new method helps verify code as developers write it while preserving speed and adaptability as it gets implemented and evolves. The post Verifying Rust cryptography in SymCr
Status: Resolved All impacted services have now fully recovered. Affected components File uploads (Operational)
Status: Resolved All impacted services have now fully recovered. Affected components Sites (Operational)
Google and AIM launched ATL Saathi, a Gemini-powered AI tool empowering Indian educators in robotics labs.
模型动态Stability AI 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-05-16T23:40:30.000Z。最后修改:2026-07-13T05:03:04.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-to-audio。
模型动态Stability AI 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-05-17T01:05:35.000Z。最后修改:2026-07-13T05:02:49.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-to-audio。
模型动态Stability AI 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-05-17T01:44:44.000Z。最后修改:2026-07-13T05:02:33.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-to-audio。
Status: Resolved All impacted services have now fully recovered. Affected components Login (Operational) Conversations (Operational)
# vLLM v0.25.0 Release Notes ## Highlights This release features 558 commits from 232 contributors (64 new)! * **Model Runner V2 is now the default for all dense models** (#44443). Building on quantized-model support from the previous release, MRv2 is now the
Agent 与开发工具Transformers 发布官方来源 # Patch release v5.13.1 This patch is focused on enabling `transformers` for the latest release of vllm! - Be more defensive with remap_legacy_layer_types for custom models (#47245) from @hmellor - Fix custom code which doesn't know about the new linear layer
模型动态MiniMax 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-06-02T07:50:45.000Z。最后修改:2026-07-11T09:09:23.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:image-text-to-text。
This is a patch release on top of [v1.27.0](https://github.com/microsoft/onnxruntime/releases/tag/v1.27.0), containing targeted bug fixes, a CUDA QMoE decode-path optimization, and CI/build infrastructure fixes. ## Bug Fixes - [MLAS] Fixed an `igemm` regressio
# Highlights **GLM-5.2 NVFP4, tuned for production**: We took time this cycle to tune GLM-5.2 NVFP4 on Blackwell for optimized production serving. It now runs at **500+ tok/s/user on 8x B300, 450 on 4x GB300** (bs=1). Run GLM-5.2 with our [cookbook](https://do
## Table of Contents - [Dialect & Frontend](#dialect--frontend) - [Backend & Compiler](#backend--compiler) - [AMD/HIP Backend](#amdhip-backend) - [NVIDIA Backend](#nvidia-backend) - [Gluon & Layout Improvements](#gluon--layout-improvements) - [Kernels & Benchm
Out-of-tree LLIR scheduler + amdgcnas plugins. Removes the in-tree LLIR scheduler and amdgcnas post-assembly tool (now shipped as plugins in the gfx950-gluon-tutorials repo) and keeps only the plugin-enabling hooks: the LLVM_PASS_PLUGIN_KEEP_TARGET_MACHINE gat
前沿研究Microsoft Research官方来源 Aurora 1.5 adds 22 more variables, hourly temporal resolution, and probabilistic ensemble forecasting to the Aurora foundation model, making it more useful for real-world weather, climate, and energy applications. The post Aurora 1.5: Extending open foundation
# Introducing Dify Agent *(Experimental)*: A New Agent Experience in Dify > [!WARNING] > You should provide Dify Agent services ⚠️**only to trusted, non-malicious users**⚠️. As many of you know, the shell-based LLM agent paradigm has brought a major leap in ag
来源已更新,查看详情及原始发布记录。
漏洞与安全LangChain 安全公告社区 / 第三方 ## Summary Several LangChain components that resolve filesystem paths or expand search patterns do not consistently confine the *resolved* path to the intended root directory. Affected behaviors include: a file-search agent middleware that validates a starting
前沿研究Microsoft Research官方来源 Short chart specifications are easy to write, but often produce uninspiring results. Flint is an open-source visualization language that offers a middle path, letting AI agents create expressive charts from compact, human-editable specifications. The post Flin
LLIR scheduler + amdgcnas + RA-hints, rebased onto upstream main (63a5e1f). Adds MMRA-fence-aware barrier handling (keeps release/acquire fences glued to the s.barrier) and decouples the LLVM misched-disable from the env var (driven off whether the LLIR schedu
Agent 与开发工具AgentScope 发布官方来源 ## Highlight **Agent** - Support agent interruption with context handling #1995 **Longterm Memory Integration** - Integrate ReMe based longterm memory #1972 - Integrate agentic longterm memory #1927 **TTS** - Support OpenAI TTS API #1878 - Support DashScope Co
模型动态DeepSeek 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-06-27T02:27:36.000Z。最后修改:2026-07-04T03:15:12.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。
模型动态DeepSeek 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-06-27T03:02:56.000Z。最后修改:2026-07-04T03:14:46.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。 任务:text-generation。
Agent 与开发工具Transformers 发布官方来源 # Release v5.13.0 ## New Model additions ### KimiK 2.5, 2.6, and 2.7 This release includes the architecture for Kimi 2.5 which is used by 2.5-2.7: Kimi K2.5 is an open-source, native multimodal agentic model that advances practical capabilities in long-horizon
来源已更新,查看详情及原始发布记录。
模型动态字节 Seed 模型仓库(镜像)社区 / 第三方 模型仓库动态。仓库创建:2026-07-03T00:07:55.000Z。最后修改:2026-07-03T03:05:35.000Z。仓库修改不等同于模型正式发布。 数据经第三方 Hugging Face 镜像采集,需以原始模型页面复核。
漏洞与安全Transformers 安全公告社区 / 第三方 A critical remote code execution vulnerability exists in all versions of the HuggingFace transformers library prior to version 5.3.0. The vulnerability allows an attacker to craft a malicious `config.json` file containing the `_attn_implementation_internal` fi
### Added - 💭 **Streamed reasoning display.** Models that emit thinking or reasoning now show that content as it streams, and it renders correctly in the chat overview and in exported conversations. [Commit](https://github.com/open-webui/open-webui/commit/0b7