接入 API · 个人 AI 解读连接自己的模型解读资讯,浏览新闻无需配置。

NVIDIA · 英伟达 最新消息

汇集 NVIDIA 显卡、GPU 算力、CUDA、AI 工作站与驱动更新。

164 条资讯按来源发布或更新时间排序
算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11049

test-llama-archs : generate dummy test vocab (#29084) Assisted-by: pi:llama.cpp/DeepSeek-V4-Flash-Vision-Exp **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b11049/llama-

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11048

metal : support qwen4exp hc ops (#29000) Add support for the new DSV4 HC op variants used by qwen4exp: - hc_pre with per-element sigmoid gate (gated variant) - hc_post with identity mixing (comb == nullptr) Assisted-by: pi:llama.cpp/Qwen3.8-27B **Website:** -

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11047

cuda : fix CUB argsort corruption caused by in-place keys (#28389) argsort_f32_i32_cuda_cub called the one-shot DeviceRadixSort::SortPairs API with d_keys_in == d_keys_out (temp_keys, temp_keys). CUB's internal double-buffer ping-pong requires distinct key buf

行业与政策IT之家 AI 与硬件媒体报道

英伟达黄仁勋反对放缓 AI 发展的呼声,称 2030 年不会是世界末日

IT之家 9 月 19 日消息,在本周放出的 CBS《周日早间》节目中,英伟达首席执行官黄仁勋表示“2030 年不会是世界末日”,并反对放缓 AI 发展的呼声。 在过去 1 周时间里,恐慌浪潮席卷 AI 行业,这股浪潮由 Anthropic CEO 达里奥 · 阿莫代伊于 9 月 12 日发表的文章《我们必须控制前沿》所引发,文章呼吁对前沿 AI 模型的开发进行协调管控、放缓节奏。 OpenAI CEO 萨姆 · 奥尔特曼对文章观点予以背书,埃隆 · 马斯克也表示认同,这在这些行业最激烈的竞争对手之间实属罕见的同声

故障与 BugIT之家 AI 与硬件媒体报道

英伟达 GPU 驱动 Bug 被修复:RPCS3 模拟游戏帧率提升 37%、显存占用降低 32%

IT之家 9 月 19 日消息,科技媒体 Wccftech 昨日(9 月 18 日)发布博文,报道称 RPCS3 开发团队修复 NVIDIA GPU 驱动程序的一个 Bug,模拟游戏测试发现帧率可以提高 37%、显存占用降低 32%。 IT之家援引博文介绍,RPCS3 开发团队指出在 NVIDIA GPU 驱动程序中,存在影响模拟器性能的问题,会限制部分游戏的帧率,并在多数测试中推高显存占用。 贡献者 Yahfz 找到修复方案,可以变通缓解该驱动 Bug 影响,从而提高模拟器性能。开发团队在《GT 赛车 5》开放道

显卡与工作站Tom's Hardware 显卡与工作站媒体报道

AMD shares first official benchmarks for EPYC 'Venice' CPUs, targets Nvidia — company claims 256-core chip is more than twice as fast as Nvidia Vera, 96-core model 20% faster per-core

AMD has released several benchmarks for its EPYC 'Venice' CPUs in a clear shot at Nvidia. <p>Jake Roach has been bending pins and busting solder joints since the mid-2000s. From trying to run scratched CDs of <em>Delta Force </em>and <em>Unreal Tournament </em

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11046

opencl: add support for bin kernel `flash_attn_f32_f16_bin` (#29046) * opencl: add `flash_attn_f32_f16_bin` * opencl: guarded prefill fa **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/rel

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11045

hexagon: add ROLL op support (#29105) **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b11045/llama-b11045-bin-macos-arm64.tar.gz) - macOS Apple Silicon (arm64, KleidiAI e

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11044

hexagon: im2col update (#29103) * ggml-hexagon: accept 1D and padded IM2COL ops * ggml-hexagon: make pure-DDR IM2COL kernel is_2D-aware * ggml-hexagon: extend IM2COL DMA patch-embed fast path to 1D * ggml-hexagon: add blocked-staging general IM2COL DMA kernel

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11043

hexagon: HMX flash-attention head_dim padding (support DK=DV=72) (#26539) Allow HMX flash-attention to run with head_dim not a multiple of 64 (e.g. SigLIP head_dim=72), by operating on DK/DV rounded up to 64 with zero-filled tail lanes. **Website:** - **Attest

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11042

opencl: add bin kernel `kernel_gemm_noshuffle_q6_k_f32_32b_trans_ila_a8_bin` (#28678) * opencl: add A8 Q6_K non-MoE binary kernel * opencl: fix layout compatibility **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://githu

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11040

ggml : check for allocation failures to prevent crashes (#28149) * ggml : check for allocation failures to prevent crashes * wording **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/release

显卡与工作站Tom's Hardware 显卡与工作站媒体报道

Modder gets Nvidia's DLSS 5 working in a web browser using WebGPU — 147MB browser port runs on non-Nvidia GPUs and macOS but takes two seconds per render

A developer's live WebGPU demo runs Nvidia's DLSS 5 neural rendering in a web browser, but also apparently runs on macOS. <p>Shane has a background in computer engineering and has worked as a freelance consultant in multiple industries. He has a strong affecti

显卡与工作站Tom's Hardware 显卡与工作站媒体报道

Save $300 on this 4K gaming PC with a 9800X3D and RTX 5070 Ti, now $2,599 — powerhouse ABS Stratos II rig ships with 32GB DDR5 and a 2TB SSD

A 4K-capable gaming machine from ABS, featuring the powerful RTX 5070 Ti, AMD Ryzen 7 9800X3D, 32GB DDR5, and a 2TB SSD, all for $2,599.99. <p>Ben Stockton is a deals writer at Tom’s Hardware. Previously a hardware writer at PCGamesN, Ben’s been writing about

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11037

ggml-webgpu: fix supports_op condition for GET_ROWS (#28978) * fix get_rows vec4 handling * Add src strides checking to vec4_aligned of get_rows and the new test case. **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://gi

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11036

ggml : handle graph buffer reservation failure (#26070) **Website:** - **Attestations:** - **macOS/iOS:** - [macOS Apple Silicon (arm64)](https://github.com/ggml-org/llama.cpp/releases/download/b11036/llama-b11036-bin-macos-arm64.tar.gz) - macOS Apple Silicon

算力与芯片llama.cpp 发布官方来源

llama.cpp 发布: b11035

vulkan: add IQ3_S MMQ matmul kernels (#28822) * vulkan: add IQ3_S MMQ matmul kernels * Make block_a_to_shmem do 2-byte loads (110 bytes is divisible by 2) * Align the check, IQ3_S is also using K tile size **Website:** - **Attestations:** - **macOS/iOS:** - [m

算力与芯片Tom's Hardware 显卡与工作站媒体报道

Huawei details AI accelerator roadmap, pulls in next-generation Ascend NPUs by several quarters — FP4 performance of the Ascend 960PR doubles expectations

Huawei's mimics Nvidia's approach to AI factories, unveils details about next-generation Ascend NPUs, Kunpeng CPUs, scale-up and scale-out connectivity solutions. <p>Anton Shilov has been in the PC industry since 1990s playing games, building PCs, and writing

显卡与工作站Tom's Hardware 显卡与工作站媒体报道

Investigative report details how export-restricted Nvidia AI chips reach China — public records reveal how Chinese entities skirt US sanctions

American nonprofit C4ADS, a monitoring organization funded mostly by the U.S. government, produced a report shedding light on the many ways that American AI accelerators reach China. <p>Bruno Ferreira's journey kicked off with the venerable ZX Spectrum, a cass

算力与芯片Tom's Hardware 显卡与工作站媒体报道

Apple eyes Nvidia NVLink to power its new custom M8 Ultra AI servers — historically bitter rivals reportedly team up for 2029 data center push

Apple is reportedly interested in using Nvidia's NVLink Fusion for its own data center platforms. <p>Anton Shilov has been in the PC industry since 1990s playing games, building PCs, and writing stories about pretty much everything that relates to PCs, Macs, s

把 AI 雷达放到桌面

在支持安装的浏览器中,可以将本站作为应用打开。

安装入口取决于浏览器;应用和网站使用同一份最新内容。

查看完整安装指南