llama.cpp Releases更新日

[Pre-release / 継続的ビルド] b11448

[Pre-release / 継続的ビルド] cuda: BF16/FP16 conversion to f32 chunking (#29442) * ggml-cuda: chunk large BF16/FP16 to F32 conversions * Update ggml/src/ggml-cuda/ggml-cuda.cu Co-authored-by: Johannes Gäßler * Update…

Pre-release / 継続的ビルド: これは実験的なビルドであり、安定版リリースではありません。

cuda: BF16/FP16からf32への変換のチャンク化 (#29442) * ggml-cuda: 大きなBF16/FP16からF32への変換をチャンク化 * ggml/src/ggml-cuda/ggml-cuda.cu を更新 Co-authored-by: Johannes Gäßler * ggml/src/ggml-cuda/ggml-cuda.cu を更新 Co-authored-by: Johannes Gäßler * ggml-cuda: チャンク化したcuBLAS matmulでdstストライドを尊重 --------- Co-authored-by: Johannes Gäßler

ウェブサイト: - https://llama.app

Attestations: - https://github.com/ggml-org/llama.cpp/attestations/53292370

macOS/iOS用: - macOS Apple Silicon(arm64版) - macOS Apple Silicon(arm64、KleidiAI有効) 無効 - macOS Intel(x64版) - iOS向けXCFramework

Linux: - Ubuntu x64 (CPU) - Ubuntu arm64 (CPU) - Ubuntu s390x (CPU) - Ubuntu x64 (Vulkan) - Ubuntu arm64 (Vulkan) - Ubuntu x64 (CUDA 12) - CUDA 12.8 ライブラリ - Ubuntu x64 (CUDA 13) - CUDA 13.4 ライブラリ - Ubuntu arm64 (CUDA 13) - CUDA 13.4 ライブラリ - Ubuntu x64 (ROCm 10.0) - Ubuntu x64 (OpenVINO) - Ubuntu x64 (SYCL FP32) - Ubuntu x64 (SYCL FP16) - Linux arm64 (Snapdragon: CPU、Adreno GPU、Hexagon NPU) - セットアップガイド

Android: - Android arm64(CPU) - Android arm64(Snapdragon: CPU、Adreno GPU、Hexagon NPU) - セットアップガイド

Windows版: - Windows x64(CPU) - Windows arm64(CPU) - Windows arm64(OpenCL Adreno) - Windows x64(CUDA 12) - CUDA 12.4 DLLファイル - Windows x64(CUDA 13) - CUDA 13.4 DLLファイル - Windows arm64(CUDA 13) - CUDA 13.4 DLLファイル - Windows x64(Vulkan) - Windows arm64(Vulkan) - Windows x64(OpenVINO) - Windows x64(SYCL) - Windows x64(ROCm 10.0)

openEuler: - 無効 - openEuler x86(310p向け) - openEuler x86(910b、ACL Graph向け) - openEuler aarch64(310p向け) - openEuler aarch64(910b、ACL Graph向け)

UI: - UI

原文の出典

llama.cpp Releases

内容について

原文の公開と権利は出典元に帰属します。

機械翻訳 · 原文をご参照ください