所屬事件:
llama.cpp 為 Clef 模型加入文字與視覺輸入支援(1 個來源 · 2 則報導)
2026 年 10 月 3 日,llama.cpp 發布建置 b11371,合併 PR #29831,為 clef 決策模型加入初始支援,目前僅限純文字輸入;該改動同時更新 gguf-py 常量並清理靜態圖相關程式碼。隨附建置產物覆蓋多個平台,包括 macOS 的 Apple Silicon(arm64)與 Intel… 看完整事件
llama.cpp b11418:server 支援 Clef 視覺輸入
llama.cpp 發布建置 b11418,其 server 新增對 Clef 的視覺輸入支援(PR #29969),並修復影像 token 上限、abort 與 yield_to_queue 資料變更等問題。
官方 server 為 Clef 加上視覺輸入支援並修正影像 token 上限等問題,同時列出各平台預編譯產物與本次停用的後端,便於在本地跑多模態模型的人判斷是否升級。
原標題:b11418
閱讀原文
同一事件共有 2 則報導(1 個來源),看事件全貌
| 評分 | 65 / 67(平均 66,門檻 60) |
| 狀態 | 精選 |
|---|
全文翻譯
server:為 Clef 支援視覺輸入(#29969)
server:為 Clef 支援視覺輸入
將 input_attn_causal 移至私有
擴充套件舊的 server_batch::embd
server_batch::token::pos 改為多維
小修
修復中止
修復影像 token 上限
修復 yield_to_queue 修改資料
網站:
https://llama.app
證明:
https://github.com/ggml-org/llama.cpp/attestations/52850734
macOS/iOS:
- macOS Apple Silicon(arm64)
- macOS Apple Silicon(arm64,啟用 KleidiAI)已停用
- macOS Intel(x64)
- iOS XCFramework
Linux:
- Ubuntu x64(CPU)
- Ubuntu arm64(CPU)
- Ubuntu s390x(CPU)
- Ubuntu x64(Vulkan)
- Ubuntu arm64(Vulkan)
- Ubuntu x64(CUDA 12)- CUDA 12.8 庫
- Ubuntu x64(CUDA 13)- CUDA 13.4 庫
- Ubuntu arm64(CUDA 13)- CUDA 13.4 庫
- Ubuntu x64(ROCm 10.0)
- Ubuntu x64(OpenVINO)
- Ubuntu x64(SYCL FP32)
- Ubuntu x64(SYCL FP16)
- Linux arm64(Snapdragon:CPU、Adreno GPU、Hexagon NPU)- 安裝指南
Android:
- Android arm64(CPU)
- Android arm64(Snapdragon:CPU、Adreno GPU、Hexagon NPU)- 安裝指南
Windows:
- Windows x64(CPU)
- Windows arm64(CPU)
- Windows arm64(OpenCL Adreno)
- Windows x64(CUDA 12)- CUDA 12.4 DLL
- Windows x64(CUDA 13)- CUDA 13.4 DLL
- Windows arm64(CUDA 13)- CUDA 13.4 DLL
- Windows x64(Vulkan)
- Windows arm64(Vulkan)
- Windows x64(OpenVINO)
- Windows x64(SYCL)
- Windows x64(ROCm 10.0)
openEuler:
- 已停用
- openEuler x86(310p)
- openEuler x86(910b,ACL Graph)
- openEuler aarch64(310p)
- openEuler aarch64(910b,ACL Graph)
UI:
- UI
由 AI 翻譯,以原文為準。
原文
server: support vision input for Clef ( #29969 )
server: support vision input for Clef
move input_attn_causal to private
extend old server_batch::embd
server_batch::token::pos to multi dim
nits
fix abort
fix img tokens cap
fix yield_to_queue mutate data
Website:
https://llama.app
Attestations:
https://github.com/ggml-org/llama.cpp/attestations/52850734
macOS/iOS:
macOS Apple Silicon (arm64)
macOS Apple Silicon (arm64, KleidiAI enabled) DISABLED
macOS Intel (x64)
iOS XCFramework
Linux:
Ubuntu x64 (CPU)
Ubuntu arm64 (CPU)
Ubuntu s390x (CPU)
Ubuntu x64 (Vulkan)
Ubuntu arm64 (Vulkan)
Ubuntu x64 (CUDA 12) - CUDA 12.8 libraries
Ubuntu x64 (CUDA 13) - CUDA 13.4 libraries
Ubuntu arm64 (CUDA 13) - CUDA 13.4 libraries
Ubuntu x64 (ROCm 10.0)
Ubuntu x64 (OpenVINO)
Ubuntu x64 (SYCL FP32)
Ubuntu x64 (SYCL FP16)
Linux arm64 (Snapdragon: CPU, Adreno GPU, Hexagon NPU) - setup guide
Android:
Android arm64 (CPU)
Android arm64 (Snapdragon: CPU, Adreno GPU, Hexagon NPU) - setup guide
Windows:
Windows x64 (CPU)
Windows arm64 (CPU)
Windows arm64 (OpenCL Adreno)
Windows x64 (CUDA 12) - CUDA 12.4 DLLs
Windows x64 (CUDA 13) - CUDA 13.4 DLLs
Windows arm64 (CUDA 13) - CUDA 13.4 DLLs
Windows x64 (Vulkan)
Windows arm64 (Vulkan)
Windows x64 (OpenVINO)
Windows x64 (SYCL)
Windows x64 (ROCm 10.0)
openEuler:
DISABLED
openEuler x86 (310p)
openEuler x86 (910b, ACL Graph)
openEuler aarch64 (310p)
openEuler aarch64 (910b, ACL Graph)
UI:
UI
相關報導
llama.cpp releases10/3 00:57AI 評分25
llama.cpp 的 server 元件新增了 /v1/systemone API,支援 laya、julia-1、lev、openjev、kev 五個模型。該 PR 包含模型轉換、服務端程式碼、共享提示字首、視覺支援,並新增 openjev tiny 模型用於測試。建置覆蓋 macOS、Linux、Windows、Android 等平台,但明確不支援 date_facts。
llama.cpp releases10/5 15:50AI 評分42
llama.cpp 發布 b11424 建置版本,修復 Vulkan 後端 Flash Attention 的共享記憶體越界寫問題(#29988)。該版本照例提供 macOS/iOS、Linux、Windows、Android 的預編譯包,涵蓋 Vulkan、CUDA 12/13、ROCm 10.0、OpenVINO、SYCL、OpenCL 等後端,並附驍龍 CPU/Adreno GPU/Hexagon NPU 的安裝指引。
llama.cpp releases10/4 15:43AI 評分30
llama.cpp 發布 b11392 版本,主要變更是 CI 設定預設權限(#29945)。該版本提供覆蓋 macOS、iOS、Linux、Android、Windows 與 openEuler 的預編譯產物,包括 Apple Silicon arm64、CUDA 12/13、ROCm 10.0、Vulkan、OpenVINO、SYCL 等後端;
llama.cpp releases10/4 09:06AI 評分20
llama.cpp 發布 b11384 版本,官網為 llama.app,並附 GitHub 建置證明。該版本提供 macOS、iOS、Linux、Android、Windows 及 openEuler 的預編譯包,涵蓋 Apple Silicon、CUDA 12/13、ROCm 10.0、Vulkan、SYCL、OpenVINO 與 Snapdragon 的 CPU、Adreno GPU、Hexagon NPU 等後端。
Ollama releases● 精選10/2 19:28AI 評分72
Ollama 發布 v0.35.1,通過 /v1/systemone 介面支援 Cloudflare 新開源的決策模型 Clef(27B)與 Clef Flash(9B),兩者均為多模態,請求可在文字 state 之外附帶圖片,所有問題共享該狀態並聯合打分。同一版本把聯網搜尋上限從每次回應 3 次提高到 10 次,Modelfile 新增 CAPABILITY 宣告,並更新了 llama.cpp 與 MLX 引擎。