AISpot

v0.40.1

Ollama releases10/7 21:04版本更新地端推論開源

Ollama 發布 v0.40.1,服務端新增代理雲端用量與餘額 API,客戶端可藉此讀取帳戶用量與餘額。該版本還把帳號步驟從 CLI 新手引導中移除,修復 Windows 上 llama 讀取超過 2GiB 的 clef head 問題,並讓 manifest 在 Windows 上避免使用符號連結。MLX 方面刪除了已進入上游的 Metal residency 補丁,文件修復了 README 社群整合列表中的 6 個失效連結和應用 README 的下載連結。

原標題:v0.40.1
閱讀原文

評分42 / 44(平均 43,門檻 60)
狀態未入選

原文

What's Changed server: proxy cloud usage and balance APIs by @drifkin in #18829 llama: fix clef head reads past 2GiB on windows by @Gigrise in #18777 cmd: remove account step from CLI onboarding by @hoyyeva in #18826 manifest: avoid symlinks on Windows by @dhiltgen in #18852 docs: fix 6 dead links in README community integrations list by @aniketkrs in #18814 docs: fix broken download links in app README by @chenlichao in #18233 mlx: drop carried metal residency patch now that it is upstream by @dhiltgen in #18854 New Contributors @Gigrise made their first contribution in #18777 @aniketkrs made their first contribution in #18814 @chenlichao made their first contribution in #18233 Full Changelog : v0.40.0...v0.40.1-rc0

相關報導

Ollama releases● 精選10/6 14:47AI 評分82

Ollama v0.40.0:Apple Silicon 預設用 MLX 執行模型

Ollama 發布 v0.40.0,在 Apple Silicon 裝置上,MLX 執行時支援的模型架構會自動改用 MLX 執行。官方範例為 ollama pull qwen3.8 與 ollama run qwen3.8,其他可用模型還包括 gemma4、qwen3.6 和 qwen3.5;決策模型 Nimble、tev1、clef、clef-flash 也已支援 MLX,並新增嵌入模型 embeddinggemma-2 支援。官方表示會繼續測試並啟用更多模型。

Ollama releases● 精選10/2 14:28AI 評分72

Ollama v0.35.1 支援 Cloudflare 決策模型 Clef

Ollama 發布 v0.35.1,通過 /v1/systemone 介面支援 Cloudflare 新開源的決策模型 Clef(27B)與 Clef Flash(9B),兩者均為多模態,請求可在文字 state 之外附帶圖片,所有問題共享該狀態並聯合打分。同一版本把聯網搜尋上限從每次回應 3 次提高到 10 次,Modelfile 新增 CAPABILITY 宣告,並更新了 llama.cpp 與 MLX 引擎。

Ollama releases● 精選10/6 11:08AI 評分76

Ollama v0.40.0-rc6 在 MLX 上加入多模態嵌入

Ollama 發布 v0.40.0-rc6,在 MLX runner 上實現 EmbeddingGemma2Model 架構,使本地嵌入支援多模態輸入。/api/embed 介面現在可通過 input dict 為每個條目單獨傳入媒體。該模型為 24 層雙向文字編碼器,帶 PLE,共享 gemma4 的視覺與音訊塔,輸出經 mean-pool 與 L2 歸一化。該標籤由 pdevine 於 10 月 6 日打上,對應提交 #18820,仍屬 rc 預發布版本。