AISpot
所屬事件:Google 發布 EmbeddingGemma 2 端側多模態嵌入模型(6 個來源 · 8 則報導)

10 月 6 日,Google DeepMind 發布 EmbeddingGemma 2,官方稱其為首個原生多模態的端側嵌入開放模型。該模型為 740M 引數,基於 Gemma 4 架構,把文字、程式碼、影像、音訊與影片對映到統一的 768 維向量空間,以 Apache 2.0 許可開源,權重已上線 Hugging F… 看完整事件

Ollama v0.40.0-rc6 在 MLX 上加入多模態嵌入

Ollama releases10/6 16:08版本更新模型地端推論開發工具開源

Ollama 發布 v0.40.0-rc6,在 MLX runner 上實現 EmbeddingGemma2Model 架構,使本地嵌入支援多模態輸入。/api/embed 介面現在可通過 input dict 為每個條目單獨傳入媒體。該模型為 24 層雙向文字編碼器,帶 PLE,共享 gemma4 的視覺與音訊塔,輸出經 mean-pool 與 L2 歸一化。該標籤由 pdevine 於 10 月 6 日打上,對應提交 #18820,仍屬 rc 預發布版本。

多模態嵌入進入 Ollama 的 MLX 推理路徑,/api/embed 開始接受逐條媒體輸入,對在本地用嵌入模型做檢索、需要圖文混合向量的開發者有直接參考價值。

原標題:v0.40.0-rc6: model: add multimodal embeddings (#18820)
閱讀原文

同一事件共有 8 則報導(6 個來源),看事件全貌

評分74 / 78(平均 76,門檻 60)
狀態精選

全文翻譯

ollama / ollama 公開 通知 你必須登入才能更改通知設定 復刻 18.1k 星標 182k v0.40.0-rc6 v0.40.0-rc6 efe43c5 已驗證 此提交是在 GitHub.com 上建立,並使用 GitHub 的已驗證簽名簽署。 GPG 金鑰 ID:B5690EEEBB952194 已驗證 瞭解警惕模式 選擇一個標籤進行比較 抱歉,出了點問題。 篩選 載入中 抱歉,出了點問題。 哎呀! 載入時出錯。請重新載入此頁面。 未找到結果 檢視所有標簽 v0.40.0-rc6:模型:新增多模態嵌入(#18820) v0.40.0-rc6 efe43c5 選擇一個標籤進行比較 抱歉,出了點問題。 篩選 載入中 抱歉,出了點問題。 哎呀! 載入時出錯。請重新載入此頁面。 未找到結果 檢視所有標簽 已驗證 此提交是在 GitHub.com 上建立,並使用 GitHub 的已驗證簽名簽署。 GPG 金鑰 ID:B5690EEEBB952194 已驗證 瞭解警惕模式 pdevine 為此打了標簽 10 月 6 日 15:08 在 MLX 執行器上實現 EmbeddingGemma2Model 架構:24 層 帶 PLE 的雙向文字編碼器,共享的 gemma4 視覺/音訊塔, 均值池化 + L2 輸出。/api/embed 通過輸入字典接受逐項媒體。 資源 2 載入中 哎呀! 載入時出錯。請重新載入此頁面。

由 AI 翻譯,以原文為準。

原文
ollama / ollama Public Notifications You must be signed in to change notification settings Fork 18.1k Star 182k v0.40.0-rc6 v0.40.0-rc6 efe43c5 Verified This commit was created on GitHub.com and signed with GitHub’s verified signature . GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode Choose a tag to compare Sorry, something went wrong. Filter Loading Sorry, something went wrong. Uh oh! There was an error while loading. Please reload this page . No results found View all tags v0.40.0-rc6: model: add multimodal embeddings (#18820) v0.40.0-rc6 efe43c5 Choose a tag to compare Sorry, something went wrong. Filter Loading Sorry, something went wrong. Uh oh! There was an error while loading. Please reload this page . No results found View all tags Verified This commit was created on GitHub.com and signed with GitHub’s verified signature . GPG key ID: B5690EEEBB952194 Verified Learn about vigilant mode pdevine tagged this 06 Oct 15:08 Implements the EmbeddingGemma2Model architecture on the MLX runner: 24-layer bidirectional text encoder with PLE, shared gemma4 vision/audio towers, mean-pool + L2 output. /api/embed accepts per-item media via input dicts. Assets 2 Loading Uh oh! There was an error while loading. Please reload this page .

相關報導

Ollama releases● 精選10/6 19:47AI 評分82

Ollama v0.40.0:Apple Silicon 預設用 MLX 執行模型

Ollama 發布 v0.40.0,在 Apple Silicon 裝置上,MLX 執行時支援的模型架構會自動改用 MLX 執行。官方範例為 ollama pull qwen3.8 與 ollama run qwen3.8,其他可用模型還包括 gemma4、qwen3.6 和 qwen3.5;決策模型 Nimble、tev1、clef、clef-flash 也已支援 MLX,並新增嵌入模型 embeddinggemma-2 支援。官方表示會繼續測試並啟用更多模型。

Ollama releases● 精選10/2 19:28AI 評分72

Ollama v0.35.1 支援 Cloudflare 決策模型 Clef

Ollama 發布 v0.35.1,通過 /v1/systemone 介面支援 Cloudflare 新開源的決策模型 Clef(27B)與 Clef Flash(9B),兩者均為多模態,請求可在文字 state 之外附帶圖片,所有問題共享該狀態並聯合打分。同一版本把聯網搜尋上限從每次回應 3 次提高到 10 次,Modelfile 新增 CAPABILITY 宣告,並更新了 llama.cpp 與 MLX 引擎。