此主題的繁體中文文章,最新優先。English posts in this topic, newest first.
Ollama 0.30 透過 llama.cpp 深化 GGUF 引擎:NVIDIA 吞吐最高提升 20%(RTX 5090 實測條件)、Vulkan 預設開啟讓 AMD 與 Intel 開箱即用、LFM 與 Prism 家族及 Unsloth 微調可直接執行,tool calling 能力沿用並可掛上 coding agent。
Ollama 0.30 deepens its GGUF engine via llama.cpp: up to 20% faster NVIDIA throughput under a stated RTX 5090 condition, Vulkan on by default, and tool calling that carries to coding agents.