Ollama Tool 10 related events
Aug 16 · Sun
-
ConfirmedOpen Source
ollama v0.32.14
What's Changed llm: transcode WebP images for llama-server renderers/qwen: tolerate non-leading system messages Full Changelog : v0.32.13...v0.32.14
GitHub Releases · ollama/ollama Aug 15 · Sat
-
ConfirmedOpen Source
ollama v0.32.14-rc0
mlx update ( #17761 )
GitHub Releases · ollama/ollama -
ConfirmedOpen Source
ollama v0.32.12
Qwen 3.8 27B This release adds the support of Qwen 3.8 27B . Qwen3.8 delivers substantial gains across coding, professional work, research, and long-horizon agentic tasks. ollama run qwen3.8:27b For…
GitHub Releases · ollama/ollama Aug 14 · Fri
-
ConfirmedOpen Source
ollama v0.32.13
What's Changed qwen3.8: support developer instructions Full Changelog : v0.32.12...v0.32.13
GitHub Releases · ollama/ollama -
ConfirmedOpen Source
ollama v0.32.11
What's Changed ollama launch dsh now supports DeepSeek Harness, DeepSeek's open-source agent harness ollama launch muse now supports Muse Code , Meta's agentic coding CLI The OpenAI-compatible Respon…
GitHub Releases · ollama/ollama Aug 13 · Thu
-
ConfirmedOpen Source
ollama v0.32.10
What's Changed Models that don't set a repeat_penalty now default to 1.0 (off) instead of 1.1, matching other engines and speeding up speculative decoding; set a per-model parameter if an older model…
GitHub Releases · ollama/ollama Aug 12 · Wed
-
ConfirmedOpen Source
ollama v0.32.10-rc1: mlx: avoid pulling MLX models when MLX is missing (#17710)
As we look to bring Linux and Windows MLX support online, instead of blocking downloads at the registry to avoid users wasting time downloading a model they can't run, shift the logic to the local si…
GitHub Releases · ollama/ollama -
ConfirmedOpen Source
ollama v0.32.10-rc0: nn: speed up prefill on double-scale nvfp4 models
ModelOpt checkpoints apply a float32 global scale to every projection output on top of the per-group quantization scales. Running the multiply and the cast back to the activation dtype as separate ea…
GitHub Releases · ollama/ollama Aug 11 · Tue
-
ConfirmedOpen Source
ollama v0.32.9
NVIDIA Nemotron 3.5 Lightning NVIDIA Nemotron 3.5 Lightning is an open 30B mixture-of-experts (MoE) model with 3B active parameters built for that execution layer of always-on agents. It is designed…
GitHub Releases · ollama/ollama -
ConfirmedOpen Source
ollama v0.32.8
Muse Glimmer Muse Glimmer is now available on all platforms. Muse Glimmer can power coding agent applications such as Claude Code, Codex, Pi and more, as well as long-running personal assistants such…
GitHub Releases · ollama/ollama