Ollama

ollama/ollama last check 210 releases recent
Notes
Release notes
v0.33.1 · recent
view on github

What's Changed

  • MLX: Qwen3.8 Flash Next support
  • cmake: make external compat patches idempotent
  • MLX and llama.cpp update
  • mlxrunner: add structured output support
  • mlxrunner: avoid Metal GPU timeouts when loading models from slow storage

New Contributors

Full Changelog: https://github.com/ollama/ollama/compare/v0.33.0...v0.33.1