Ollama releases · · 1 min read

v0.33.1

Mirrored from Ollama releases for archival readability. Support the source by reading on the original site.

What's Changed

  • MLX: Qwen3.8 Flash Next support
  • cmake: make external compat patches idempotent
  • MLX and llama.cpp update
  • mlxrunner: add structured output support
  • mlxrunner: avoid Metal GPU timeouts when loading models from slow storage

New Contributors

Full Changelog: v0.33.0...v0.33.1

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from Ollama releases