v0.33.1
Mirrored from Ollama releases for archival readability. Support the source by reading on the original site.
What's Changed
- MLX: Qwen3.8 Flash Next support
- cmake: make external compat patches idempotent
- MLX and llama.cpp update
- mlxrunner: add structured output support
- mlxrunner: avoid Metal GPU timeouts when loading models from slow storage
New Contributors
Full Changelog: v0.33.0...v0.33.1
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.