O

Ollama

Ollama v0.32.6 — ## What's Changed - Qwen3.5 is faster on Apple GPUs: the MLX engine now uses the model's MTP head for speculative decoding automatically - `/v1/chat/completions`…

Local modelsOpen source4.5 / 5
Visit site
WA

What it's genuinely good at

  • Open source
  • Active release cadence

Where it falls apart

    Alternatives to Ollama

    More Local models tools