Software Watch

Ollama (Stable) 0.34.1

公開 · 検出 2026-09-19

What's Changed * MLX safetensors ollama create no longer experimental. GGUF model creation now requires using llama.cpp tooling for safetensor conversion and quantization. * Improved MLX memory handling on Apple Silicon * Runaway repeat token detection now requires 100 repeat tokens for reduced false positives (e.g. OCR) * /api/tags is much faster on large model libraries (3.1 s → 294 ms cold in testing), and model capabilities are now reported consistently. * Deprecated typical_p: it can n…

公式のリリース情報 → · Ollama (Stable) の履歴 · Ollama まとめ