Ollama has released version 0.34.1, introducing several updates. The release makes MLX safetensors "ollama create" non-experimental and improves MLX memory handling on Apple Silicon. Additionally, runaway repeat token detection has been refined to reduce false positives, and the `/api/tags` endpoint now offers significantly faster performance for large model libraries. AI
IMPACT Improves the user experience and performance for local LLM deployment via Ollama.
RANK_REASON This is a software release for a tool that facilitates running LLMs locally, not a frontier model release or significant industry event.
Read on Mastodon — fosstodon.org →
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →