Improved performance and model support with GGUF in Ollama

Ollama has released version 0.30, significantly enhancing performance and expanding model compatibility through GGUF integration via llama.cpp. This update complements Ollama’s existing MLX engine on Apple silicon, allowing a broader array of models to run on diverse hardware configurations.

Source: Ollama