Change8

v0.33.3-rc1

📦 ollamaView on GitHub →
2 features

Summary

This release introduces features for reporting cached prompt tokens and honoring GGUF model default parameters. It also includes updates to MLX, MLX-C, and llama.cpp.

✨ New Features

  • Report cached prompt tokens
  • Honor GGUF model defined default parameters