quant
Found in 1 package: llama-cpp
llama-cpp(3 releases)
v0.2.0This release bumps the llama.cpp version to 0.2.0 and ggml to 0.21.0, incorporating numerous backend improvements, bug fixes, and new model support across various hardware accelerators like SYCL, OpenCL, Metal, and Vulkan.
b10037This release introduces enhanced quantization capabilities, allowing manual tensor types with the --pure flag. It also provides pre-compiled binaries for various platforms and hardware configurations.
b7807This release introduces a fix where manual tensor type overrides now correctly take precedence within the quant module and provides extensive pre-compiled binaries for macOS, Linux, Windows, and openEuler across various hardware and accelerator configurations.
Track Symbol Changes
Use the Change8 MCP server or GitHub Action to get notified when quant changes.
Learn More