OpenAI-compatible API
Found in 2 packages: unsloth, ollama
unsloth(2 releases)
v0.1.806-betaThis release significantly enhances performance for Qwen and GLM models with MTP, introduces new local media APIs, and improves audio capabilities. It also brings numerous bug fixes and performance optimizations across various features like MLX inference and chat functionalities.
v0.1.805-betaThis release introduces significant performance improvements for Qwen and GLM models with MTP enabled by default, alongside new local media APIs for video and audio, and enhanced MLX inference on Apple Silicon. It also includes numerous bug fixes and reliability upgrades across various features like chat, training, and hardware support.
ollama(5 releases)
v0.18.0-rc2This release introduces documentation for reasoning_effort support in the OpenAI-compatible API and fixes several issues related to cloud model handling and launch command integration.
v0.18.0This release introduces documentation for `reasoning_effort` support in the OpenAI-compatible API and includes several fixes related to cloud model handling and launch command integration.
v0.11.5This release introduces significant memory management improvements for GPU scheduling and multi-GPU setups, alongside performance optimizations for gpt-oss models and reduced installation sizes.
v0.10.0BreakingOllama v0.10.0 introduces a new desktop app, significant performance optimizations for gemma3n and multi-GPU setups, and critical fixes for tool calling and API image support.
v0.5.12This release introduces the Perplexity R1 1776 model and improves the OpenAI-compatible API with tool calling support. It also includes several Linux-specific bug fixes and performance restorations for Intel Xeon processors.
Track Symbol Changes
Use the Change8 MCP server or GitHub Action to get notified when OpenAI-compatible API changes.
Learn More