Mamba
Found in 3 packages: vllm, transformers, localai
vllm(4 releases)
v0.27.0BreakingvLLM v0.27.0 introduces Kimi K3 support, new model additions like Qwen3.5 and VaultGemma, and a significant upgrade to PyTorch 2.13.0. The release also enhances performance and features across various areas including FlashAttention 4, Model Runner V2, KV offloading, and hardware enablement.
v0.24.0Breakingv0.24.0 introduces extensive support and performance optimizations for new models like MiniMax-M3 and DeepSeek-V4, matures the Model Runner V2 with default quantization support, and overhauls device selection by removing internal use of CUDA_VISIBLE_DEVICES.
v0.22.0This release focuses heavily on DeepSeek V4 maturity with new kernel support and packaging, significant advancements in Model Runner V2, and the introduction of an experimental Rust frontend. Performance saw notable gains from batch-invariant inference with Cutlass FP8 and the rollout of multi-tier KV cache offloading.
v0.17.0BreakingvLLM v0.17.0 introduces a major upgrade to PyTorch 2.10, integrates FlashAttention 4, and significantly matures Model Runner V2 with features like Pipeline Parallelism. This release also adds full support for the Qwen3.5 model family and introduces new performance tuning flags.
transformers(3 releases)
v5.15.0BreakingThis release introduces several new models including Meta Muse Glimmer, GraniteMoeSWA, A.X-K1/K2, and Cosmos3 Edge. It also includes significant updates to attention mechanisms, vision processing, and generation capabilities, alongside important breaking changes in kernel opt-in and cache cropping.
v4.55.3BreakingPatch release 4.55.3 focuses on stability improvements for FlashAttention-2 on Ascend NPU, FSDP sharding fixes, and critical bug fixes for GPT-OSS and Mamba models.
4.54.1BreakingA maintenance patch release focused on fixing regressions in cache inheritance, device placement, and distributed training (TP/device-mesh) across various model architectures like ModernBERT, GPT2, and Mamba.
localai(1 releases)
Track Symbol Changes
Use the Change8 MCP server or GitHub Action to get notified when Mamba changes.
Learn More