Data
ML frameworks
-
mlx-vlm v0.7.0 Released — Expanded Model Support and inference Performance Optimizations
The latest version of MLX-VLM, v0.7.0, has been released. This update adds support for new models such as LLaVA-OneVision and DeepSeek-V4 Flash Vision EXP, alongside multiple optimizations for inference processes.
-
mlx-vlm v0.7.0rc0 Released — MTP Support for Qwen3.8-Flash-Next and Numerous Bug Fixes
A release candidate featuring MTP and continuous batching support for Qwen3.8-Flash-Next, an APC redesign, and fixes for tool calling and streaming responses.
-
mlx-vlm v0.6.16–17 Released: New Model Support and Metal Resource Leak Fixes
mlx-vlm has released v0.6.16 and v0.6.17. These updates add support for new models such as GLM-5.3-Flash and Qwen3.8-27B, while fixing multiple bugs including Metal buffer leaks and RoPE position errors.
-
mlx v0.32.2 Released — Floating-Point Bug Fixes and NAX Attention Optimizations
Fixes truncation in floating-point divmod and NaN dropping in median. Adds fused attention paths for NAX devices and memory read optimizations for gqa-8 decoding.