MLX Swift v0.32.2 has been released. This update includes significant improvements to distributed support, CUDA build performance on Linux, and several API changes designed to better align the Swift implementation with the original Python MLX library.
MLX Swift v0.32.2 Public — Distributed support and CUDA performance improvements
New Features and Improvements
- Added distributed support for MLX.
- Improved CUDA NVCC spawn performance on Linux.
- Added fp8 conversion utilities.
- Added global scale support to quantized layers.
- Added
countNonzeroto the Swift API. - Integrated more robust integration tests.
Bug Fixes and Behavioral Changes
- Fixed a deadlock between
CompiledFunction.lockandevalLock, as well as a data race in the compiler cache. - Fixed memory leaks in
MLXArray(rawPointer:_:dtype:finalizer:). - Fixed
convolvepadRight behavior for even kernels in.samemode. - Fixed
MLXNN.Poolpadding width issues. - Updated
tensordotdefault axes andnanToNumdefaults to match the Python MLX implementation. - Updated
linspaceto handle dtype arguments, ensuring integer arguments default to float32. - Adjusted several optimizer and layer behaviors, including
ALiBi,GRU,Adafactor, andLion, to be more consistent with the Python version.
Sources
- 0.32.2 (2026-09-28)