Real Atlassian solutions to real problems — no fluff, no SEO spam.
NVFP4 has a second scaling level. mlx-lm never sets it, so NVFP4 converts to the worst 4-bit result on the table — worse than a format using fewer bits. Set it and NVFP4 wins. I nearly published this as a defect in NVIDIA's spec; it was my own unit error.