← Search

Jatin Chhugani

1 accepted papers

2026

Unveiling the Potential of Quantization with MXFP4: Strategies for Quantization Error Reduction

ICML 2026poster

Large Language Models (LLMs) have intensified the need for low-precision formats that enable efficient, large-scale inference. The Open Compute Project (OCP) Microscaling (MX) standard is attractive due to its favorable hardware efficiency, but its 4-bit variant (MXFP4) lags behind NVIDIA’s NVFP4 in…

Cited by 0SourceScholar