LLM quantization

SpinQuant

Superseded baseline#7 of 80 most-superseded

Superseded — cited as a baseline and beaten by newer methods

3 papers critique it · 8 beat it on benchmarks

What papers say

Verbatim critique sentences, each from a paper that cites SpinQuant as a baseline.

QuaRot, SpinQuant, and ButterflyQuant do not engage with directly [the regime of per-head q_norm/RoPE compatibility failures]
Influence-Inspired Spectral Rotations for Extreme Low-Bit LLM Quantization
previous gradient-based optimization~spinquant cannot easily explore permutation invariance, as permutation creates symmetric local optima in a non-convex fashion.
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
Unlike other learnable methods (e.g., SpinQuant liu2024spinquant) that optimize over the full Stiefel manifold with high computational cost, our sparse parameterization guarantees orthogonality by construction, enabling stable and efficient optimization.
ButterflyQuant: Ultra-low-bit LLM Quantization through Learnable Orthogonal Butterfly Transforms

Beaten on benchmarks

Head-to-head results where a newer method reports beating SpinQuant. Values are copied from the source paper's tables — verify against the cited paper.

What to use instead

Recent methods in the same sub-problem, not yet superseded in the knowledge base — arXiv benchmark leaders, not vetted production recommendations.