LLM quantization
SplitQuantV2: Enhancing Low-Bit Quantization of LLMs Without GPUs
Present, but with little supersession signal in the knowledge base
0 papers critique it · 0 beat it on benchmarks