Perplexity Can Miss SAE Feature Damage Under Quantization
arXiv:2606.03002v2 Announce Type: replace Abstract: Quantization is a standard path to deploying large language models, and quantized models are typically judged acceptable when perplexity or downstream accuracy remains close to the full-precision original. But behavioral parity need not imply feature…
